Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 14 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 11 inbound Pith citation observations for arXiv:2312.09187.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-14T04:40:11.923452Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-04T09:59:44.592995Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation 6f759d14-e0f2-4a46-984d-4cf3474c1a56 · inbound
LLM-Based Offline Learning for Embodied Agents via Consistency-Guided Reward Ensemble Vision-Language Models as a Source of Rewards
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f14040fc-c140-4551-af25-53001d9673c0 · inbound
STEVE-Audio: Expanding the Goal Conditioning Modalities of Embodied Agents in Minecraft Vision-Language Models as a Source of Rewards
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bcb625d3-95da-4688-af60-76847536ef5c · inbound
YingSound: Video-Guided Sound Effects Generation with Multi-modal Chain-of-Thought Controls Vision-Language Models as a Source of Rewards
Reference 102
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b83bc419-0129-477e-90b2-e188a5197b0d · inbound
RobotSmith: Generative Robotic Tool Design for Acquisition of Complex Manipulation Skills Vision-Language Models as a Source of Rewards
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 393ff599-eb2f-415e-9f3d-295793eb7661 · inbound
FOUNDER: Grounding Foundation Models in World Models for Open-Ended Embodied Decision Making Vision-Language Models as a Source of Rewards
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9ff64ef4-43cb-4187-8a91-b9730e3e8d1b · inbound
SOLE-R1: Video-Language Reasoning as the Sole Reward for On-Robot Reinforcement Learning Vision-Language Models as a Source of Rewards
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8f96b5e0-a3ca-448d-a974-045bb3a59bb0 · inbound
Learning Process Rewards via Success Visitation Matching for Efficient RL Vision-Language Models as a Source of Rewards
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 5e767bed-48c0-49f0-b88f-1a9050aebd65 · inbound
QVal: Cheaply Evaluating Dense Supervision Signals for Long-Horizon LLM Agents Vision-Language Models as a Source of Rewards
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation f7291943-708c-4944-9fe9-03b5eeb91fad · inbound
Progress Reward Modeling for Robotic Learning: A Comprehensive Survey Vision-Language Models as a Source of Rewards
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fccf0884-caea-4685-8570-debabaf3e1fa · inbound
A Unified Framework for Dynamic Reward Shaping in Reinforcement Learning Vision-Language Models as a Source of Rewards
Reference 2017
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0884635c-e168-4f7e-8c97-f4efbcac098c · inbound
TrustRoboReward: Preference-Ordered Isotonic Score Editing for Multi-Paradigm Robot Reward Models Vision-Language Models as a Source of Rewards
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.