Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 6 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 6 inbound Pith citation observations for arXiv:2402.04764.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-05T06:32:48.257954+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-05T12:24:02.379856Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z
0 of 0 outbound references displayed
External citation measurements
1
arxiv_reference, observed 2026-08-05T02:28:24.338817Z
No outbound reference observations are available for this paper version.
Observation 5ef0f259-b734-480b-83f0-32aa46ffb8de · inbound
Reinforcement Learning for Machine Learning Engineering Agents Code as Reward: Empowering Reinforcement Learning with VLMs
Reference 60
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8e582b40-d12d-421b-8bb9-ac4cc3893c17 · inbound
SOLE-R1: Video-Language Reasoning as the Sole Reward for On-Robot Reinforcement Learning Code as Reward: Empowering Reinforcement Learning with VLMs
Reference 56
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 71bedf21-e11d-47d6-abf7-0821cf3432ae · inbound
Beyond NL2Code: A Structured Survey of Multimodal Code Intelligence Code as Reward: Empowering Reinforcement Learning with VLMs
Reference 77
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 6e83b4c1-aca7-4359-b8df-6992451c2968 · inbound
Learning Process Rewards via Success Visitation Matching for Efficient RL Code as Reward: Empowering Reinforcement Learning with VLMs
Reference 85
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation d7d7e84b-5cbe-4546-be36-ed692400209a · inbound
QVal: Cheaply Evaluating Dense Supervision Signals for Long-Horizon LLM Agents Code as Reward: Empowering Reinforcement Learning with VLMs
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation fd35c397-5693-4c02-86b2-8a0975e6cef4 · inbound
Weights or Skills? A Survey of Robot-Learning Techniques: from Action-Predicting Weights to Robots that Write their Own Skills Code as Reward: Empowering Reinforcement Learning with VLMs
Reference 245
Source-reported events for the cited work
Unavailable: canonical work link unavailable.