Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 5 inbound Pith citation observations for arXiv:2505.03792.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-07-31T21:55:17.309418Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-02T12:16:56.911098Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation 5bd1a5f0-d1bd-45c8-99e5-55f276d9e0e9 · inbound
Odysseus: Scaling VLMs to 100+ Turn Decision-Making in Games via Reinforcement Learning Towards Efficient Online Tuning of VLM Agents via Counterfactual Soft Reinforcement Learning
Reference 70
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6b48a64f-f88f-4bbb-acb2-db4e987ba59b · inbound
StraTA: Incentivizing Agentic Reinforcement Learning with Strategic Trajectory Abstraction Towards Efficient Online Tuning of VLM Agents via Counterfactual Soft Reinforcement Learning
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e0c6c9f2-3dda-46d5-8b59-7eb88bd81ba9 · inbound
SIRI: Self-Internalizing Reinforcement Learning with Intrinsic Skills for LLM Agent Training Towards Efficient Online Tuning of VLM Agents via Counterfactual Soft Reinforcement Learning
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 3c679b43-a9f6-4690-8251-134ba644d44a · inbound
When Denser Credit Is Not Enough: Evidence-Calibrated Policy Optimization for Long-Horizon LLM Agent Training Towards Efficient Online Tuning of VLM Agents via Counterfactual Soft Reinforcement Learning
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f948133b-c362-4c92-aa50-e1a8598e1a00 · inbound
MARS-RA: Rank Aggregation for Credit Assignment via Multimodal Comparisons in Embodied Multi-Agent Cooperation Towards Efficient Online Tuning of VLM Agents via Counterfactual Soft Reinforcement Learning
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.