Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 4 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 16 inbound Pith citation observations for arXiv:1406.5979.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-04T06:34:03.388597+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-03T18:30:49.157907Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-07-03T08:17:45.248813Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation 6a53853e-3cbc-4f46-8be1-ca49315e7df5 · inbound
Learning Belief Representations for Imitation Learning in POMDPs Reinforcement and Imitation Learning via Interactive No-Regret Learning
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 43814708-aa17-4c39-bdbc-ecd51f5ff171 · inbound
Rewarding Progress: Scaling Automated Process Verifiers for LLM Reasoning Reinforcement and Imitation Learning via Interactive No-Regret Learning
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 8b12cbb0-c742-479d-ae10-1566c49eac71 · inbound
State-Conditional Adversarial Learning: An Off-Policy Visual Domain Transfer Method for End-to-End Imitation Learning Reinforcement and Imitation Learning via Interactive No-Regret Learning
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation b325bd8a-c583-4862-85b8-638755cfab4a · inbound
State-Conditional Adversarial Learning: An Off-Policy Visual Domain Transfer Method for End-to-End Imitation Learning Reinforcement and Imitation Learning via Interactive No-Regret Learning
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6e09e307-c4c6-4f66-808a-f4b7891234c6 · inbound
BridgeSim: Unveiling the OL-CL Gap in End-to-End Autonomous Driving Reinforcement and Imitation Learning via Interactive No-Regret Learning
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 42dd62c3-8446-439d-8c22-aed3812b466b · inbound
Vision-Language-Action Jump-Starting for Reinforcement Learning Robotic Agents Reinforcement and Imitation Learning via Interactive No-Regret Learning
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 22e50df1-5037-4ef0-8ab1-927df1d955fe · inbound
Vision-Language-Action Jump-Starting for Reinforcement Learning Robotic Agents Reinforcement and Imitation Learning via Interactive No-Regret Learning
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9d5f0629-4a9c-41d1-85c7-20e3b10681a8 · inbound
The Essence of Balance for Self-Improving Agents in Vision-and-Language Navigation Reinforcement and Imitation Learning via Interactive No-Regret Learning
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation c4838a9d-3bb8-4cb4-b8ff-e2e84a43988c · inbound
Provable imitation learning for control of instability in partially-observed Vlasov--Poisson equations Reinforcement and Imitation Learning via Interactive No-Regret Learning
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 3bf08f22-8d83-49dd-b4ca-e8470761cfcc · inbound
Revisiting DAgger in the Era of LLM-Agents Reinforcement and Imitation Learning via Interactive No-Regret Learning
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 748ecd4f-f159-4602-a947-a3c939abad0d · inbound
Inverse Reinforcement Learning without an Optimal Demonstrator: A Feasible Reward Set Approach Reinforcement and Imitation Learning via Interactive No-Regret Learning
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation abd24304-83cf-4192-80fa-107bc7b467b7 · inbound
Speculative Rollback Correction for Quality-Diverse Web Agent Imitation Reinforcement and Imitation Learning via Interactive No-Regret Learning
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 3a4556c6-937b-4a99-b035-2d7504350ce5 · inbound
When Does Online Imitation Learning Help in LLM Post-Training? The Role of (Non-)Realizability Beyond Horizon Reinforcement and Imitation Learning via Interactive No-Regret Learning
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 344e9a74-af39-4882-8b3d-4db7ce7e542f · inbound
Behavior Cloning is Not All You Need: The Optimality of On-Policy Distillation for Noisy Expert Feedback Reinforcement and Imitation Learning via Interactive No-Regret Learning
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation a7b4f5be-d21c-4bb6-9bb3-2a82f49e70c9 · inbound
A Few Teacher Steps Go a Long Way: Cost-Efficient On-Policy Data Augmentation for Agent Post-Training Reinforcement and Imitation Learning via Interactive No-Regret Learning
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d8c4c4d5-f76d-4998-a8ce-c3d174e4721d · inbound
CAST: Game Solvers as Turn-Level Teachers for LLM Agents Reinforcement and Imitation Learning via Interactive No-Regret Learning
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.