Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-05T23:41:31.838029Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 31 of 31 outbound references and 18 inbound Pith citation observations for arXiv:2508.06571.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-05T23:41:31.838029Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-03T16:27:24.257956Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-07-10T08:36:59.779636Z
31 of 31 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 19a43edb-9e70-49d0-8bd3-e205ac0bd3a0 · outbound
IRL-VLA: Training an Vision-Language-Action Policy via Reward World Model GPT-4 Technical Report
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 73b426f7-a4dc-4022-a6de-0d63445c8c70 · outbound
IRL-VLA: Training an Vision-Language-Action Policy via Reward World Model Training diffusion models with reinforcement learning, 2024
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 10117937-04c5-4fdc-a9e3-327b56b8ff8a · outbound
IRL-VLA: Training an Vision-Language-Action Policy via Reward World Model Pseudo- simulation for autonomous driving
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 12831453-a6e2-4209-8fce-de07c480d253 · outbound
IRL-VLA: Training an Vision-Language-Action Policy via Reward World Model Transfuser: Imitation with transformer-based sensor fusion for au- tonomous driving
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 1c3578b4-b05c-4f19-97c9-5a75c2cb28ba · outbound
IRL-VLA: Training an Vision-Language-Action Policy via Reward World Model Parting with misconceptions about learning-based vehicle motion planning
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e3db77b6-c139-4555-9b31-0fb3986738a0 · outbound
IRL-VLA: Training an Vision-Language-Action Policy via Reward World Model Navsim: Data-driven non-reactive autonomous vehicle simulation and benchmarking
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6bc197cc-12e3-47c6-bc39-60c10520e2c1 · outbound
IRL-VLA: Training an Vision-Language-Action Policy via Reward World Model ORION: A Holistic End-to-End Autonomous Driving Framework by Vision-Language Instructed Action Generation
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation be38b2d6-bcb0-4ed9-94cd-b5159061e992 · outbound
IRL-VLA: Training an Vision-Language-Action Policy via Reward World Model Rad: Training an end-to-end driving policy via large-scale 3dgs-based reinforcement learning
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 48f5bac0-284e-4315-9f00-8a476dcdee68 · outbound
IRL-VLA: Training an Vision-Language-Action Policy via Reward World Model DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fb25938f-71dc-4de2-8767-58b5d5feba6a · outbound
IRL-VLA: Training an Vision-Language-Action Policy via Reward World Model Planning-oriented autonomous driving
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 9a9efb1e-292e-48fb-8693-635745dd3494 · outbound
IRL-VLA: Training an Vision-Language-Action Policy via Reward World Model EMMA: End-to-End Multimodal Model for Autonomous Driving
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e9227cc0-15a1-4ff1-b4f9-4abff732a5c1 · outbound
IRL-VLA: Training an Vision-Language-Action Policy via Reward World Model Planning with Diffusion for Flexible Behavior Synthesis
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 539c03c9-b0de-49b4-acd6-2035043bae97 · outbound
IRL-VLA: Training an Vision-Language-Action Policy via Reward World Model DiffVLA: Vision-Language Guided Diffusion Planning for Autonomous Driving
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 213b3c2c-5e5c-4467-bafe-94ad9546f0c9 · outbound
IRL-VLA: Training an Vision-Language-Action Policy via Reward World Model Senna: Bridging Large Vision-Language Models and End-to-End Autonomous Driving
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1ff95cbe-0d21-43ac-b262-001535e5e840 · outbound
IRL-VLA: Training an Vision-Language-Action Policy via Reward World Model Vad: Vectorized scene representation for efficient autonomous driving
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 230a0b04-8927-41b3-9e69-ca7de315e0e4 · outbound
IRL-VLA: Training an Vision-Language-Action Policy via Reward World Model End-to-End Driving with Online Trajectory Evaluation via BEV World Model
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b9d1c1f7-27d1-4edd-ae28-ea730f060092 · outbound
IRL-VLA: Training an Vision-Language-Action Policy via Reward World Model ReCogDrive: A Reinforced Cognitive Framework for End-to-End Autonomous Driving
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation acc7543c-7e1d-4d01-898f-e197281976e8 · outbound
IRL-VLA: Training an Vision-Language-Action Policy via Reward World Model Hydra-MDP: End-to-end Multimodal Planning with Multi-target Hydra-Distillation
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 71701eec-44e2-461e-9382-3e4e5b9274f0 · outbound
IRL-VLA: Training an Vision-Language-Action Policy via Reward World Model Generalized Trajectory Scoring for End-to-end Multimodal Planning
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2be68a7b-7bbd-49c4-9637-7eca3de7b726 · outbound
IRL-VLA: Training an Vision-Language-Action Policy via Reward World Model DiffusionDrive: Truncated Diffusion Model for End-to-End Autonomous Driving
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 46338617-f4df-429c-b157-1ed26c088973 · outbound
IRL-VLA: Training an Vision-Language-Action Policy via Reward World Model Diffusiondrive: Trun- cated diffusion model for end-to-end autonomous driv- ing
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c3cab5c8-2364-4ba4-b226-026b9df6c532 · outbound
IRL-VLA: Training an Vision-Language-Action Policy via Reward World Model Diffusion Policy Policy Optimization
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a7fe2bb6-be6c-4f05-bd13-76e76d5c13a4 · outbound
IRL-VLA: Training an Vision-Language-Action Policy via Reward World Model Simlingo: Vision-only closed-loop au- tonomous driving with language-action alignment
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation bd531091-cd19-4492-8439-a8e50e69d93c · outbound
IRL-VLA: Training an Vision-Language-Action Policy via Reward World Model High-dimensional continuous control using generalized advantage estima- tion, 2018
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 39ef19ad-c823-45e8-a107-2e31eae148bc · outbound
IRL-VLA: Training an Vision-Language-Action Policy via Reward World Model Unresolved cited work
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3bfb75c9-dc1a-4684-9f11-68baacb63df7 · outbound
IRL-VLA: Training an Vision-Language-Action Policy via Reward World Model SparseDrive: End-to-End Autonomous Driving via Sparse Scene Representation
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation be67406b-29b3-46b7-8ea6-1bf9730e4512 · outbound
IRL-VLA: Training an Vision-Language-Action Policy via Reward World Model Diffsemanticfusion: Semantic raster bev fusion for autonomous driving via online hd map diffusion, 2025
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 64093467-0a56-4ce8-aa72-b159878d7a16 · outbound
IRL-VLA: Training an Vision-Language-Action Policy via Reward World Model Efficient Reinforcement Learning for Autonomous Driving with Parameterized Skills and Priors
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6a08dcfb-c46b-421c-951d-930a25083ecc · outbound
IRL-VLA: Training an Vision-Language-Action Policy via Reward World Model Carplanner: Consistent auto-regressive trajectory plan- ning for large-scale reinforcement learning in au- tonomous driving
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 915dc76c-a89d-4f05-af92-afa619e7bb84 · outbound
IRL-VLA: Training an Vision-Language-Action Policy via Reward World Model Accelerating reinforcement learning for autonomous driving using task-agnostic and ego- centric motion skills
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 02553d1d-7218-4414-8240-7aa7760f6556 · outbound
IRL-VLA: Training an Vision-Language-Action Policy via Reward World Model Opendrivevla: Towards end-to-end autonomous driving with large vision language action model
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0bf917a1-ce2c-4b1c-9be6-5d5b4f46a41b · inbound
World-Env: Leveraging World Model as a Virtual Environment for VLA Post-Training IRL-VLA: Training an Vision-Language-Action Policy via Reward World Model
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b8797e64-ab29-46c2-9fbb-9996a361131f · inbound
Alpamayo-R1: Bridging Reasoning and Action Prediction for Generalizable Autonomous Driving in the Long Tail IRL-VLA: Training an Vision-Language-Action Policy via Reward World Model
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b1fde534-10ca-400a-bfb6-388f62c45197 · inbound
AVA-VLA: Improving Vision-Language-Action models with Active Visual Attention IRL-VLA: Training an Vision-Language-Action Policy via Reward World Model
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation bba8b92a-e0e6-496f-b302-b97ab1d846c4 · inbound
Latent Chain-of-Thought World Modeling for End-to-End Driving IRL-VLA: Training an Vision-Language-Action Policy via Reward World Model
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 78427e88-2f27-4152-ac07-e3e37463355d · inbound
MindDrive: A Vision-Language-Action Model for Autonomous Driving via Online Reinforcement Learning IRL-VLA: Training an Vision-Language-Action Policy via Reward World Model
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cc3c1ee7-953f-4aaa-a642-ed20e715d474 · inbound
Vision-and-Language Navigation for UAVs: Progress, Challenges, and a Research Roadmap IRL-VLA: Training an Vision-Language-Action Policy via Reward World Model
Reference 169
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 726b3100-d940-45dd-bc8e-612e43ff0621 · inbound
Human Cognition in Machines: A Unified Perspective of World Models IRL-VLA: Training an Vision-Language-Action Policy via Reward World Model
Reference 80
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 906c2d25-f002-4967-b119-fad978051e4b · inbound
SpanVLA: Efficient Action Bridging and Learning from Negative-Recovery Samples for Vision-Language-Action Model IRL-VLA: Training an Vision-Language-Action Policy via Reward World Model
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6140c01f-3951-42f6-806c-f72b4e4d7bb9 · inbound
Latency Analysis and Optimization of Alpamayo 1 via Efficient Trajectory Generation IRL-VLA: Training an Vision-Language-Action Policy via Reward World Model
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a9519e4e-7fde-4811-a07e-b7089c25c52e · inbound
Distill to Think, Foresee to Act: Cognitive-Physical Reinforcement Learning for Autonomous Driving IRL-VLA: Training an Vision-Language-Action Policy via Reward World Model
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b49abf41-3de4-495c-973f-050d5cf5a7f8 · inbound
Distill to Think, Foresee to Act: Cognitive-Physical Reinforcement Learning for Autonomous Driving IRL-VLA: Training an Vision-Language-Action Policy via Reward World Model
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 0e8534ad-08b3-4702-9e8b-cbba6dfdff07 · inbound
NTR: Neural Token Reconstruction for Scene Token Bottleneck in End-to-End Driving IRL-VLA: Training an Vision-Language-Action Policy via Reward World Model
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b00599c5-6f3f-430f-bbc2-6a4f40027e71 · inbound
IDOL: Inverse-Dynamics-Guided Future Prediction for End-to-End Autonomous Driving IRL-VLA: Training an Vision-Language-Action Policy via Reward World Model
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ceaf1399-6b15-4a6f-9991-67721a9a07c6 · inbound
World Models for Robotic Manipulation: A Survey IRL-VLA: Training an Vision-Language-Action Policy via Reward World Model
Reference 106
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 860069f4-8aa2-41d2-9a69-2b774bdb3957 · inbound
Discrete-WAM: Unified Discrete Vision-Action Token Editing for World-Policy Learning IRL-VLA: Training an Vision-Language-Action Policy via Reward World Model
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 673c745e-6bdb-452c-9c15-9afbabf66cdb · inbound
World Engine: Towards the Era of Post-Training for Autonomous Driving IRL-VLA: Training an Vision-Language-Action Policy via Reward World Model
Reference 101
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5a592479-c7e2-440b-bb5b-c89a8e3a29bf · inbound
Post-Training in End-to-End Autonomous Driving IRL-VLA: Training an Vision-Language-Action Policy via Reward World Model
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 29aca45e-edf5-4daa-8fdf-06915c8773ba · inbound
WCog-VLA: A Dual-Level World-Cognitive Vision-Language-Action Model for End-to-End Autonomous Driving IRL-VLA: Training an Vision-Language-Action Policy via Reward World Model
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.