Pith. sign in

Paper Citation Record · LEDGER

RAD-2: Scaling Reinforcement Learning in a Generator-Discriminator Framework

As of 6 August 2026, this Paper Citation Record lists 67 of 67 outbound references and 4 inbound Pith citation observations for arXiv:2604.15308.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2604.15308 v1

Coverage vector

measured 67 of 67 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-05-10T11:40:26.649975Z

measured 71 of 71 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-06T06:34:29.942622+00:00

measured 4 of 4 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-07-01T05:09:44.224340Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-01T10:45:42.632381Z

Reference resolution

67 of 67 outbound references displayed

  • verified exact40
  • verified fuzzy22
  • unresolved1
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch4

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 840b0c73-cf82-4a07-be4b-883d745314d3 · outbound

This paper cites Lan- guage models are few-shot learners.Advances in neural in- formation processing systems, 33:1877–1901.

RAD-2: Scaling Reinforcement Learning in a Generator-Discriminator Framework Lan- guage models are few-shot learners.Advances in neural in- formation processing systems, 33:1877–1901

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-05-19T13:42:19.940862Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T11:40:26.649975Z digest=sha256:36eebe22f91f7cce4218a66604ef5952d2aca118541f492ee72d9d6bcad5432a

Observation f0dbca0e-be35-4dca-893c-7892d2c8cc93 · outbound

This paper cites VADv2: End-to-End Vectorized Autonomous Driving via Probabilistic Planning.

RAD-2: Scaling Reinforcement Learning in a Generator-Discriminator Framework VADv2: End-to-End Vectorized Autonomous Driving via Probabilistic Planning

Reference 2

Resolution
metadata mismatch
local_arxiv, observed 2026-05-10T11:50:21.041065Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T11:40:26.649975Z digest=sha256:609824f5ecd4b1e143d8cd23ecb1ac7d2c2bcde7fa6fc3788d558bb6d237102f

Observation 92b86fd1-531d-45c6-9395-4c3209775aae · outbound

This paper cites Transfuser: Imitation with transformer-based sensor fusion for autonomous driv- ing.IEEE transactions on pattern analysis and machine in- telligence, 45(11):12878–12895.

RAD-2: Scaling Reinforcement Learning in a Generator-Discriminator Framework Transfuser: Imitation with transformer-based sensor fusion for autonomous driv- ing.IEEE transactions on pattern analysis and machine in- telligence, 45(11):12878–12895

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-05-19T13:42:19.954402Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T11:40:26.649975Z digest=sha256:1313e53d7902db0ce1ac4414b9da5890adb915d5aa7d94a9ad59d76ea1700bd8

Observation 2d2cf40f-1149-4bad-8976-4cf465f8bffb · outbound

This paper cites Carla: An open urban driv- ing simulator.

RAD-2: Scaling Reinforcement Learning in a Generator-Discriminator Framework Carla: An open urban driv- ing simulator

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-05-19T13:42:19.956583Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T11:40:26.649975Z digest=sha256:38284f2bcb993dea88827302ca5c67f853fb89f5e32e394f178bf8d0ee9d3e59

Observation d07e678b-9544-4407-a69c-0161020672db · outbound

This paper cites Resisting Stochastic Risks in Diffusion Planners with the Trajectory Aggregation Tree.

RAD-2: Scaling Reinforcement Learning in a Generator-Discriminator Framework Resisting Stochastic Risks in Diffusion Planners with the Trajectory Aggregation Tree

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-10T11:50:21.045184Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T11:40:26.649975Z digest=sha256:5f6574769cbf90adc8fe8c6b66ac3764a9c774fd05a2a1049d0911f17ca9c557

Observation 51c2ff29-f615-4989-acf3-288d6436912b · outbound

This paper cites MindDrive: A Vision-Language-Action Model for Autonomous Driving via Online Reinforcement Learning.

RAD-2: Scaling Reinforcement Learning in a Generator-Discriminator Framework MindDrive: A Vision-Language-Action Model for Autonomous Driving via Online Reinforcement Learning

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-07-20T02:18:22.862599Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T11:40:26.649975Z digest=sha256:251de1e60d7f5941ddb204bb15db31fce5fde0c032bd6ef91b34007fed1ee6de

Observation fc49d740-28e1-4fa7-91d7-4034eefa377b · outbound

This paper cites Rad: Training an end-to-end driving policy via large-scale 3dgs-based reinforcement learning.

RAD-2: Scaling Reinforcement Learning in a Generator-Discriminator Framework Rad: Training an end-to-end driving policy via large-scale 3dgs-based reinforcement learning

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-10T11:50:21.049149Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T11:40:26.649975Z digest=sha256:6470ecf3c03fbe885468044240f8d7cf7142a0355840076379ea55fe4aef8f6f

Observation 80eb5342-4900-445e-8868-3e4d60b7950e · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

RAD-2: Scaling Reinforcement Learning in a Generator-Discriminator Framework DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 8

Resolution
verified exact
local_arxiv, observed 2026-05-10T11:50:21.064344Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T11:40:26.649975Z digest=sha256:c0fc30b09e1bfc36c26f1dd5425136e324256603c4b01f683a38b475f4a3ff50

Observation 1d83db5a-3405-479e-97ba-50cdab6d643c · outbound

This paper cites iPad: Iterative Proposal-centric End-to-End Autonomous Driving.

RAD-2: Scaling Reinforcement Learning in a Generator-Discriminator Framework iPad: Iterative Proposal-centric End-to-End Autonomous Driving

Reference 9

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T11:50:21.056914Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T11:40:26.649975Z digest=sha256:e53cfc2724639e51084c7bbeb3c7670fac9a23b9f04f7843da5da6862cf47125

Observation 16eb8e52-696c-4f20-ad40-8be71aa10b6b · outbound

This paper cites Denoising dif- fusion probabilistic models.Advances in neural information processing systems, 33:6840–6851.

RAD-2: Scaling Reinforcement Learning in a Generator-Discriminator Framework Denoising dif- fusion probabilistic models.Advances in neural information processing systems, 33:6840–6851

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-05-19T13:42:19.948732Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T11:40:26.649975Z digest=sha256:fa48deefddb9c01d65ff5abe9eb1af62fb35b6cb09cb9d45b65aacc5791abe26

Observation b2802603-c05d-4535-b324-e3b446ad262d · outbound

This paper cites GAIA-1: A Generative World Model for Autonomous Driving.

RAD-2: Scaling Reinforcement Learning in a Generator-Discriminator Framework GAIA-1: A Generative World Model for Autonomous Driving

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-12T07:15:10.779550Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T11:40:26.649975Z digest=sha256:1eee377be51ee1b4658941eef5ed1043940036e7ecc2bb4fb728d5266b78afdd

Observation 8b408c0a-1490-4320-8198-03231d5407de · outbound

This paper cites Planning-oriented autonomous driving.

RAD-2: Scaling Reinforcement Learning in a Generator-Discriminator Framework Planning-oriented autonomous driving

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-05-19T13:42:19.943020Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T11:40:26.649975Z digest=sha256:824a163ad768e30cb4bd2162bd1c0eaf206baed8ee02fece1a8213a0ab25104d

Observation 3548b779-b234-4869-80b6-91adb1501a60 · outbound

This paper cites Efficient deep reinforcement learning with imitative expert priors for au- tonomous driving.IEEE Transactions on Neural Networks and Learning Systems, 34(10):7391–7403.

RAD-2: Scaling Reinforcement Learning in a Generator-Discriminator Framework Efficient deep reinforcement learning with imitative expert priors for au- tonomous driving.IEEE Transactions on Neural Networks and Learning Systems, 34(10):7391–7403

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-05-19T13:42:19.945020Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T11:40:26.649975Z digest=sha256:5c72497a4f28f489f49a9113b8795553badc382a359b6cdaad2c135e6d899b4b

Observation 70ba084a-20ee-4e53-ab87-c041e02fb532 · outbound

This paper cites Spatial transformer networks.Advances in neural informa- tion processing systems, 28.

RAD-2: Scaling Reinforcement Learning in a Generator-Discriminator Framework Spatial transformer networks.Advances in neural informa- tion processing systems, 28

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-05-19T13:42:19.958457Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T11:40:26.649975Z digest=sha256:4fcf8dcdfc8d0a980c3012f99008b53d3cd528c2bb56b4a8a8ab157adb0af4a8

Observation 9c2763b5-0114-431c-8799-11ea5c3d6578 · outbound

This paper cites OpenAI o1 System Card.

RAD-2: Scaling Reinforcement Learning in a Generator-Discriminator Framework OpenAI o1 System Card

Reference 15

Resolution
verified exact
local_arxiv, observed 2026-05-10T11:50:21.098215Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T11:40:26.649975Z digest=sha256:751155dbeba6e19406544b5bafecc058c283f92c9238f0f5d0aedfaccc2d4846

Observation 39d8feef-cfeb-46f1-908b-6b481ed0f38a · outbound

This paper cites Vad: Vectorized scene representa- tion for efficient autonomous driving.

RAD-2: Scaling Reinforcement Learning in a Generator-Discriminator Framework Vad: Vectorized scene representa- tion for efficient autonomous driving

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-05-19T13:42:19.946880Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T11:40:26.649975Z digest=sha256:3024c40bc42f39fa77bd564506f4c5667d2c52284d86f0f33400f04f9c093a69

Observation 569dbf75-583f-4a1b-a9e2-508011a630fb · outbound

This paper cites Senna: Bridging Large Vision-Language Models and End-to-End Autonomous Driving.

RAD-2: Scaling Reinforcement Learning in a Generator-Discriminator Framework Senna: Bridging Large Vision-Language Models and End-to-End Autonomous Driving

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-05-15T15:24:24.181007Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T11:40:26.649975Z digest=sha256:2ad1fbc988ef1ba0f86d0554f3f2d3e5443f2f154d67657f770127fe6735b713

Observation 4770d6a5-9a45-489f-aa13-5625ee58a909 · outbound

This paper cites AlphaDrive: Unleashing the Power of VLMs in Autonomous Driving via Reinforcement Learning and Reasoning.

RAD-2: Scaling Reinforcement Learning in a Generator-Discriminator Framework AlphaDrive: Unleashing the Power of VLMs in Autonomous Driving via Reinforcement Learning and Reasoning

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-05-16T20:06:27.342141Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T11:40:26.649975Z digest=sha256:cf5b9c26059e92767a55ae571b92d61ce2da100278704061623786fc683dd15c

Observation a7875584-d594-4914-b0d9-2090c549625c · outbound

This paper cites Learning to drive in a day.

RAD-2: Scaling Reinforcement Learning in a Generator-Discriminator Framework Learning to drive in a day

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-05-19T13:42:19.934668Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T11:40:26.649975Z digest=sha256:5414cc650e4b11d13e6fbe259143bb223755ca6b07e9227068fee6ce93d5aaaa

Observation d748a0f8-3a40-4758-8e41-0bf85737e925 · outbound

This paper cites 3d gaussian splatting for real-time radiance field rendering.ACM Trans.

RAD-2: Scaling Reinforcement Learning in a Generator-Discriminator Framework 3d gaussian splatting for real-time radiance field rendering.ACM Trans

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-05-19T13:42:19.936700Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T11:40:26.649975Z digest=sha256:789c2ca319ded685f293a5ea88c76449b457bbb06543a39ddd89205bb750401c

Observation d56b62e9-6713-4cd1-9581-b6e503732534 · outbound

This paper cites Refining dif- fusion planner for reliable behavior synthesis by automatic detection of infeasible plans.Advances in Neural Informa- tion Processing Systems, 36:24223–24246.

RAD-2: Scaling Reinforcement Learning in a Generator-Discriminator Framework Refining dif- fusion planner for reliable behavior synthesis by automatic detection of infeasible plans.Advances in Neural Informa- tion Processing Systems, 36:24223–24246

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-05-19T13:42:19.930951Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T11:40:26.649975Z digest=sha256:3fa9a8266096e415f54f45c9f70873fa65d7641312d6090db990c4e4a193b315

Observation 0ed22d82-7023-4b08-bbd6-f7433d96b1a5 · outbound

This paper cites Reinforcement Learning and Control as Probabilistic Inference: Tutorial and Review.

RAD-2: Scaling Reinforcement Learning in a Generator-Discriminator Framework Reinforcement Learning and Control as Probabilistic Inference: Tutorial and Review

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-05-13T18:27:03.097041Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T11:40:26.649975Z digest=sha256:5366a5f1f48a4675a554d1f7608a8e60380fa49df79b677e64c3bffbf39dae42

Observation 4f1d6be8-e7eb-4c6a-aa5d-2907f89ace15 · outbound

This paper cites OmniNWM: Omniscient Driving Navigation World Models.

RAD-2: Scaling Reinforcement Learning in a Generator-Discriminator Framework OmniNWM: Omniscient Driving Navigation World Models

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-06-23T03:12:33.346294Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T11:40:26.649975Z digest=sha256:2b2f295e7601c0cc1d3c88d7c227ffe0ebc38d80e6e2adb2dcbba79725450376

Observation 5176e58d-a71c-462e-bd17-fa5c7e09eb5e · outbound

This paper cites Hydra-MDP++: Advancing End-to-End Driving via Expert-Guided Hydra-Distillation.

RAD-2: Scaling Reinforcement Learning in a Generator-Discriminator Framework Hydra-MDP++: Advancing End-to-End Driving via Expert-Guided Hydra-Distillation

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-05-10T11:50:21.117631Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T11:40:26.649975Z digest=sha256:b575982399ac20e12738c256e2b747d66b29007e39d2b20dbf46a0d64997539e

Observation 77658704-a00f-465e-b097-1037e2b88b7c · outbound

This paper cites Reinforcement Learning with Action Chunking.

RAD-2: Scaling Reinforcement Learning in a Generator-Discriminator Framework Reinforcement Learning with Action Chunking

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-05-12T03:42:01.554293Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T11:40:26.649975Z digest=sha256:c9630135952b4b66a9c1d2ae56f0e01577c62b6d9002454a0ff967307f18e17d

Observation e70560a9-a09a-419a-93da-e5ce0af06da3 · outbound

This paper cites Back to Basics: Let Denoising Generative Models Denoise.

RAD-2: Scaling Reinforcement Learning in a Generator-Discriminator Framework Back to Basics: Let Denoising Generative Models Denoise

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-05-11T22:16:43.095562Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T11:40:26.649975Z digest=sha256:2bd22a75270bfa3c661b5d934c74c179df614c72d275487c66a3bda91159c67c

Observation fca5608c-9329-4f03-b45c-44aa2e34dba7 · outbound

This paper cites Iterative linear quadratic regulator design for nonlinear biological movement systems.

RAD-2: Scaling Reinforcement Learning in a Generator-Discriminator Framework Iterative linear quadratic regulator design for nonlinear biological movement systems

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-05-19T13:42:19.962411Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T11:40:26.649975Z digest=sha256:814c5e7373c6f900feda9486de6f3719a34dae6ffa184b0ebd5c3e60b8c20c75

Observation da3db91f-4450-4c79-a101-06bbe6bca610 · outbound

This paper cites End-to-End Driving with Online Trajectory Evaluation via BEV World Model.

RAD-2: Scaling Reinforcement Learning in a Generator-Discriminator Framework End-to-End Driving with Online Trajectory Evaluation via BEV World Model

Reference 28

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T11:50:21.125285Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T11:40:26.649975Z digest=sha256:86a313111dcd31124d502adc146ea59e7898038dd0f5fb0fc3c8e93b5c8aff87

Observation 868d851c-5245-4722-b7cb-f0bdc877a8b6 · outbound

This paper cites ReCogDrive: A Reinforced Cognitive Framework for End-to-End Autonomous Driving.

RAD-2: Scaling Reinforcement Learning in a Generator-Discriminator Framework ReCogDrive: A Reinforced Cognitive Framework for End-to-End Autonomous Driving

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-05-15T07:36:24.555133Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T11:40:26.649975Z digest=sha256:66a3a0f23950493e4e6388487a60843f5b98b908e44087e17919de5f82c505ee

Observation bef77172-a39c-4f90-8318-603eb9ce6374 · outbound

This paper cites Hydra-MDP: End-to-end Multimodal Planning with Multi-target Hydra-Distillation.

RAD-2: Scaling Reinforcement Learning in a Generator-Discriminator Framework Hydra-MDP: End-to-end Multimodal Planning with Multi-target Hydra-Distillation

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-05-13T23:13:55.802367Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T11:40:26.649975Z digest=sha256:d409fe7fc4f60c33d301f4d26761cf30fe7d2480f8a7861051cf2f62f667b002

Observation 6d3dca14-9aa3-4b72-9527-3c5077ffc45d · outbound

This paper cites an unresolved cited work.

RAD-2: Scaling Reinforcement Learning in a Generator-Discriminator Framework Unresolved cited work

Reference 31

Resolution
unresolved
raw_fallback, observed 2026-05-19T13:42:19.932827Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T11:40:26.649975Z digest=sha256:4f0ff9ffffb249b91ff2a860f9aa7466c5bcee879779ce8af5500e4049d6b93c

Observation 4b5101a2-0834-4ec0-bb95-8f160defba92 · outbound

This paper cites Zero-Human Demonstration End-to-end Autonomous Driving with Trajectory Scorer.

RAD-2: Scaling Reinforcement Learning in a Generator-Discriminator Framework Zero-Human Demonstration End-to-end Autonomous Driving with Trajectory Scorer

Reference 32

Resolution
verified exact
arxiv_id, observed 2026-07-09T02:19:48.260809Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T11:40:26.649975Z digest=sha256:4344ce375556fbd57442a3fc0b7522c82107f2bc1a80b3ce5847ac487fb1f4bf

Observation b593187a-3442-42c8-8df4-9db4c667bd30 · outbound

This paper cites Generalized Trajectory Scoring for End-to-end Multimodal Planning.

RAD-2: Scaling Reinforcement Learning in a Generator-Discriminator Framework Generalized Trajectory Scoring for End-to-end Multimodal Planning

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-05-10T11:50:21.075872Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T11:40:26.649975Z digest=sha256:cbd4aa1a897949bd9c73c6afa253b5917b492aa0d35f73643fa302a5e140a7e4

Observation c0863630-3113-4b41-ad4e-dc60c09b293f · outbound

This paper cites Fine-tuning is Not Enough: A Parallel Framework for Collaborative Imitation and Reinforcement Learning in End-to-end Autonomous Driving.

RAD-2: Scaling Reinforcement Learning in a Generator-Discriminator Framework Fine-tuning is Not Enough: A Parallel Framework for Collaborative Imitation and Reinforcement Learning in End-to-end Autonomous Driving

Reference 34

Resolution
verified exact
local_arxiv, observed 2026-05-10T11:50:21.121209Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T11:40:26.649975Z digest=sha256:7f92dd0bcf69757d86f0a050ee0a9f17160810b022909205d72b157009d99d0c

Observation 20618d5e-cd4f-4b22-827f-9c0f9dcfea21 · outbound

This paper cites Cirl: Controllable imitative reinforcement learning for vision-based self-driving.

RAD-2: Scaling Reinforcement Learning in a Generator-Discriminator Framework Cirl: Controllable imitative reinforcement learning for vision-based self-driving

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-05-19T13:42:19.960509Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T11:40:26.649975Z digest=sha256:d4fe02bb60d5338b7070aa930d409a07a80b9881c980ef12e0bc11a751986661

Observation f9b14ead-3dcd-4472-a0ac-f427fe1325db · outbound

This paper cites Diffusiondrive: Truncated diffusion model for end-to-end autonomous driving.

RAD-2: Scaling Reinforcement Learning in a Generator-Discriminator Framework Diffusiondrive: Truncated diffusion model for end-to-end autonomous driving

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-05-19T13:42:19.950598Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T11:40:26.649975Z digest=sha256:dfdc1c02424d74e17a2f5b6f9bcd0646272a1ed144360313a24849195ea4fdcf

Observation 0ae4f41e-7c48-42c0-8ebc-876224a2831f · outbound

This paper cites Continuous control with deep reinforcement learning.

RAD-2: Scaling Reinforcement Learning in a Generator-Discriminator Framework Continuous control with deep reinforcement learning

Reference 37

Resolution
verified exact
arxiv_id, observed 2026-05-11T15:43:36.523954Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T11:40:26.649975Z digest=sha256:4ddd7f7229b0b1fb68d926ffe527576083974b42af45b5169dc12fe88034d807

Observation ac48b991-b899-43d0-a6e8-55ff5e622b0b · outbound

This paper cites Reinforced Refinement with Self-Aware Expansion for End-to-End Autonomous Driving.

RAD-2: Scaling Reinforcement Learning in a Generator-Discriminator Framework Reinforced Refinement with Self-Aware Expansion for End-to-End Autonomous Driving

Reference 38

Resolution
verified exact
arxiv_id, observed 2026-05-10T11:50:21.101934Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T11:40:26.649975Z digest=sha256:2764946dbed95bbae216b5486e2cefe884966c0d3cd87c4611f812b53b2cd373

Observation 984f6165-9b3d-4c99-8d80-16f4ec300a04 · outbound

This paper cites Imitation is not enough: Robustifying imitation with reinforcement learn- ing for challenging driving scenarios.

RAD-2: Scaling Reinforcement Learning in a Generator-Discriminator Framework Imitation is not enough: Robustifying imitation with reinforcement learn- ing for challenging driving scenarios

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-05-19T13:42:19.938636Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T11:40:26.649975Z digest=sha256:3902041786b62409b2a6a17d581de92342cdef2b8e02a54184fa907251f5be86

Observation bec0fbcf-48ea-4449-8ea9-80319859acf8 · outbound

This paper cites ReconDreamer-RL: Enhancing Reinforcement Learning via Diffusion-based Scene Reconstruction.

RAD-2: Scaling Reinforcement Learning in a Generator-Discriminator Framework ReconDreamer-RL: Enhancing Reinforcement Learning via Diffusion-based Scene Reconstruction

Reference 40

Resolution
verified exact
arxiv_id, observed 2026-05-10T11:50:21.155676Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T11:40:26.649975Z digest=sha256:26e1da0fefaa61f640684eb29d3adc0b2e11581796edb1a19d27660eda5b5014

Observation 743b89a7-c38e-4a68-8787-7823417a73f6 · outbound

This paper cites Lift, splat, shoot: Encoding images from arbitrary camera rigs by implicitly unprojecting to 3d.

RAD-2: Scaling Reinforcement Learning in a Generator-Discriminator Framework Lift, splat, shoot: Encoding images from arbitrary camera rigs by implicitly unprojecting to 3d

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-05-19T13:42:19.928932Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T11:40:26.649975Z digest=sha256:f4d45b0bb1322085d074741e71f4e21fb929252fa5148b3ec34de57be5b89f5d

Observation d3bcc715-ef98-474e-95c5-44b690ea2303 · outbound

This paper cites Proximal Policy Optimization Algorithms.

RAD-2: Scaling Reinforcement Learning in a Generator-Discriminator Framework Proximal Policy Optimization Algorithms

Reference 42

Resolution
verified exact
local_arxiv, observed 2026-05-10T11:50:21.178700Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T11:40:26.649975Z digest=sha256:6b36f3cd5c3602c9e519b3e2ca205d00f05e550cafe296ccb0e2477f0da1a6ae

Observation 1855e6ef-8bea-4628-a884-2f5d5ad16837 · outbound

This paper cites Drivedpo: Policy learning via safety dpo for end-to-end autonomous driving.

RAD-2: Scaling Reinforcement Learning in a Generator-Discriminator Framework Drivedpo: Policy learning via safety dpo for end-to-end autonomous driving

Reference 43

Resolution
verified exact
arxiv_id, observed 2026-05-10T11:50:21.147098Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T11:40:26.649975Z digest=sha256:228912cacc7aeb719521cd7c11e90bf5b2195aa36f3ef762af8400fd727bad4e

Observation 2ee83fe9-c1a6-4a79-86e3-38f11ecea743 · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

RAD-2: Scaling Reinforcement Learning in a Generator-Discriminator Framework DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 44

Resolution
verified exact
local_arxiv, observed 2026-05-10T11:50:21.143111Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T11:40:26.649975Z digest=sha256:88e066438c7fa6a6f3c9f7a0dd5b13e6479bc520f127d6666183ffc594dd6b8f

Observation b79b221d-8bfa-4951-a37c-73dcbf68dac0 · outbound

This paper cites Senna-2: Aligning vlm and end-to-end driving policy for consistent decision making and planning.

RAD-2: Scaling Reinforcement Learning in a Generator-Discriminator Framework Senna-2: Aligning vlm and end-to-end driving policy for consistent decision making and planning

Reference 45

Resolution
verified exact
arxiv_id, observed 2026-05-10T11:50:21.186727Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T11:40:26.649975Z digest=sha256:2491dd0f5a2f9aad68df678ac8122b85256a1df3863b42b3b3e4cd2e04f9bf30

Observation 96d24e68-a977-4ed6-a8b7-7e38d4c183f7 · outbound

This paper cites Sparsedrivev2: Scoring is all you need for end-to-end autonomous driving.

RAD-2: Scaling Reinforcement Learning in a Generator-Discriminator Framework Sparsedrivev2: Scoring is all you need for end-to-end autonomous driving

Reference 46

Resolution
verified exact
arxiv_id, observed 2026-05-10T11:50:21.175328Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T11:40:26.649975Z digest=sha256:73af15a97faf198bf5c274523925560cdef50f744ded8b32ae931aec84726c2a

Observation c628a69a-466d-46eb-a2eb-96448df2df94 · outbound

This paper cites Driving into the future: Multiview visual forecasting and planning with world model for au- tonomous driving.

RAD-2: Scaling Reinforcement Learning in a Generator-Discriminator Framework Driving into the future: Multiview visual forecasting and planning with world model for au- tonomous driving

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-05-19T13:42:19.924714Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T11:40:26.649975Z digest=sha256:9c00fe4ba8d14e7ed433af0fc7b4189d2aeb062a7b67ac18965bf757de29c24a

Observation 679a7699-20e5-4e0c-9c69-1b6194e38217 · outbound

This paper cites Driving into the future: Multiview visual forecasting and planning with world model for au- tonomous driving.

RAD-2: Scaling Reinforcement Learning in a Generator-Discriminator Framework Driving into the future: Multiview visual forecasting and planning with world model for au- tonomous driving

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-05-19T13:42:19.926618Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T11:40:26.649975Z digest=sha256:71b0d014afd69477f3fc589778a8b3633692099ac794ea9da0fbfb9f40b715cf

Observation 1d44b207-bfd9-432f-a151-4ef13a81ccc2 · outbound

This paper cites Para-drive: Parallelized architecture for real- time autonomous driving.

RAD-2: Scaling Reinforcement Learning in a Generator-Discriminator Framework Para-drive: Parallelized architecture for real- time autonomous driving

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-05-19T13:42:19.964237Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T11:40:26.649975Z digest=sha256:64cbee8dbdba6bfd82eee6b03828c6f8438425836853908dacdaad03b4e117c2

Observation 3408d938-2324-4c75-b29d-9b505b746050 · outbound

This paper cites DriveLaW:Unifying Planning and Video Generation in a Latent Driving World.

RAD-2: Scaling Reinforcement Learning in a Generator-Discriminator Framework DriveLaW:Unifying Planning and Video Generation in a Latent Driving World

Reference 50

Resolution
verified exact
local_arxiv, observed 2026-05-10T11:50:21.204147Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T11:40:26.649975Z digest=sha256:2d221bba8d7a3c8d9e06207f2a28079445232bf99d52efcb583ef3c3d0a43bf1

Observation 74b02d06-d359-474a-9226-ae2d2378bc19 · outbound

This paper cites Ad-r1: Closed-loop reinforcement learning for end-to-end autonomous driving with impartial world models.

RAD-2: Scaling Reinforcement Learning in a Generator-Discriminator Framework Ad-r1: Closed-loop reinforcement learning for end-to-end autonomous driving with impartial world models

Reference 51

Resolution
verified exact
arxiv_id, observed 2026-05-10T11:50:21.189985Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T11:40:26.649975Z digest=sha256:08e3646f080cb390a46c7f1da5c5e254677813e6dedb0264219b1a084e3a160a

Observation 6aa3edd0-f19a-4655-9dbd-d72de32df77d · outbound

This paper cites Worldrft: Latent world model planning with reinforcement fine-tuning for autonomous driving.

RAD-2: Scaling Reinforcement Learning in a Generator-Discriminator Framework Worldrft: Latent world model planning with reinforcement fine-tuning for autonomous driving

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-05-19T13:42:19.952425Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T11:40:26.649975Z digest=sha256:7e809befc366c4b05f111e6db38bccf4e99164e280b85b6f580d5b8cda3871d0

Observation 216911d1-6cc0-4cf6-85b8-172bf4654b73 · outbound

This paper cites Dreamerad: Efficient re- inforcement learning via latent world model for autonomous driving.

RAD-2: Scaling Reinforcement Learning in a Generator-Discriminator Framework Dreamerad: Efficient re- inforcement learning via latent world model for autonomous driving

Reference 53

Resolution
verified exact
arxiv_id, observed 2026-05-10T11:50:21.171465Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T11:40:26.649975Z digest=sha256:21021cdfd1d2db5cc6fcd80757919a66b84b7c397e15f91e5d593f9e0e3ca710

Observation 7a1271f1-1074-4bc1-86ee-efaf29257272 · outbound

This paper cites Drivesuprim: Towards precise trajectory selection for end-to-end planning.

RAD-2: Scaling Reinforcement Learning in a Generator-Discriminator Framework Drivesuprim: Towards precise trajectory selection for end-to-end planning

Reference 54

Resolution
verified exact
arxiv_id, observed 2026-05-10T11:50:21.182568Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T11:40:26.649975Z digest=sha256:71e3aad63f6608d80b2c7d8b0274f3625c274d311006f55b4b42921a6da0da61

Observation 419b87fb-db5d-4124-be81-f42815cd8489 · outbound

This paper cites DAPO: An Open-Source LLM Reinforcement Learning System at Scale.

RAD-2: Scaling Reinforcement Learning in a Generator-Discriminator Framework DAPO: An Open-Source LLM Reinforcement Learning System at Scale

Reference 55

Resolution
verified exact
local_arxiv, observed 2026-05-10T11:50:21.197191Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T11:40:26.649975Z digest=sha256:bf1c7deadfe33efc37b7e8b16911add73a1a6a9e12776cbcffc01ec96d844d1d

Observation cf179533-b3ae-4a4f-950b-57e98827f6f0 · outbound

This paper cites Generative Planning for Temporally Coordinated Exploration in Reinforcement Learning.

RAD-2: Scaling Reinforcement Learning in a Generator-Discriminator Framework Generative Planning for Temporally Coordinated Exploration in Reinforcement Learning

Reference 56

Resolution
verified exact
arxiv_id, observed 2026-05-10T11:50:21.167640Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T11:40:26.649975Z digest=sha256:3380d6215fb7b9941c5d69542b41e4250f05b473e60cdd7d787a35058d8a2952

Observation 9c1cbdf6-983c-4f5a-a4ec-3a0e1ada0167 · outbound

This paper cites Drivedreamer4d: World models are effective data machines for 4d driving scene rep- resentation.

RAD-2: Scaling Reinforcement Learning in a Generator-Discriminator Framework Drivedreamer4d: World models are effective data machines for 4d driving scene rep- resentation

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-05-19T13:42:19.967956Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T11:40:26.649975Z digest=sha256:abf50e720a11f215ade252ab8e0d71188e2fa483bc6aeb952b9b2fa9ac820142

Observation a64745a4-ee6f-49d6-8778-4e83741f652d · outbound

This paper cites ReconDreamer++: Harmonizing Generative and Reconstructive Models for Driving Scene Representation.

RAD-2: Scaling Reinforcement Learning in a Generator-Discriminator Framework ReconDreamer++: Harmonizing Generative and Reconstructive Models for Driving Scene Representation

Reference 58

Resolution
verified exact
arxiv_id, observed 2026-05-10T11:50:21.200657Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T11:40:26.649975Z digest=sha256:1a2b9113f7ec29c75c65c846772cc9de66526739a28964370df88dd370916b87

Observation e4390d3e-6b73-43ce-8830-3e51dceb0bb8 · outbound

This paper cites Group Sequence Policy Optimization.

RAD-2: Scaling Reinforcement Learning in a Generator-Discriminator Framework Group Sequence Policy Optimization

Reference 59

Resolution
verified exact
arxiv_id, observed 2026-05-10T19:22:54.227899Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T11:40:26.649975Z digest=sha256:df27ca7455b386f5708b3a2b37e600f508e6850f17e487b50421e19d22ffaa7a

Observation e32df56e-55aa-4f52-85a2-65fbaa968f02 · outbound

This paper cites Genad: Generative end-to-end au- tonomous driving.

RAD-2: Scaling Reinforcement Learning in a Generator-Discriminator Framework Genad: Generative end-to-end au- tonomous driving

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-05-19T13:42:19.966099Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T11:40:26.649975Z digest=sha256:13f7e79b245ac9d9cb10000230684325d041de58eed326081da8ba39f4b47d2d

Observation 6c264c51-5309-4664-9bcf-d5ea6ddc073f · outbound

This paper cites Diffusion-Based Planning for Autonomous Driving with Flexible Guidance.

RAD-2: Scaling Reinforcement Learning in a Generator-Discriminator Framework Diffusion-Based Planning for Autonomous Driving with Flexible Guidance

Reference 61

Resolution
verified exact
arxiv_id, observed 2026-05-10T11:50:21.159683Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T11:40:26.649975Z digest=sha256:7e540f62b337d0da81930e85b80c0ef421f679cd0e542b5150725960d6b01d9f

Observation 55dcca5a-0239-424b-b7e9-ec7dfd902dc4 · outbound

This paper cites Unleashing the Potential of Diffusion Models for End-to-End Autonomous Driving.

RAD-2: Scaling Reinforcement Learning in a Generator-Discriminator Framework Unleashing the Potential of Diffusion Models for End-to-End Autonomous Driving

Reference 62

Resolution
verified exact
arxiv_id, observed 2026-05-20T00:03:06.805723Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T11:40:26.649975Z digest=sha256:6fce407d5756144b86b03f38adfa2dfb3ea01ef81156544cba969ef365d69dc8

Observation 4f97e69b-e206-408d-acc2-3bb63369d86e · outbound

This paper cites Resad: Normalized residual trajectory modeling for end-to-end autonomous driving.arXiv preprint arXiv:2510.08562.

RAD-2: Scaling Reinforcement Learning in a Generator-Discriminator Framework Resad: Normalized residual trajectory modeling for end-to-end autonomous driving.arXiv preprint arXiv:2510.08562

Reference 63

Resolution
verified exact
arxiv_id, observed 2026-05-10T11:50:21.151276Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T11:40:26.649975Z digest=sha256:f822892580c926795608e7859e1f73dd7e469c0e8c37b851aafe1ab969cbab1f

Observation 93b8efd6-75ca-4464-b539-da29aae3e44c · outbound

This paper cites HUGSIM: A Real-Time, Photo-Realistic and Closed-Loop Simulator for Autonomous Driving.

RAD-2: Scaling Reinforcement Learning in a Generator-Discriminator Framework HUGSIM: A Real-Time, Photo-Realistic and Closed-Loop Simulator for Autonomous Driving

Reference 64

Resolution
verified exact
arxiv_id, observed 2026-05-10T11:50:21.134298Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T11:40:26.649975Z digest=sha256:d6afebb54b9dc9b9f235b330eed6686363c86ca1b9d438dc14f019dca1c13fc7

Observation 58949104-f078-4e4d-a38f-da0d850632ba · outbound

This paper cites SMARTS: Scalable Multi-Agent Reinforcement Learning Training School for Autonomous Driving.

RAD-2: Scaling Reinforcement Learning in a Generator-Discriminator Framework SMARTS: Scalable Multi-Agent Reinforcement Learning Training School for Autonomous Driving

Reference 65

Resolution
verified exact
arxiv_id, observed 2026-05-10T11:50:21.194000Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T11:40:26.649975Z digest=sha256:2a2c45ee0611b21914cda080833a076f8a5392b462aad6dbeae1ef8c1ddb48bf

Observation 29082637-b118-4b0a-a719-402304e5f8fe · outbound

This paper cites AutoVLA: A Vision-Language-Action Model for End-to-End Autonomous Driving with Adaptive Reasoning and Reinforcement Fine-Tuning.

RAD-2: Scaling Reinforcement Learning in a Generator-Discriminator Framework AutoVLA: A Vision-Language-Action Model for End-to-End Autonomous Driving with Adaptive Reasoning and Reinforcement Fine-Tuning

Reference 66

Resolution
verified exact
arxiv_id, observed 2026-05-14T21:46:44.401546Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T11:40:26.649975Z digest=sha256:457920b10ffb6f52ed482c49f41549fb261bb7b0541a910a4d7c83bfebcf3eef

Observation c1670164-64e3-4a9c-a313-cbe79657ab58 · outbound

This paper cites re- gions important for driving.

RAD-2: Scaling Reinforcement Learning in a Generator-Discriminator Framework re- gions important for driving

Reference 67

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T11:50:21.052852Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T11:40:26.649975Z digest=sha256:bae142b378f6729362f3bbf0e71983762ab6d56e7dc9d0900b61ce2b0cae27c6

Pith citing papers

Observation dff84d96-0b57-448f-a037-f7452849df0d · inbound

CRAFT: Counterfactual-to-Interactive Reinforcement Fine-Tuning for Driving Policies cites this paper.

CRAFT: Counterfactual-to-Interactive Reinforcement Fine-Tuning for Driving Policies RAD-2: Scaling Reinforcement Learning in a Generator-Discriminator Framework

Reference 27

Resolution
verified exact
local_arxiv, observed 2026-05-11T17:41:06.621120Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-08T17:20:55.364175Z digest=sha256:5feccde1ff25ebe82258670a9dbec658462de9782c8845b07897c06a7dc2ac47

Observation c6a744b5-127e-45c8-ac57-7e87bab40e60 · inbound

Action Emergence from Streaming Intent cites this paper.

Action Emergence from Streaming Intent RAD-2: Scaling Reinforcement Learning in a Generator-Discriminator Framework

Reference 35

Resolution
metadata mismatch
local_arxiv, observed 2026-05-14T21:02:59.140621Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-14T20:59:42.456958Z digest=sha256:2a0784153df3d2dce0486f45d5048c8dbf09eb948d4b6bc6df8a8feda98548c0

Observation 139ed23e-ac61-40bc-9be3-8b9a2ac03116 · inbound

Action Emergence from Streaming Intent cites this paper.

Action Emergence from Streaming Intent RAD-2: Scaling Reinforcement Learning in a Generator-Discriminator Framework

Reference 35

Resolution
metadata mismatch
local_arxiv, observed 2026-05-15T05:15:03.010464Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-15T05:13:44.795483Z digest=sha256:5675d806c69069dbfdcde0f4c729f2f30cce9c295969537fa51c612b8997c9c3

Observation 91497563-66f8-4f8f-b736-bffaf38cb914 · inbound

Bridging Local Observation and Global Simulation in Closed-Loop Traffic Modeling cites this paper.

Bridging Local Observation and Global Simulation in Closed-Loop Traffic Modeling RAD-2: Scaling Reinforcement Learning in a Generator-Discriminator Framework

Reference 11

Resolution
verified exact
local_arxiv, observed 2026-07-01T10:45:42.633736Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-01T05:09:44.224340Z digest=sha256:2f1344073076722a47319be7077c1f6ab148d769688b2848e56a87d8239e2072