Pith. sign in

Paper Citation Record · LEDGER

From Representational Complementarity to Dual Systems: Synergizing VLM and Vision-Only Backbones for End-to-End Driving

As of 20 August 2026, this Paper Citation Record lists 21 of 21 outbound references and 0 inbound Pith citation observations for arXiv:2602.10719.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2602.10719 v2

Coverage vector

measured 21 of 21 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-04T06:07:10.575230Z

measured 21 of 21 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

21 of 21 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved21
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 43feb6cd-7ddf-42f2-babb-c72b2997c47a · outbound

This paper cites Is a 3D-Tokenized LLM the Key to Reliable Autonomous Driving?.

From Representational Complementarity to Dual Systems: Synergizing VLM and Vision-Only Backbones for End-to-End Driving Is a 3D-Tokenized LLM the Key to Reliable Autonomous Driving?

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-04T06:07:07.593375Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T06:07:07.593375Z digest=sha256:bb8274b2a362465d265bc4699c10a9802967cf6f86ce2deacc4dee71be0e2f70

Observation 6adc54cf-6f3f-4c74-8727-90923c910f12 · outbound

This paper cites iPad: Iterative Proposal-centric End-to-End Autonomous Driving.

From Representational Complementarity to Dual Systems: Synergizing VLM and Vision-Only Backbones for End-to-End Driving iPad: Iterative Proposal-centric End-to-End Autonomous Driving

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-04T06:07:08.238644Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T06:07:08.238644Z digest=sha256:bc1604763a9b7d3837be2089af91daa9dfe896eb610950e6fbf06ac694b908c6

Observation 59cf92a5-7e6e-4af5-959e-72f8d2a590b3 · outbound

This paper cites EMMA: End-to-End Multimodal Model for Autonomous Driving.

From Representational Complementarity to Dual Systems: Synergizing VLM and Vision-Only Backbones for End-to-End Driving EMMA: End-to-End Multimodal Model for Autonomous Driving

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-04T06:07:08.383889Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T06:07:08.383889Z digest=sha256:4b3e991e31a3ad7c424b2138288cccb99082d331047ea1f76e6e14e946bbe25b

Observation 5972ab4a-f7d4-49ba-ba11-0f6a6ddb0b40 · outbound

This paper cites Senna: Bridging Large Vision-Language Models and End-to-End Autonomous Driving.

From Representational Complementarity to Dual Systems: Synergizing VLM and Vision-Only Backbones for End-to-End Driving Senna: Bridging Large Vision-Language Models and End-to-End Autonomous Driving

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-04T06:07:08.496067Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T06:07:08.496067Z digest=sha256:83cd5542408fce477c1325f98cb88bc9d413782435e899ae32b5eef81289c554

Observation 905b0922-fe69-4b3a-9d06-0aebd8182a00 · outbound

This paper cites Driving on registers.arXiv preprint arXiv:2601.05083,.

From Representational Complementarity to Dual Systems: Synergizing VLM and Vision-Only Backbones for End-to-End Driving Driving on registers.arXiv preprint arXiv:2601.05083,

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-04T06:07:08.603921Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T06:07:08.603921Z digest=sha256:92831c57ebfdb838f6de03ff7ebbbbfb538c32bc9d3ca4b00d082f5d910b7ab6

Observation 03d3ef13-b7f2-426f-b044-eff232a4788b · outbound

This paper cites DiffusionDrive: Truncated Diffusion Model for End-to-End Autonomous Driving.

From Representational Complementarity to Dual Systems: Synergizing VLM and Vision-Only Backbones for End-to-End Driving DiffusionDrive: Truncated Diffusion Model for End-to-End Autonomous Driving

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-04T06:07:08.966984Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T06:07:08.966984Z digest=sha256:ae37a429aca3fa54c1897b34eec0174dc90e6dbd424035aee291718cfa2a54a7

Observation e266002b-2b20-49ed-bfaa-f055d82514ec · outbound

This paper cites GPT-Driver: Learning to Drive with GPT.

From Representational Complementarity to Dual Systems: Synergizing VLM and Vision-Only Backbones for End-to-End Driving GPT-Driver: Learning to Drive with GPT

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-04T06:07:09.140508Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T06:07:09.140508Z digest=sha256:18656ad00af1cd264cb2a75ace16fa81b87b8ca6bdfc75d9367d689909989a25

Observation 852dc445-123f-4aad-97c6-fa0e80d250bf · outbound

This paper cites Reason2Drive: Towards Interpretable and Chain-based Reasoning for Autonomous Driving.

From Representational Complementarity to Dual Systems: Synergizing VLM and Vision-Only Backbones for End-to-End Driving Reason2Drive: Towards Interpretable and Chain-based Reasoning for Autonomous Driving

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-04T06:07:09.272985Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T06:07:09.272985Z digest=sha256:394717a05db85c19a7ee4d248fefd2181db6ef7ffaa3228cc51b23094f0102cc

Observation 04c9f689-ac34-47c7-a01d-e8440510b18d · outbound

This paper cites A Survey on Vision-Language-Action Models for Autonomous Driving.

From Representational Complementarity to Dual Systems: Synergizing VLM and Vision-Only Backbones for End-to-End Driving A Survey on Vision-Language-Action Models for Autonomous Driving

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-04T06:07:09.468768Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T06:07:09.468768Z digest=sha256:807b1877906ff396f41a1b20be6dc78299ba066b924dc0a9e827eb2031e5757b

Observation e756eab8-0838-4aa4-ba82-eee6c167b0fd · outbound

This paper cites Wang, S., Yu, Z., Jiang, X., Lan, S., Shi, M., Chang, N., Kautz, J., Li, Y ., and Alvarez, J.

From Representational Complementarity to Dual Systems: Synergizing VLM and Vision-Only Backbones for End-to-End Driving Wang, S., Yu, Z., Jiang, X., Lan, S., Shi, M., Chang, N., Kautz, J., Li, Y ., and Alvarez, J

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-04T06:07:09.663477Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T06:07:09.663477Z digest=sha256:7c1a8fd1e2542342a85a63801440dfc0091ba3823ab9b53d6d7ffc8b1abdeeeb

Observation 75dfcd14-b7c2-419b-b79f-22c7a5592e7f · outbound

This paper cites Wam-diff: A masked diffusion vla framework with moe and online reinforce- ment learning for autonomous driving.arXiv preprint arXiv:2512.11872,.

From Representational Complementarity to Dual Systems: Synergizing VLM and Vision-Only Backbones for End-to-End Driving Wam-diff: A masked diffusion vla framework with moe and online reinforce- ment learning for autonomous driving.arXiv preprint arXiv:2512.11872,

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-04T06:07:09.749744Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T06:07:09.749744Z digest=sha256:2fd5c982b7be06d9f4199f657c2a5c2cab7c9b9150962f195ba872ced3225b3d

Observation c2e44f28-07b0-4ea9-bf21-9f51f3188c58 · outbound

This paper cites DRAMA: An Efficient End-to-end Motion Planner for Autonomous Driving with Mamba.

From Representational Complementarity to Dual Systems: Synergizing VLM and Vision-Only Backbones for End-to-End Driving DRAMA: An Efficient End-to-end Motion Planner for Autonomous Driving with Mamba

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-04T06:07:09.820418Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T06:07:09.820418Z digest=sha256:611b82c460f55f43df5ce17e33ee64fa9cc25056507298ff63aeb71bb31c788b

Observation 87d1d350-20c6-4b6e-b2e1-31c4d1a8ee87 · outbound

This paper cites Fine-Grained Evaluation of Large Vision-Language Models in Autonomous Driving.

From Representational Complementarity to Dual Systems: Synergizing VLM and Vision-Only Backbones for End-to-End Driving Fine-Grained Evaluation of Large Vision-Language Models in Autonomous Driving

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-04T06:07:10.012899Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T06:07:10.012899Z digest=sha256:98e12093670a478a14cf2e359fb394959db33dd0d1ff92c16a184d675da0499d

Observation 1a4aa5fe-9ec7-4bfa-8e87-23e337e62e15 · outbound

This paper cites Drive-P2D: A Progressive Perception-to-Decision Benchmark for VLMs in Autonomous Driving.

From Representational Complementarity to Dual Systems: Synergizing VLM and Vision-Only Backbones for End-to-End Driving Drive-P2D: A Progressive Perception-to-Decision Benchmark for VLMs in Autonomous Driving

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-04T06:07:10.099362Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T06:07:10.099362Z digest=sha256:3c5001eb402e5ad2b08d1409482d4df1088aeedf53be9f8f9fa9d5d86dc82f50

Observation f99483ac-4145-48d8-81c4-5621eb0c1358 · outbound

This paper cites Sce2DriveX: A Generalized MLLM Framework for Scene-to-Drive Learning.

From Representational Complementarity to Dual Systems: Synergizing VLM and Vision-Only Backbones for End-to-End Driving Sce2DriveX: A Generalized MLLM Framework for Scene-to-Drive Learning

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-04T06:07:10.218338Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T06:07:10.218338Z digest=sha256:c60f867e51b0123752965986908b70d5cf574b61030770b96ca45e9daa6ac213

Observation 7d52e9d4-252f-46b1-9725-d3c32a6f991f · outbound

This paper cites AutoVLA: A Vision-Language-Action Model for End-to-End Autonomous Driving with Adaptive Reasoning and Reinforcement Fine-Tuning.

From Representational Complementarity to Dual Systems: Synergizing VLM and Vision-Only Backbones for End-to-End Driving AutoVLA: A Vision-Language-Action Model for End-to-End Autonomous Driving with Adaptive Reasoning and Reinforcement Fine-Tuning

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-04T06:07:10.397329Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T06:07:10.397329Z digest=sha256:c544ee4b250c0897fdd4abd0ab1f67df5c242a417867ad8b3fedc43365100433

Observation 10f28bc0-d57d-4069-a387-53ae7baf74fe · outbound

This paper cites Diffusiondrivev2: Rein- forcement learning-constrained truncated diffusion mod- eling in end-to-end autonomous driving.arXiv preprint arXiv:2512.07745,.

From Representational Complementarity to Dual Systems: Synergizing VLM and Vision-Only Backbones for End-to-End Driving Diffusiondrivev2: Rein- forcement learning-constrained truncated diffusion mod- eling in end-to-end autonomous driving.arXiv preprint arXiv:2512.07745,

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-04T06:07:10.575230Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T06:07:10.575230Z digest=sha256:fb4ffaad211161f8d2c9e82295243e3ea1a999e492333bed73826356c03f3a4d

Observation 3d8c5845-fb10-4cb0-ac89-1b1178910c48 · outbound

This paper cites DriveLM: Driving with Graph Visual Question Answering.

From Representational Complementarity to Dual Systems: Synergizing VLM and Vision-Only Backbones for End-to-End Driving DriveLM: Driving with Graph Visual Question Answering

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-04T06:07:07.991904Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T06:07:07.991904Z digest=sha256:10cd565345fbbdd7aca3d7ad21a6594f0139febecb5a501098315e4f98237cc7

Observation 72651e63-9180-4f92-9514-1ac3b6dfdf4b · outbound

This paper cites VADv2: End-to-End Vectorized Autonomous Driving via Probabilistic Planning.

From Representational Complementarity to Dual Systems: Synergizing VLM and Vision-Only Backbones for End-to-End Driving VADv2: End-to-End Vectorized Autonomous Driving via Probabilistic Planning

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-04T06:07:07.706398Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T06:07:07.706398Z digest=sha256:158c9f7b31b0b650211f0bc348c55d4872292cc24d900803aaa1d11c1ffd5862

Observation 16d582f5-0a8b-4957-9a6d-fe77161a0b8f · outbound

This paper cites ORION: A Holistic End-to-End Autonomous Driving Framework by Vision-Language Instructed Action Generation.

From Representational Complementarity to Dual Systems: Synergizing VLM and Vision-Only Backbones for End-to-End Driving ORION: A Holistic End-to-End Autonomous Driving Framework by Vision-Language Instructed Action Generation

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-04T06:07:08.095703Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T06:07:08.095703Z digest=sha256:c832f942cd19db587ddcc78b897ad6e3115c75e93d820732383d48d43671b829

Observation ac926fb9-9d60-4637-9d52-1bbfd0f44f72 · outbound

This paper cites Imagidrive: A unified imagination-and-planning framework for autonomous driving.arXiv preprint arXiv:2508.11428, 2025a.

From Representational Complementarity to Dual Systems: Synergizing VLM and Vision-Only Backbones for End-to-End Driving Imagidrive: A unified imagination-and-planning framework for autonomous driving.arXiv preprint arXiv:2508.11428, 2025a

Reference 2026

Resolution
unresolved
no resolver link, observed 2026-08-04T06:07:08.761015Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T06:07:08.761015Z digest=sha256:4431fa3b91fe31229d62a5c0bbe5c00078e77137a25e7cfd3d3b6b894ebca329

Pith citing papers

No inbound Pith citation observations are available.