Pith. sign in

Paper Citation Record · LEDGER

A Survey on Vision-Language-Action Models for Autonomous Driving

As of 7 August 2026, this Paper Citation Record lists 100 of 166 outbound references and 14 inbound Pith citation observations for arXiv:2506.24044.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.24044 v1

Coverage vector

measured 100 of 166 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T21:31:04.674245Z

measured 114 of 114 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-06T06:34:29.942622+00:00

measured 14 of 14 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T12:39:02.315303Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-05T11:41:02.543365Z

Reference resolution

100 of 166 outbound references displayed

  • verified exact1
  • verified fuzzy0
  • unresolved99
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 95972b4e-af7f-4426-b799-a6a82ba1c909 · outbound

This paper cites Flamingo: a visual language model for few-shot learn- ing.

A Survey on Vision-Language-Action Models for Autonomous Driving Flamingo: a visual language model for few-shot learn- ing

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.317875Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.317875Z digest=sha256:1074d73c91fe27e5c1e40795b582d8c7d9d8bdc3ab0f1f7c96b960da15ec0a5c

Observation 9683c8e4-9fa6-4ba5-b4de-b343ce145803 · outbound

This paper cites An lstm net- work for highway trajectory prediction.

A Survey on Vision-Language-Action Models for Autonomous Driving An lstm net- work for highway trajectory prediction

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.321970Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.321970Z digest=sha256:f87d77b63aa74d355a9086d6d6b0a2f3f62f0f261e71d54b89ea059190659bf3

Observation 1d2f59fb-38cf-4392-b6ad-e788b6dbb978 · outbound

This paper cites VaViM and VaVAM: Autonomous Driving through Video Generative Modeling.

A Survey on Vision-Language-Action Models for Autonomous Driving VaViM and VaVAM: Autonomous Driving through Video Generative Modeling

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.325481Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.325481Z digest=sha256:be44a5072c7d0b3f694a67be6b2290094b8fe2f0c3c6dd0d7ca38b4b9668777e

Observation 3be9daf3-30aa-47f2-8f53-03ea0dfa3524 · outbound

This paper cites $\pi_0$: A Vision-Language-Action Flow Model for General Robot Control.

A Survey on Vision-Language-Action Models for Autonomous Driving $\pi_0$: A Vision-Language-Action Flow Model for General Robot Control

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.329364Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.329364Z digest=sha256:163ae5517fa8b412af1d2f8824150a0edb5f4ccf1a812d1059349b73cdac0c5e

Observation a940b45e-a0dc-4940-be05-64a3d59804c0 · outbound

This paper cites Fine-grained affective processing capabilities emerging from large lan- guage models.

A Survey on Vision-Language-Action Models for Autonomous Driving Fine-grained affective processing capabilities emerging from large lan- guage models

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.333208Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.333208Z digest=sha256:a145008f27959027833131528960785182e534cb8502df1c1d24fdef278464d9

Observation 0c255595-3398-4239-bed9-78afb5ad81ef · outbound

This paper cites Language models are few-shot learners.

A Survey on Vision-Language-Action Models for Autonomous Driving Language models are few-shot learners

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.336827Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.336827Z digest=sha256:0564acf51a207670700ae6fc1a4c15f4691a685e95e9e04316a637bc1ba81659

Observation eddffb62-ec8c-4c99-9782-f80d4cdc0146 · outbound

This paper cites nuscenes: A multimodal dataset for autonomous driving.

A Survey on Vision-Language-Action Models for Autonomous Driving nuscenes: A multimodal dataset for autonomous driving

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.341206Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.341206Z digest=sha256:0ea0bf3d5d81f9d75d50a31c87467cfd713192254e1cb430608fb01b2c719255

Observation 7fa2c5dd-9bfd-4e66-9c0b-c682eeeca2fa · outbound

This paper cites nuscenes: A mul- timodal dataset for autonomous driving.

A Survey on Vision-Language-Action Models for Autonomous Driving nuscenes: A mul- timodal dataset for autonomous driving

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.344745Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.344745Z digest=sha256:4e252c2d7c2288eb14d4345c88bec7c086711ba7781d00090cbdd36e13080a2f

Observation 87ae5ca0-c84f-4946-9287-e7bf4952b552 · outbound

This paper cites Learning from all vehi- cles.

A Survey on Vision-Language-Action Models for Autonomous Driving Learning from all vehi- cles

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.349406Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.349406Z digest=sha256:14775fa962259a48a6cd3b11b37d2f78db8d481e783859d3e65117935a8a7a84

Observation 15618d1e-cae1-4536-a0a4-4ad30489a0fb · outbound

This paper cites Insight: Enhancing autonomous driving safety through vision-language models on context-aware hazard detection and edge case evaluation.

A Survey on Vision-Language-Action Models for Autonomous Driving Insight: Enhancing autonomous driving safety through vision-language models on context-aware hazard detection and edge case evaluation

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.352905Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.352905Z digest=sha256:753c3f0aef0699727bfeaa78fa0a915d6261ded8411033aafb21e10659502607

Observation 88023e48-d173-4a4c-a9d3-d4966c1b92f8 · outbound

This paper cites TS-VLM: Text-Guided SoftSort Pooling for Vision-Language Models in Multi-View Driving Reasoning.

A Survey on Vision-Language-Action Models for Autonomous Driving TS-VLM: Text-Guided SoftSort Pooling for Vision-Language Models in Multi-View Driving Reasoning

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.357239Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.357239Z digest=sha256:05a5e08840cfb02b9d568403be29f604180f0061e36428abb8e2abae2db87160

Observation 4e41fc49-42a9-42da-9ee8-3ce930c479e6 · outbound

This paper cites What data do we need for training an av motion planner? In 2021 IEEE Inter- national Conference on Robotics and Automation (ICRA) , pages 1066–1072.

A Survey on Vision-Language-Action Models for Autonomous Driving What data do we need for training an av motion planner? In 2021 IEEE Inter- national Conference on Robotics and Automation (ICRA) , pages 1066–1072

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.361431Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.361431Z digest=sha256:9b8224b72c6275ff0950a7dfaeb74be8ce7a2a27376b53351308293df5e95802

Observation b5cd6e6b-7abf-4712-a9a2-783a5470fee6 · outbound

This paper cites End-to-end autonomous driving: Challenges and frontiers.

A Survey on Vision-Language-Action Models for Autonomous Driving End-to-end autonomous driving: Challenges and frontiers

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.364887Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.364887Z digest=sha256:f2362a7f1bc22d7a4216dd885bcad6632e9017d9c85c088e6391552006722f1a

Observation f48e424e-dd98-4372-b145-c8ce827bcbc7 · outbound

This paper cites VADv2: End-to-End Vectorized Autonomous Driving via Probabilistic Planning.

A Survey on Vision-Language-Action Models for Autonomous Driving VADv2: End-to-End Vectorized Autonomous Driving via Probabilistic Planning

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.369131Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.369131Z digest=sha256:4a50e0b518a2efd14bf590524977fe8cad7362360c3b7b9edfe64b70387fe567

Observation 1d70c450-28a6-4dcc-8af5-14c10c1389aa · outbound

This paper cites Asynchronous large language model en- hanced planner for autonomous driving.

A Survey on Vision-Language-Action Models for Autonomous Driving Asynchronous large language model en- hanced planner for autonomous driving

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.373472Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.373472Z digest=sha256:7dc83327cf06d9f6db616c9483c816905d6c268f6385033e9c726e383d50c3b2

Observation 206166b7-1c33-4c8c-971f-57f8bf472cfb · outbound

This paper cites Ppad: Iterative interactions of prediction and planning for end-to-end autonomous driving.

A Survey on Vision-Language-Action Models for Autonomous Driving Ppad: Iterative interactions of prediction and planning for end-to-end autonomous driving

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.377211Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.377211Z digest=sha256:d244063eecd46aff2b15c39c3d584e06eb447ada5a2d25f6c58cae68377b4015

Observation 981fc864-c778-4305-a2b6-26039998b808 · outbound

This paper cites CoVLA: Comprehensive vision-language-action dataset for au- tonomous driving.

A Survey on Vision-Language-Action Models for Autonomous Driving CoVLA: Comprehensive vision-language-action dataset for au- tonomous driving

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.381473Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.381473Z digest=sha256:8ae69174f05d004d63d888f28cd49a246cccdb1f29d59a0c0a2b5c6b826ee32a

Observation ae9035c8-cebe-40f7-8752-3275413763fa · outbound

This paper cites Impromptu VLA: Open Weights and Open Data for Driving Vision-Language-Action Models.

A Survey on Vision-Language-Action Models for Autonomous Driving Impromptu VLA: Open Weights and Open Data for Driving Vision-Language-Action Models

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.384791Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.384791Z digest=sha256:95afa723c8d588cb777481579da380dbffe85b3ecb846604d4d922bc17525ddd

Observation f0a55e4d-3b65-4899-9f39-c0b41d12bc6b · outbound

This paper cites Neat: Neural attention fields for end-to-end autonomous driving.

A Survey on Vision-Language-Action Models for Autonomous Driving Neat: Neural attention fields for end-to-end autonomous driving

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.388971Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.388971Z digest=sha256:2f42652a819e3f3af5166e41e5b1c66449ac6885ad8a4cda64780176ed41779d

Observation 16c070ea-5acb-463f-9fea-4626f7e83d4f · outbound

This paper cites Transfuser: Imita- tion with transformer-based sensor fusion for autonomous driving.

A Survey on Vision-Language-Action Models for Autonomous Driving Transfuser: Imita- tion with transformer-based sensor fusion for autonomous driving

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.392370Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.392370Z digest=sha256:2008f2d068f52a3fbbfa467670692676b931a16f80ccb21b4200fa098c789a3a

Observation 230a8c27-d8da-41f2-863a-2fe5460e25b3 · outbound

This paper cites Talk2bev: Language-enhanced bird’s- eye view maps for autonomous driving.

A Survey on Vision-Language-Action Models for Autonomous Driving Talk2bev: Language-enhanced bird’s- eye view maps for autonomous driving

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.395495Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.395495Z digest=sha256:327bfab80255e4030fd2393be7a071ee3809551be8fbf8147528310e856c7836

Observation d5233365-64c2-4c2b-9811-1693fd109563 · outbound

This paper cites A survey on multimodal large lan- guage models for autonomous driving.

A Survey on Vision-Language-Action Models for Autonomous Driving A survey on multimodal large lan- guage models for autonomous driving

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.398942Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.398942Z digest=sha256:05b42bae6ca747c38a81f1e9d4dcd164ce2e351ef84d6b37b4a9d1f3e4212b40

Observation bb544cc9-b712-435f-b5f1-ee1b4e55acd0 · outbound

This paper cites Chain-of-Thought for Autonomous Driving: A Comprehensive Survey and Future Prospects.

A Survey on Vision-Language-Action Models for Autonomous Driving Chain-of-Thought for Autonomous Driving: A Comprehensive Survey and Future Prospects

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.402197Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.402197Z digest=sha256:ddca4122408fb6e8e15ab792087c00a75d1ab9f1cad6d44f5c676ea7302680a1

Observation 97a7f9ab-fabd-4d1e-ba7a-1558608947d7 · outbound

This paper cites Hint-AD: Holistically Aligned Interpretability in End-to-End Autonomous Driving.

A Survey on Vision-Language-Action Models for Autonomous Driving Hint-AD: Holistically Aligned Interpretability in End-to-End Autonomous Driving

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.405773Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.405773Z digest=sha256:ca193789c0721cf6a8b76cc8fe02053d23b17e25a02026481f99bd3257440ae1

Observation 400d1b16-2441-4279-a5c5-cb36963b63cf · outbound

This paper cites Dualad: Disentangling the dynamic and static world for end-to-end driving.

A Survey on Vision-Language-Action Models for Autonomous Driving Dualad: Disentangling the dynamic and static world for end-to-end driving

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.409296Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.409296Z digest=sha256:a3e41d74f6b662d9481f8600aa9a9fe4f8c91303edb2fb6ac35b5d40119f4d28

Observation 043f54c0-3ae0-49c0-b9c5-106b8097b215 · outbound

This paper cites Carla: An open urban driv- ing simulator.

A Survey on Vision-Language-Action Models for Autonomous Driving Carla: An open urban driv- ing simulator

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.412510Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.412510Z digest=sha256:0e2750f7b4aa65f4a503e274ca1352e0e50c0284c0b62ba76c6bfe383948773f

Observation 18721357-3ffb-480b-a858-92e7014b9fc1 · outbound

This paper cites On the road to portability: Compressing end-to-end motion planner for autonomous driving.

A Survey on Vision-Language-Action Models for Autonomous Driving On the road to portability: Compressing end-to-end motion planner for autonomous driving

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.415798Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.415798Z digest=sha256:91830e76dfcc94b5143d28814b87dc758a2bbbd7569eda0d9f5845522113520a

Observation bf68921b-35b3-4fb8-984d-61e1486323c8 · outbound

This paper cites Polarpoint-bev: Bird-eye- view perception in polar points for explainable end-to-end autonomous driving.

A Survey on Vision-Language-Action Models for Autonomous Driving Polarpoint-bev: Bird-eye- view perception in polar points for explainable end-to-end autonomous driving

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.418970Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.418970Z digest=sha256:b029367d148df540110873e4baa3defc2eeac7bc161f7caf59e3a8dde9c7fdff

Observation c5da9684-8de4-4242-87fb-5b4afdb9c75e · outbound

This paper cites Drive like a human: Rethinking autonomous driving with large language models.

A Survey on Vision-Language-Action Models for Autonomous Driving Drive like a human: Rethinking autonomous driving with large language models

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.422092Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.422092Z digest=sha256:65d697c844ebdf85d1a927d0ee07a92da26a64746caa5a9f369c8f4a62a79a04

Observation 25067f8d-5942-446f-b9ee-06a64cc0550d · outbound

This paper cites ORION: A Holistic End-to-End Autonomous Driving Framework by Vision-Language Instructed Action Generation.

A Survey on Vision-Language-Action Models for Autonomous Driving ORION: A Holistic End-to-End Autonomous Driving Framework by Vision-Language Instructed Action Generation

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.429364Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.429364Z digest=sha256:2025a9d97df75c7a022f268e0c8e50db96ad8687802586a2f9bac39dcf7486fc

Observation ff5b3305-7a88-4afd-9ab0-2fc46ae66821 · outbound

This paper cites A Survey for Foundation Models in Autonomous Driving.

A Survey on Vision-Language-Action Models for Autonomous Driving A Survey for Foundation Models in Autonomous Driving

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.432603Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.432603Z digest=sha256:461192a0c381ecf0e3963281b14277b384ad0b29c78ad20022b6f118cdf09150

Observation e064c0f6-803e-497d-814f-5ec440ab0104 · outbound

This paper cites LangCoop: Collaborative Driving with Language.

A Survey on Vision-Language-Action Models for Autonomous Driving LangCoop: Collaborative Driving with Language

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.436251Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.436251Z digest=sha256:bd1da02e8bf9c5a6e5436ffff18ce9e5dce5433e0a7b2e8957d86743823562c4

Observation a06dc6e1-527a-4ff7-84eb-a8eb02c12f7c · outbound

This paper cites A review of motion planning techniques for au- tomated vehicles.

A Survey on Vision-Language-Action Models for Autonomous Driving A review of motion planning techniques for au- tomated vehicles

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.439666Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.439666Z digest=sha256:751dea246c3d336dbb793704c0aead49b53034719952b375835f43ad7fb7d083

Observation eb80006e-0204-4af8-9719-477ef2246b51 · outbound

This paper cites iPad: Iterative Proposal-centric End-to-End Autonomous Driving.

A Survey on Vision-Language-Action Models for Autonomous Driving iPad: Iterative Proposal-centric End-to-End Autonomous Driving

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.442877Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.442877Z digest=sha256:0dc0f1689b79c3f3fcf7844a1aa86b7744548304bb51797a6cdcb003e7ca447d

Observation f6b14606-f44c-4317-aff4-815a937b3744 · outbound

This paper cites End-to-End Autonomous Driving without Costly Modularization and 3D Manual Annotation.

A Survey on Vision-Language-Action Models for Autonomous Driving End-to-End Autonomous Driving without Costly Modularization and 3D Manual Annotation

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.446259Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.446259Z digest=sha256:89ec4b862596400b4b069c6316faa79cff644e57b80138d7f4c9a150e2ccc88a

Observation cbe35daa-af34-47e7-9f69-a5ff07a714ed · outbound

This paper cites SURDS: Benchmarking Spatial Understanding and Reasoning in Driving Scenarios with Vision Language Models.

A Survey on Vision-Language-Action Models for Autonomous Driving SURDS: Benchmarking Spatial Understanding and Reasoning in Driving Scenarios with Vision Language Models

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.449729Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.449729Z digest=sha256:ae527c9e5fa0677869b396a4b140d8d7e8499fb55523c240e5400a41272347ba

Observation c8a4693a-d8ac-41fc-86b0-14116f3e8d94 · outbound

This paper cites Dme-driver: Integrating human decision logic and 3d scene perception in autonomous driving.

A Survey on Vision-Language-Action Models for Autonomous Driving Dme-driver: Integrating human decision logic and 3d scene perception in autonomous driving

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.453118Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.453118Z digest=sha256:17bf035e5c8736331a9dfcc90132cf359c279c2a35339d586cbeef2fb2096278

Observation ed0d1377-4c9b-413f-b84e-4ff74c4aeca0 · outbound

This paper cites Driveaction: A benchmark for exploring human-like driving decisions in vla models.

A Survey on Vision-Language-Action Models for Autonomous Driving Driveaction: A benchmark for exploring human-like driving decisions in vla models

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.456365Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.456365Z digest=sha256:f9ed950eda91050d78352e2cd9463358dc3a461f9c6c7b0070cefc14092130a2

Observation 625c298e-8009-4e1a-815c-aeb0cfe79849 · outbound

This paper cites Urban driving with conditional imitation learning.

A Survey on Vision-Language-Action Models for Autonomous Driving Urban driving with conditional imitation learning

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.459454Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.459454Z digest=sha256:b83eb94150a500e2a07b3f61ad942486ba144eeb5861127449a2c817aee046a4

Observation 49fdd285-272c-4511-a055-efdfe7f3aef7 · outbound

This paper cites DriveAgent: Multi-Agent Structured Reasoning with LLM and Multimodal Sensor Fusion for Autonomous Driving.

A Survey on Vision-Language-Action Models for Autonomous Driving DriveAgent: Multi-Agent Structured Reasoning with LLM and Multimodal Sensor Fusion for Autonomous Driving

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.462537Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.462537Z digest=sha256:1e06cb4876d3386a6be338baef385af7a5081ea6ab0c54a228c14a28e1746587

Observation 7222be9c-5737-4c26-b8fb-be9cc01874f8 · outbound

This paper cites Lora: Low-rank adaptation of large language models.

A Survey on Vision-Language-Action Models for Autonomous Driving Lora: Low-rank adaptation of large language models

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.465990Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.465990Z digest=sha256:842b9f33d67f3ce772172109fa9478a40e009c572832cf278c5d963f29536ff5

Observation 3b76c5c3-ee14-47a7-a98d-07ad99bca1d9 · outbound

This paper cites Planning-oriented autonomous driving.

A Survey on Vision-Language-Action Models for Autonomous Driving Planning-oriented autonomous driving

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.469144Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.469144Z digest=sha256:822b2437f0d71a8fb89da5046151f0b8ac1c3e02d73819f45366f143d31b167f

Observation 25810685-f58b-4f16-a052-e0c8c25afd6c · outbound

This paper cites Planning-oriented autonomous driving.

A Survey on Vision-Language-Action Models for Autonomous Driving Planning-oriented autonomous driving

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.472273Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.472273Z digest=sha256:dd241fa640a346d6658d412ab98aac641e561651c9dbabad76114134d4ba1258

Observation 3f37d0c4-e48d-4bbc-9443-1f08c6123840 · outbound

This paper cites A survey on trajectory-prediction methods for autonomous driving.

A Survey on Vision-Language-Action Models for Autonomous Driving A survey on trajectory-prediction methods for autonomous driving

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.475355Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.475355Z digest=sha256:18b458503be706d8f5969d211a421483bd25434c47e5a03eb321b074bb27a269

Observation a181b55c-bd64-43f7-8c89-c9e8f06f9781 · outbound

This paper cites RoboTron-Drive: All-in-One Large Multimodal Model for Autonomous Driving.

A Survey on Vision-Language-Action Models for Autonomous Driving RoboTron-Drive: All-in-One Large Multimodal Model for Autonomous Driving

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.478639Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.478639Z digest=sha256:1bf9b63aa8b74ef4eeff8359a90d4a0443bc9a6888d9ff679b6674a927ddfdfb

Observation c46c8d11-f331-4e38-8b22-09967ed31c8b · outbound

This paper cites VLM-RL: A Unified Vision Language Models and Reinforcement Learning Framework for Safe Autonomous Driving.

A Survey on Vision-Language-Action Models for Autonomous Driving VLM-RL: A Unified Vision Language Models and Reinforcement Learning Framework for Safe Autonomous Driving

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.482447Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.482447Z digest=sha256:f322e713db4fbb416c8d808fb0ef1d651f0cd77566afe2c64a92c91718d7e984

Observation 0bc08dcf-5eaf-4ec0-9951-e0c68652fccf · outbound

This paper cites NORA: A Small Open-Sourced Generalist Vision Language Action Model for Embodied Tasks.

A Survey on Vision-Language-Action Models for Autonomous Driving NORA: A Small Open-Sourced Generalist Vision Language Action Model for Embodied Tasks

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.485889Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.485889Z digest=sha256:e40c43c3909abddae7c89d9223509a31daf48bc1bb5f0540ca706c72cd978611

Observation 569b3479-5401-4fbc-8485-2bc3187b5b8a · outbound

This paper cites Yolo-v1 to yolo-v8, the rise of yolo and its complementary nature toward digital manufacturing and industrial defect detection.

A Survey on Vision-Language-Action Models for Autonomous Driving Yolo-v1 to yolo-v8, the rise of yolo and its complementary nature toward digital manufacturing and industrial defect detection

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.489388Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.489388Z digest=sha256:291cc5eccd01d633950aaecde4ac29734c8921f6a62898bb1d7102b94fdfebcb

Observation 5224ef4c-f19f-4cbf-b687-c2984152e145 · outbound

This paper cites EMMA: End-to-End Multimodal Model for Autonomous Driving.

A Survey on Vision-Language-Action Models for Autonomous Driving EMMA: End-to-End Multimodal Model for Autonomous Driving

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.492816Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.492816Z digest=sha256:361379614d49ed1a1e6bad65383221adad5d7dd0ce258acd9ac9944460494a06

Observation 559edfd9-9390-47cf-841f-54cc8482836a · outbound

This paper cites DriveLMM-o1: A Step-by-Step Reasoning Dataset and Large Multimodal Model for Driving Scenario Understanding.

A Survey on Vision-Language-Action Models for Autonomous Driving DriveLMM-o1: A Step-by-Step Reasoning Dataset and Large Multimodal Model for Driving Scenario Understanding

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.496591Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.496591Z digest=sha256:61484e135505f797f8553df1d5469831dc7a0661ae83a25c24d19cf3260a3491

Observation aacfba20-a326-47dd-945b-f37dc9f8ee14 · outbound

This paper cites Narrate: Versatile lan- guage architecture for optimal control in robotics.

A Survey on Vision-Language-Action Models for Autonomous Driving Narrate: Versatile lan- guage architecture for optimal control in robotics

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.499865Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.499865Z digest=sha256:36f971abb9377e399dbb3857d8070233eb41a7709365e78dd1d3e52d683de20a

Observation d92e1b2c-cbdb-4b14-9bc7-14bf466ee07e · outbound

This paper cites ADriver-I: A General World Model for Autonomous Driving.

A Survey on Vision-Language-Action Models for Autonomous Driving ADriver-I: A General World Model for Autonomous Driving

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.503085Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.503085Z digest=sha256:f1b606a6441ca68f91fe01678ecc7326cb2bc03760184f8b6a85a4dad2b1254c

Observation 334a87f1-6670-47ce-9a9c-084f0356cfd2 · outbound

This paper cites Think twice be- fore driving: Towards scalable decoders for end-to-end au- tonomous driving.

A Survey on Vision-Language-Action Models for Autonomous Driving Think twice be- fore driving: Towards scalable decoders for end-to-end au- tonomous driving

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.506379Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.506379Z digest=sha256:fb5405e1cad5de77ab75c802f54b3164df102c992bf104aa441df71ab66d7809

Observation 38578a71-14e3-43f1-aca0-8dd29613d89c · outbound

This paper cites Bench2Drive: Towards Multi-Ability Benchmarking of Closed-Loop End-To-End Autonomous Driving.

A Survey on Vision-Language-Action Models for Autonomous Driving Bench2Drive: Towards Multi-Ability Benchmarking of Closed-Loop End-To-End Autonomous Driving

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.509692Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.509692Z digest=sha256:f3702a7f50f25cb1151488cc13c945938c9598872e431a0503025c901e7bfa7c

Observation 41eabb91-c06c-40d7-9c42-900df6541d45 · outbound

This paper cites DriveTransformer: Unified Transformer for Scalable End-to-End Autonomous Driving.

A Survey on Vision-Language-Action Models for Autonomous Driving DriveTransformer: Unified Transformer for Scalable End-to-End Autonomous Driving

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.513157Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.513157Z digest=sha256:7de7f677077d52c0fd4da6dcb8c17f8272f21944e5ce2f6b7ea26a399759b2c7

Observation 224a316f-d4fa-4a7b-a4cd-896d8c42f4ef · outbound

This paper cites DiffVLA: Vision-Language Guided Diffusion Planning for Autonomous Driving.

A Survey on Vision-Language-Action Models for Autonomous Driving DiffVLA: Vision-Language Guided Diffusion Planning for Autonomous Driving

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.516550Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.516550Z digest=sha256:797f6fd509d430b6f5bea0940be51a6e2721360035cf797f2a7f400b6479e9f6

Observation e42c38b3-5667-4a3b-9987-316e834e7550 · outbound

This paper cites Senna: Bridging Large Vision-Language Models and End-to-End Autonomous Driving.

A Survey on Vision-Language-Action Models for Autonomous Driving Senna: Bridging Large Vision-Language Models and End-to-End Autonomous Driving

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.520794Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.520794Z digest=sha256:e2409a2f3e1af2b35aee2119ff63efb2a4e5a558f332bca5ac68443702cf21af

Observation 5c35b144-26d3-4866-b0c6-8ca291c57bda · outbound

This paper cites Vad: Vectorized scene rep- resentation for efficient autonomous driving.

A Survey on Vision-Language-Action Models for Autonomous Driving Vad: Vectorized scene rep- resentation for efficient autonomous driving

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.524393Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.524393Z digest=sha256:96a27eb8ab9602a2d5d231460691837d572297279e08c6f93ce9a4d04976bc53

Observation 8269ba50-47e2-4843-b48a-d34b15368cda · outbound

This paper cites Koma: Knowledge-driven multi- agent framework for autonomous driving with large lan- guage models.

A Survey on Vision-Language-Action Models for Autonomous Driving Koma: Knowledge-driven multi- agent framework for autonomous driving with large lan- guage models

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.527846Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.527846Z digest=sha256:0ac23cceb063d18d9543a96c7c2d102032fc3707d4d7b2ad987cf0f547daad0d

Observation 848168b4-dbea-4e9c-8ddf-c10b85559472 · outbound

This paper cites Communication-Aware Reinforcement Learning for Cooperative Adaptive Cruise Control.

A Survey on Vision-Language-Action Models for Autonomous Driving Communication-Aware Reinforcement Learning for Cooperative Adaptive Cruise Control

Reference 61

Resolution
verified exact
local_arxiv, observed 2026-08-06T21:31:06.177631Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T21:31:04.530858Z digest=sha256:88ffaef2ad0940dfdad3ccd1327b326a7a63bf1221e44fe5d7501effcfe7aab1

Observation faec5161-b732-46ed-bc48-d3785462ad2c · outbound

This paper cites Learning to drive in a day.

A Survey on Vision-Language-Action Models for Autonomous Driving Learning to drive in a day

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.534174Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.534174Z digest=sha256:062b339adb65b8e6e7ebae26286376d3f705e8ab43e15611159176c9e5ac6346

Observation 653bce07-5833-4ada-90fe-b04a282e2f8f · outbound

This paper cites an unresolved cited work.

A Survey on Vision-Language-Action Models for Autonomous Driving Unresolved cited work

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.537689Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.537689Z digest=sha256:969289eaa443d3993bd4a61220c41bd3cd17757f0a6d137e245633e3031531d7

Observation 75ea97e9-bd5c-4ff7-b378-05062533862b · outbound

This paper cites Fine-Tuning Vision-Language-Action Models: Optimizing Speed and Success.

A Survey on Vision-Language-Action Models for Autonomous Driving Fine-Tuning Vision-Language-Action Models: Optimizing Speed and Success

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.540761Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.540761Z digest=sha256:81a0050d5d0c19e17572f1df3e170b2560e718c18c2eaaeacb53aa0e31cb394a

Observation 88d37ec8-5d2f-4ecf-8d92-6d7cb7aaa2d4 · outbound

This paper cites OpenVLA: An Open-Source Vision-Language-Action Model.

A Survey on Vision-Language-Action Models for Autonomous Driving OpenVLA: An Open-Source Vision-Language-Action Model

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.544824Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.544824Z digest=sha256:a888fdf89a7439893418f19b40370d9021dc85be9002d107f381c601770d3434

Observation 0261be28-95c3-4a3f-b0dc-7f76815360f1 · outbound

This paper cites A survey on motion prediction and risk assessment for in- telligent vehicles.

A Survey on Vision-Language-Action Models for Autonomous Driving A survey on motion prediction and risk assessment for in- telligent vehicles

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.549316Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.549316Z digest=sha256:6ce8f50a29b38859f54511badaf7377b8cb010ebf8a6eb217301b67e290f9a6a

Observation 4acb9048-e194-4083-bc7f-d33873d302b6 · outbound

This paper cites PointVLA: Injecting the 3D World into Vision-Language-Action Models.

A Survey on Vision-Language-Action Models for Autonomous Driving PointVLA: Injecting the 3D World into Vision-Language-Action Models

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.553726Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.553726Z digest=sha256:774bb8128440193a638f1e0d7b42c3d1f35abe5ddec0b72cfcfb612813d584d1

Observation d42ecad3-38ae-4afa-a415-a0a6ab560be7 · outbound

This paper cites Navigation-Guided Sparse Scene Representation for End-to-End Autonomous Driving.

A Survey on Vision-Language-Action Models for Autonomous Driving Navigation-Guided Sparse Scene Representation for End-to-End Autonomous Driving

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.557074Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.557074Z digest=sha256:f177812174ce96be7692c352a676bf7c05275b6d75a156cf45e1ff9f603f55e0

Observation 004a1db2-0aa4-40a9-94f9-8ba012b716fd · outbound

This paper cites Enhancing End-to-End Autonomous Driving with Latent World Model.

A Survey on Vision-Language-Action Models for Autonomous Driving Enhancing End-to-End Autonomous Driving with Latent World Model

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.560587Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.560587Z digest=sha256:f421fb8471d151f0aafd5ff8d0a76d4230c37e05cd19035bd71652217c8b5880

Observation e3f1a3fa-a0a9-4035-9bf7-2035b9e91235 · outbound

This paper cites Recogdrive: A reinforced cognitive framework for end-to-end autonomous driving, 2025.

A Survey on Vision-Language-Action Models for Autonomous Driving Recogdrive: A reinforced cognitive framework for end-to-end autonomous driving, 2025

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.564262Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.564262Z digest=sha256:9ad50956908daa37cb6c7e13f3a40114882b76796f0f50a3ff25938ab2ff6d53

Observation 01b84402-1ceb-47f6-bd22-f4984485b8b3 · outbound

This paper cites Hydra-MDP: End-to-end Multimodal Planning with Multi-target Hydra-Distillation.

A Survey on Vision-Language-Action Models for Autonomous Driving Hydra-MDP: End-to-end Multimodal Planning with Multi-target Hydra-Distillation

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.568371Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.568371Z digest=sha256:66c11ef14c97c33f9aaa24f4aa344edc30108e37060f30871fa59617a81ae4e5

Observation 58283645-65e3-4dc8-be82-d85c9dc06e21 · outbound

This paper cites Generalized Trajectory Scoring for End-to-end Multimodal Planning.

A Survey on Vision-Language-Action Models for Autonomous Driving Generalized Trajectory Scoring for End-to-end Multimodal Planning

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.572098Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.572098Z digest=sha256:717a79efd917502331e504035acd672cd2bf197185e12c31c8bb94ce24132076

Observation 6b37605a-5907-47e3-959b-3713cd1443eb · outbound

This paper cites Pnpnet: End-to-end per- ception and prediction with tracking in the loop.

A Survey on Vision-Language-Action Models for Autonomous Driving Pnpnet: End-to-end per- ception and prediction with tracking in the loop

Reference 73

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.576167Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.576167Z digest=sha256:12c044ccd0ba327ee2acfaaacf9205109f6b7f1e67235f111c456c4eed5609a0

Observation e0c04cd3-a3d8-41b6-880b-70f0e1e154a1 · outbound

This paper cites Visual instruction tuning.

A Survey on Vision-Language-Action Models for Autonomous Driving Visual instruction tuning

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.579305Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.579305Z digest=sha256:bd41b1ee91d3d9d44dabe6b5b600ce2dd6845a411d1effcaed472da7bb0dfeb6

Observation 44ea8d2f-3f23-4c2f-ae28-109d31e4e189 · outbound

This paper cites Mtd-gpt: A multi-task decision-making gpt model for autonomous driving at unsignalized intersections.

A Survey on Vision-Language-Action Models for Autonomous Driving Mtd-gpt: A multi-task decision-making gpt model for autonomous driving at unsignalized intersections

Reference 75

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.582341Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.582341Z digest=sha256:aab2803a006c672b395352dbcdf7d5563e041f75b37e4d506a27d86e2aabb10f

Observation 25d0286c-a370-4c1f-8d37-358293a1a5d5 · outbound

This paper cites Robomamba: Ef- ficient vision-language-action model for robotic reasoning and manipulation.

A Survey on Vision-Language-Action Models for Autonomous Driving Robomamba: Ef- ficient vision-language-action model for robotic reasoning and manipulation

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.585664Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.585664Z digest=sha256:ec51b37a7488c0617ac05f65e7f8c058938a92516bd68fb654b970d261207d7a

Observation 6b80e62b-0141-4228-8bbd-3fd7451e7954 · outbound

This paper cites Fully Unified Motion Planning for End-to-End Autonomous Driving.

A Survey on Vision-Language-Action Models for Autonomous Driving Fully Unified Motion Planning for End-to-End Autonomous Driving

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.588876Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.588876Z digest=sha256:cad086409b60e706dd1b4e59fba00eab10d380c364e37de8580cce031d48ee52

Observation a78c074f-2c92-41a6-8e06-99d5163e47a5 · outbound

This paper cites Vlm-e2e: Enhancing end-to-end autonomous driv- ing with multimodal driver attention fusion.

A Survey on Vision-Language-Action Models for Autonomous Driving Vlm-e2e: Enhancing end-to-end autonomous driv- ing with multimodal driver attention fusion

Reference 78

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.592420Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.592420Z digest=sha256:ba26fdc9fcc7a56995a877d3e2d075ce205d55be854b09a8f420ef4a12800458

Observation a9f834fe-66c9-44ec-82ad-36c0632a40ce · outbound

This paper cites Reasonplan: Unified scene prediction and decision reasoning for closed-loop au- tonomous driving.

A Survey on Vision-Language-Action Models for Autonomous Driving Reasonplan: Unified scene prediction and decision reasoning for closed-loop au- tonomous driving

Reference 79

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.595814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.595814Z digest=sha256:2d582abb24fd8464e48ffa675d041ccc408ed84b6519c1dfff60a10dafe78587

Observation 2dd1e3bc-b9ca-4c7b-8123-1e5b6f7222f6 · outbound

This paper cites Bevfusion: Multi-task multi-sensor fusion with unified bird’s-eye view representation.

A Survey on Vision-Language-Action Models for Autonomous Driving Bevfusion: Multi-task multi-sensor fusion with unified bird’s-eye view representation

Reference 80

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.599077Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.599077Z digest=sha256:b794bc8bef71feb978b5f7555bf291dcd8dc4d0ea99151c95cdae6361e995e61

Observation 30478a36-9dec-494f-9251-a392c74f2b68 · outbound

This paper cites VLM-MPC: Vision Language Foundation Model (VLM)-Guided Model Predictive Controller (MPC) for Autonomous Driving.

A Survey on Vision-Language-Action Models for Autonomous Driving VLM-MPC: Vision Language Foundation Model (VLM)-Guided Model Predictive Controller (MPC) for Autonomous Driving

Reference 81

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.602293Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.602293Z digest=sha256:094da1e5c291314afda0f9654dbc6048199d7cfc2f13e89cd4dff2dbaaf4f90e

Observation f55be80c-fd31-45a6-a185-719224fa5fd1 · outbound

This paper cites ActiveAD: Planning-Oriented Active Learning for End-to-End Autonomous Driving.

A Survey on Vision-Language-Action Models for Autonomous Driving ActiveAD: Planning-Oriented Active Learning for End-to-End Autonomous Driving

Reference 82

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.605560Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.605560Z digest=sha256:178b32808cc9b32336f16677274780083cd000951b2356c333f5b7ff5520924e

Observation 6b7a9aad-23d7-4b7d-89c1-d3bf8446c354 · outbound

This paper cites Fast and fu- rious: Real time end-to-end 3d detection, tracking and mo- tion forecasting with a single convolutional net.

A Survey on Vision-Language-Action Models for Autonomous Driving Fast and fu- rious: Real time end-to-end 3d detection, tracking and mo- tion forecasting with a single convolutional net

Reference 83

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.608939Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.608939Z digest=sha256:bcb66d493a970219f96402e8391c8bff4ac77ac997f36abad241248a2abcf1fa

Observation 8342d199-fd35-4553-81a6-23c516a5613a · outbound

This paper cites Dolphins: Multimodal language model for driving.

A Survey on Vision-Language-Action Models for Autonomous Driving Dolphins: Multimodal language model for driving

Reference 84

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.612085Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.612085Z digest=sha256:0e820f084b1b736809438acac0f873a1c9500d4e3221c0a53a73ce2fe5121242

Observation 89742c4a-fee3-4efe-bba9-f3dac8bf713a · outbound

This paper cites A Survey on Vision-Language-Action Models for Embodied AI.

A Survey on Vision-Language-Action Models for Autonomous Driving A Survey on Vision-Language-Action Models for Embodied AI

Reference 85

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.615132Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.615132Z digest=sha256:2e2f1412330123de1c62be718187d43415529435aedbf947b14870afb1373c75

Observation 6e81063f-e9f2-41e9-b3cf-77ed23412124 · outbound

This paper cites LeapVAD: A Leap in Autonomous Driving via Cognitive Perception and Dual-Process Thinking.

A Survey on Vision-Language-Action Models for Autonomous Driving LeapVAD: A Leap in Autonomous Driving via Cognitive Perception and Dual-Process Thinking

Reference 86

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.618427Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.618427Z digest=sha256:75517eca2bfd48198a779145d9fc602eb596f16014b863eaeff86a968447c8b6

Observation cda65a11-b737-4623-b2a2-fbc9896fcb20 · outbound

This paper cites GPT-Driver: Learning to Drive with GPT.

A Survey on Vision-Language-Action Models for Autonomous Driving GPT-Driver: Learning to Drive with GPT

Reference 87

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.622875Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.622875Z digest=sha256:d9ab6652a4a8bb697ef444119ad010b5bf311cc0ddca15357fb38c4591ad31f9

Observation d02d2829-189c-4182-a3c9-f7aa54bf1d9c · outbound

This paper cites A Language Agent for Autonomous Driving.

A Survey on Vision-Language-Action Models for Autonomous Driving A Language Agent for Autonomous Driving

Reference 88

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.626480Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.626480Z digest=sha256:033d22c14fb5ee2dd5d36bf0aac78eb0b2a7cc19045d8372da484c5492e174eb

Observation 83a6066d-1707-4fa5-9bbb-420260d6686b · outbound

This paper cites Lingoqa: Visual question answering for autonomous driving.

A Survey on Vision-Language-Action Models for Autonomous Driving Lingoqa: Visual question answering for autonomous driving

Reference 89

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.629874Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.629874Z digest=sha256:b970627d22828e50ab6125cf831bf7c0c264d360b383a31612ede9e191786b1d

Observation 1ec33b50-ede3-45f7-a0bb-db23a2fe0d2f · outbound

This paper cites Continuously Learning, Adapting, and Improving: A Dual-Process Approach to Autonomous Driving.

A Survey on Vision-Language-Action Models for Autonomous Driving Continuously Learning, Adapting, and Improving: A Dual-Process Approach to Autonomous Driving

Reference 90

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.633849Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.633849Z digest=sha256:e648bda0930d5f85c7c9a0333814311c01859199a1dbd3e6281019417423dd3e

Observation e7f686fe-f594-4a44-a100-80d971a5004e · outbound

This paper cites Chatmpc: Natural language based mpc personalization.

A Survey on Vision-Language-Action Models for Autonomous Driving Chatmpc: Natural language based mpc personalization

Reference 91

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.638377Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.638377Z digest=sha256:90795861abbc97dfd191a25b83d1a1dceac1d7e197ea8919873947aa53409606

Observation 4ef1b1d9-cf11-4a75-bdc3-3bef5054ae1a · outbound

This paper cites Data Scaling Laws for End-to-End Autonomous Driving.

A Survey on Vision-Language-Action Models for Autonomous Driving Data Scaling Laws for End-to-End Autonomous Driving

Reference 92

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.641543Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.641543Z digest=sha256:01f62ec1b0ca6447f025aa3fc0de17fcd2d0026888799134a5b9ddd5eb2dcf29

Observation aab49f27-f584-4724-a0ef-1071af9a8d3e · outbound

This paper cites Rea- son2drive: Towards interpretable and chain-based reason- ing for autonomous driving.

A Survey on Vision-Language-Action Models for Autonomous Driving Rea- son2drive: Towards interpretable and chain-based reason- ing for autonomous driving

Reference 93

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.645337Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.645337Z digest=sha256:53cbcf182f45be3272ec565c0d17c1705fc57aa623618df9f880ab7b4d8af3f6

Observation 50a6587c-3ce2-410d-b1f3-67784e3c7a6a · outbound

This paper cites DINOv2: Learning Robust Visual Features without Supervision.

A Survey on Vision-Language-Action Models for Autonomous Driving DINOv2: Learning Robust Visual Features without Supervision

Reference 94

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.648737Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.648737Z digest=sha256:a8aeee74b7d1362af9b4be6a8f5a1b2b3223829cf3326d49cfa82a3823866e7f

Observation 2bd0575f-59dd-4132-9ead-ab10afbd4ff6 · outbound

This paper cites A survey of motion planning and control techniques for self-driving urban vehicles.

A Survey on Vision-Language-Action Models for Autonomous Driving A survey of motion planning and control techniques for self-driving urban vehicles

Reference 95

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.652503Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.652503Z digest=sha256:d22f8938a6f35e69b86513485cf7aa034c4b48059fbabdc8bdf208dc1cca97af

Observation ae903616-2e2d-4ab7-a36c-49d69a2c4e2e · outbound

This paper cites Vlp: Vision language planning for autonomous driving.

A Survey on Vision-Language-Action Models for Autonomous Driving Vlp: Vision language planning for autonomous driving

Reference 96

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.656685Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.656685Z digest=sha256:9506c45696be06c972baae74114fdd68f731461b150b5136b1d4af080dcfae4c

Observation d4f55e8d-20a7-4b11-897b-961c39ea1760 · outbound

This paper cites Lego-drive: Language-enhanced goal-oriented closed-loop end-to-end autonomous driving.

A Survey on Vision-Language-Action Models for Autonomous Driving Lego-drive: Language-enhanced goal-oriented closed-loop end-to-end autonomous driving

Reference 97

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.659895Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.659895Z digest=sha256:cac5519e47c199a9f5d6318d17cddebeb2a5859774249a374331741809e1960f

Observation 29f45af6-fa90-40c6-93e8-d9eddb9426a5 · outbound

This paper cites FAST: Efficient Action Tokenization for Vision-Language-Action Models.

A Survey on Vision-Language-Action Models for Autonomous Driving FAST: Efficient Action Tokenization for Vision-Language-Action Models

Reference 98

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.663325Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.663325Z digest=sha256:448b43503648ea9041f7918a5b1d87eac053f7fb229e30f92e1b9d99f0445a6c

Observation c2c0b6f3-5add-43e1-8983-a090cf69e547 · outbound

This paper cites Agentthink: A unified framework for tool-augmented chain-of-thought reasoning in vision- language models for autonomous driving.

A Survey on Vision-Language-Action Models for Autonomous Driving Agentthink: A unified framework for tool-augmented chain-of-thought reasoning in vision- language models for autonomous driving

Reference 99

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.667282Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.667282Z digest=sha256:8807e57fe4dbaae8182c9f7cbfb8b00eb47da2e33559bac5d1e90e7648e69ce9

Observation bd8d2423-3ca1-44ce-bdaa-d1c46b3f86ec · outbound

This paper cites FASIONAD++ : Integrating High-Level Instruction and Information Bottleneck in FAt-Slow fusION Systems for Enhanced Safety in Autonomous Driving with Adaptive Feedback.

A Survey on Vision-Language-Action Models for Autonomous Driving FASIONAD++ : Integrating High-Level Instruction and Information Bottleneck in FAt-Slow fusION Systems for Enhanced Safety in Autonomous Driving with Adaptive Feedback

Reference 100

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.670504Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.670504Z digest=sha256:fe51bc27aefbf3add1a2184479442e9b96f372605df70441d7a8e6696a91993e

Observation a020dcbf-7044-407c-8e17-0afe8753c195 · outbound

This paper cites SpatialVLA: Exploring Spatial Representations for Visual-Language-Action Model.

A Survey on Vision-Language-Action Models for Autonomous Driving SpatialVLA: Exploring Spatial Representations for Visual-Language-Action Model

Reference 101

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.674245Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.674245Z digest=sha256:a44eba9bbdee77c7c7b137cf1f1215534bdc3784f0590d4523b300dd45a3ed4b

Pith citing papers

Observation 6114a5cf-459a-4563-b467-772dcb481c99 · inbound

Research Challenges and Progress in the End-to-End V2X Cooperative Autonomous Driving Competition cites this paper.

Research Challenges and Progress in the End-to-End V2X Cooperative Autonomous Driving Competition A Survey on Vision-Language-Action Models for Autonomous Driving

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-06T12:39:02.315303Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:39:02.315303Z digest=sha256:c9680c626ba97763fa8bfcc22f38283211cdd4312f98db239d481bc26b562917

Observation 0538d1b4-ec46-4733-ba7d-7fb55c70bcfc · inbound

Efficient Vision-Language-Action Models for Embodied Manipulation: A Systematic Survey cites this paper.

Efficient Vision-Language-Action Models for Embodied Manipulation: A Systematic Survey A Survey on Vision-Language-Action Models for Autonomous Driving

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-04T09:08:07.066489Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:08:07.066489Z digest=sha256:677f6b3c4b42885ea54b4d88029af7c902ba9ea55b48b43c42978cff29a3fa9c

Observation 92d6ba2c-4203-42ed-9441-008dca2c40cc · inbound

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving cites this paper.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving A Survey on Vision-Language-Action Models for Autonomous Driving

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.416823Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.416823Z digest=sha256:b2b4637e171a1e896346465d127c9ca870db80581d5cb748f21a200e12c29cd4

Observation 3ac9cf36-3282-43d6-9301-f54beba24238 · inbound

VLSA: Vision-Language-Action Models with Plug-and-Play Safety Constraint Layer cites this paper.

VLSA: Vision-Language-Action Models with Plug-and-Play Safety Constraint Layer A Survey on Vision-Language-Action Models for Autonomous Driving

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-03T17:38:20.749342Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T17:38:20.749342Z digest=sha256:9adb6914a640af902c3a2e66c8adb2158633d4ed748fdeb864d5fe2b23f0213f

Observation ac0e4079-2379-48f1-aae3-20bfad6b6af6 · inbound

A Review of Learning-Based Motion Planning: Toward a Data-Driven Optimal Control Approach cites this paper.

A Review of Learning-Based Motion Planning: Toward a Data-Driven Optimal Control Approach A Survey on Vision-Language-Action Models for Autonomous Driving

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-03T16:52:52.413571Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:52:52.413571Z digest=sha256:3c48114eb04c3ecf5ae020b535e37d5f45ba31654ca5828f8055ef45672780bb

Observation 0b1d1384-657a-4d9f-b8a6-c3c5f9bd6585 · inbound

LinMU: Multimodal Understanding Made Linear cites this paper.

LinMU: Multimodal Understanding Made Linear A Survey on Vision-Language-Action Models for Autonomous Driving

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-16T18:33:15.213150Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T18:32:23.951566Z digest=sha256:ad4068853900831adb170d5bb142fce4d729c69be14baa3c3e63c7e588c6cb67

Observation 04c9f689-ac34-47c7-a01d-e8440510b18d · inbound

From Representational Complementarity to Dual Systems: Synergizing VLM and Vision-Only Backbones for End-to-End Driving cites this paper.

From Representational Complementarity to Dual Systems: Synergizing VLM and Vision-Only Backbones for End-to-End Driving A Survey on Vision-Language-Action Models for Autonomous Driving

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-04T06:07:09.468768Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T06:07:09.468768Z digest=sha256:e5d224f2c2acbab2df9f8e5350f6e9de851e8846723eb0e45152073b94da0e18

Observation 136b991e-d88f-48a0-9c4b-a5d941be43f3 · inbound

Steadily moving semi-infinite fracture in plane poroelasticity cites this paper.

Steadily moving semi-infinite fracture in plane poroelasticity A Survey on Vision-Language-Action Models for Autonomous Driving

Reference 48

Resolution
metadata mismatch
local_arxiv, observed 2026-07-05T11:41:02.544826Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-05T11:39:05.686584Z digest=sha256:bac20f81b71215da9854c6d43c8c66846e69f05930343702ab6a593a23e3ab26

Observation 85266423-3b9a-482c-b212-fadb1dfcaf91 · inbound

XEmbodied: A Foundation Model with Enhanced Geometric and Physical Cues for Large-Scale Embodied Environments cites this paper.

XEmbodied: A Foundation Model with Enhanced Geometric and Physical Cues for Large-Scale Embodied Environments A Survey on Vision-Language-Action Models for Autonomous Driving

Reference 48

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T05:51:10.420760Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T05:46:36.865150Z digest=sha256:949a1582ecd73861bfdaf63342f211dc017092bc468d3129b7d11ec5bb135127

Observation 24670dd2-623b-4306-9a19-12df51202285 · inbound

EgoDyn-Bench: Evaluating Ego-Motion Understanding in Vision-Centric Foundation Models for Autonomous Driving cites this paper.

EgoDyn-Bench: Evaluating Ego-Motion Understanding in Vision-Centric Foundation Models for Autonomous Driving A Survey on Vision-Language-Action Models for Autonomous Driving

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-05-10T00:19:47.159530Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T00:09:18.068337Z digest=sha256:091a4d09777dad3312918bf8f40eff3581a7ffb410306ce9ee670d671d2f975a

Observation 670b3c2b-9aab-4827-bfd5-e0b330b9271c · inbound

MVPruner: Dynamic Token Pruning for Accelerating Multi-view Vision-Language Models in Autonomous Driving cites this paper.

MVPruner: Dynamic Token Pruning for Accelerating Multi-view Vision-Language Models in Autonomous Driving A Survey on Vision-Language-Action Models for Autonomous Driving

Reference 15

Resolution
metadata mismatch
arxiv_id, observed 2026-06-29T18:23:51.152734Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-29T05:08:46.145616Z digest=sha256:a13458c11ed38deb96742c3e0c62098c30d75d6ddff71edb8a7d84d9684352f1

Observation 3158c713-bf1e-40eb-aa7b-8f6b7f88b616 · inbound

MVPruner: Dynamic Token Pruning for Accelerating Multi-view Vision-Language Models in Autonomous Driving cites this paper.

MVPruner: Dynamic Token Pruning for Accelerating Multi-view Vision-Language Models in Autonomous Driving A Survey on Vision-Language-Action Models for Autonomous Driving

Reference 15

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T21:37:24.180527Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-02T21:35:33.280591Z digest=sha256:2376d148d2dc4a3a9b20a4dcc8685f5fb538a14b4a9cfcd378df3765ccdd0046

Observation 0d0865d9-586c-436c-b941-718cebd07d84 · inbound

PriorEye: Geospatial Visual Priors for End-to-End Autonomous Driving cites this paper.

PriorEye: Geospatial Visual Priors for End-to-End Autonomous Driving A Survey on Vision-Language-Action Models for Autonomous Driving

Reference 29

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T09:55:41.492322Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-01T05:58:58.075150Z digest=sha256:da21d07c5f65d3324860855ceafaa2954c6269d3e04337b753b825b92c7290b2

Observation 8c159a26-866d-4ccf-a534-b4e5dccdc3dc · inbound

MoRAL: Sensor-Grounded BEV Reasoning for Compact VLMs toward Edge-Oriented Autonomous Driving cites this paper.

MoRAL: Sensor-Grounded BEV Reasoning for Compact VLMs toward Edge-Oriented Autonomous Driving A Survey on Vision-Language-Action Models for Autonomous Driving

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-04T07:11:20.776007Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T07:11:20.776007Z digest=sha256:2fc8acadc0e535359ed111b6afe4e6ae646443ec3358469b76c41ebd2883b069