Pith. sign in

Paper Citation Record · LEDGER

WiseAD: Knowledge Augmented End-to-End Autonomous Driving with Vision-Language Model

As of 14 August 2026, this Paper Citation Record lists 45 of 45 outbound references and 11 inbound Pith citation observations for arXiv:2412.09951.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2412.09951 v2

Coverage vector

measured 45 of 45 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-11T16:35:17.480171Z

measured 56 of 56 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00

measured 11 of 11 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:02:54.573103Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T11:08:03.455156Z

Reference resolution

45 of 45 outbound references displayed

  • verified exact0
  • verified fuzzy27
  • unresolved18
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 1880178c-caca-4d87-9ece-eb6aa29b50f1 · outbound

This paper cites com/apolloauto/apollo, note = Accessed: 2019-02-.

WiseAD: Knowledge Augmented End-to-End Autonomous Driving with Vision-Language Model com/apolloauto/apollo, note = Accessed: 2019-02-

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T16:35:17.970919Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T16:35:17.304746Z digest=sha256:babf4e84d803fd23759b9dbd00aef4b2b6fcada9c6a383a4db36fdece0eff3e1

Observation 5b23e33b-4dd9-4be2-9a21-455f2c53c504 · outbound

This paper cites Covla: Comprehensive vision-language-action dataset for autonomous driving.

WiseAD: Knowledge Augmented End-to-End Autonomous Driving with Vision-Language Model Covla: Comprehensive vision-language-action dataset for autonomous driving

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-11T16:35:17.307896Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:35:17.307896Z digest=sha256:d29d81abd8969824b77e4d829063b3d95c85074a7d2aef94472056e1e4a029fb

Observation 14e060d5-c002-4196-b079-769bd05a095e · outbound

This paper cites Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond.

WiseAD: Knowledge Augmented End-to-End Autonomous Driving with Vision-Language Model Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-11T16:35:17.310736Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:35:17.310736Z digest=sha256:68ada0ebe9e9225a9856b1e4674bd4645e02a925632a85b11eb98543d75c3d18

Observation 622d331e-4400-4ed0-ac79-0956095c41be · outbound

This paper cites nuscenes: A multimodal dataset for autonomous driving.

WiseAD: Knowledge Augmented End-to-End Autonomous Driving with Vision-Language Model nuscenes: A multimodal dataset for autonomous driving

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T16:35:17.964185Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T16:35:17.313978Z digest=sha256:ddab2ec85387ff358b21ad3014ad8303e4dc2848c6acaf0b8317733543cd5aeb

Observation 645a9e7a-831b-44a3-b211-e3f35a9fae71 · outbound

This paper cites Internvl: Scaling up vision foundation mod- els and aligning for generic visual-linguistic tasks.

WiseAD: Knowledge Augmented End-to-End Autonomous Driving with Vision-Language Model Internvl: Scaling up vision foundation mod- els and aligning for generic visual-linguistic tasks

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T16:35:17.958043Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T16:35:17.316825Z digest=sha256:2bf207294490b30f92464827e99d125f30877cf0f641bd06f1d27451c30dd778

Observation e5a368d7-abb4-4eec-8b3d-3f4d2dcd6800 · outbound

This paper cites Neat: Neural attention fields for end-to-end autonomous driving.

WiseAD: Knowledge Augmented End-to-End Autonomous Driving with Vision-Language Model Neat: Neural attention fields for end-to-end autonomous driving

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T16:35:17.951773Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T16:35:17.319482Z digest=sha256:62cb8db1baca28f4821fc4c5a3c5bf28f29bff796738b142f4034985aeeaf7d5

Observation bcc2bcb6-62c1-4c08-9a1a-1b70aebde2e1 · outbound

This paper cites MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices.

WiseAD: Knowledge Augmented End-to-End Autonomous Driving with Vision-Language Model MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-11T16:35:17.321824Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:35:17.321824Z digest=sha256:67b5f6b780357fc89899a05ca1ac0ed568d544fb86b86318bb738dc107fea781

Observation 0c86b0cd-f3a2-4af8-a5b5-890cb1b2e63b · outbound

This paper cites MobileVLM V2: Faster and Stronger Baseline for Vision Language Model.

WiseAD: Knowledge Augmented End-to-End Autonomous Driving with Vision-Language Model MobileVLM V2: Faster and Stronger Baseline for Vision Language Model

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-11T16:35:17.324427Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:35:17.324427Z digest=sha256:5c74c7b1a35825f5ffca8552775afdfcb4fe7cfee787c2e3eff1534efd9418b2

Observation 17fb9a6a-3a38-43fb-95bb-4bc160a48638 · outbound

This paper cites The cityscapes dataset for semantic urban scene understanding.

WiseAD: Knowledge Augmented End-to-End Autonomous Driving with Vision-Language Model The cityscapes dataset for semantic urban scene understanding

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T16:35:17.945615Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T16:35:17.327049Z digest=sha256:4ea93f1f490f1cd807e9542bce79686e1364c994701362e6155ac79f563a4d94

Observation 1e2a60dd-c23d-4c1e-a37e-54b30ea386c4 · outbound

This paper cites Talk2Car: Taking Control of Your Self-Driving Car.

WiseAD: Knowledge Augmented End-to-End Autonomous Driving with Vision-Language Model Talk2Car: Taking Control of Your Self-Driving Car

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-11T16:35:17.329129Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:35:17.329129Z digest=sha256:90c06b7205fe99d023a3c42f1c31c47817d2f385857cfbf2befcd9fea45e4eac

Observation 682cf184-36af-4604-9c6a-7715f8c19e1d · outbound

This paper cites Carla: An open urban driving simulator.

WiseAD: Knowledge Augmented End-to-End Autonomous Driving with Vision-Language Model Carla: An open urban driving simulator

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T16:35:17.939533Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T16:35:17.331471Z digest=sha256:a6d5b5b6abe3cd4237c51fa2c79e63baa0d948dcbfaad210e7f3ec924276359c

Observation 8e4f7473-bce7-4995-9b69-b98970c2e7e1 · outbound

This paper cites Vision meets robotics: The kitti dataset.

WiseAD: Knowledge Augmented End-to-End Autonomous Driving with Vision-Language Model Vision meets robotics: The kitti dataset

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-11T16:35:17.333563Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:35:17.333563Z digest=sha256:e4bbbb8f179dce68b85218308e9a7f511f5ebdca42b10cecb026177a7d8adbee

Observation 3301d1a6-b25e-4c8c-b185-803502d84597 · outbound

This paper cites St-p3: End-to-end vision-based au- tonomous driving via spatial-temporal feature learning.

WiseAD: Knowledge Augmented End-to-End Autonomous Driving with Vision-Language Model St-p3: End-to-end vision-based au- tonomous driving via spatial-temporal feature learning

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T16:35:17.928410Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T16:35:17.335532Z digest=sha256:78ab888fa8b946b4d469b141607a457ecc9142492fc7be34b60ccb7a583aa813

Observation c578c841-b2db-4adc-bad4-6706a7895b37 · outbound

This paper cites Planning-oriented autonomous driv- ing.

WiseAD: Knowledge Augmented End-to-End Autonomous Driving with Vision-Language Model Planning-oriented autonomous driv- ing

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T16:35:17.921246Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T16:35:17.337838Z digest=sha256:0a044d125c4b2f9af0b5e2dac16ce28d1c0d68002a8f2c77a0c519da026495b3

Observation c0cc2082-80f3-4449-9f35-5eff095b33d9 · outbound

This paper cites Vad: Vectorized scene representation for efficient autonomous driving.

WiseAD: Knowledge Augmented End-to-End Autonomous Driving with Vision-Language Model Vad: Vectorized scene representation for efficient autonomous driving

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T16:35:17.914080Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T16:35:17.339932Z digest=sha256:b40bb3208cd66763a755ee4d306eab3a74af3fb96d09a7c9d3069d8263832ddb

Observation ebc3f10e-1f2c-4ae8-a486-c8931206f787 · outbound

This paper cites Adapt: Action-aware driving caption transformer.

WiseAD: Knowledge Augmented End-to-End Autonomous Driving with Vision-Language Model Adapt: Action-aware driving caption transformer

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T16:35:17.906775Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T16:35:17.342033Z digest=sha256:48c46c4f58b461afc1aa3867721a301fe8fb6ab368eca8d96ac94a74fbf86f3e

Observation 6f1212eb-411c-4f82-90d3-72b1c328e6c8 · outbound

This paper cites Textual explanations for self-driving vehicles.

WiseAD: Knowledge Augmented End-to-End Autonomous Driving with Vision-Language Model Textual explanations for self-driving vehicles

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T16:35:17.899543Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T16:35:17.343961Z digest=sha256:1e3156ad8b12169ba9d49273cf158e6e280395e37e56ae50bee85c0ecca6db5c

Observation a4eeab96-57ef-4e2a-a61c-e3c1ab78ad37 · outbound

This paper cites Grounding human-to-vehicle advice for self- driving vehicles.

WiseAD: Knowledge Augmented End-to-End Autonomous Driving with Vision-Language Model Grounding human-to-vehicle advice for self- driving vehicles

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T16:35:17.892485Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T16:35:17.346081Z digest=sha256:37a7f993fcedc31fffdc580494f7bb6d2a82b109ad293f8c8cab696b16d46397

Observation 19023ce4-a147-41d3-8dbd-36677f08a837 · outbound

This paper cites Knowledge representation and reasoning.

WiseAD: Knowledge Augmented End-to-End Autonomous Driving with Vision-Language Model Knowledge representation and reasoning

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T16:35:17.885258Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T16:35:17.349194Z digest=sha256:619a88c3370f8f6b3d91e2e8164f49bf5a419dca714dd42f34be53fe57735488

Observation 4ddba1f5-f351-4868-bdb8-d6cf228287ce · outbound

This paper cites Towards Knowledge-driven Autonomous Driving.

WiseAD: Knowledge Augmented End-to-End Autonomous Driving with Vision-Language Model Towards Knowledge-driven Autonomous Driving

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-11T16:35:17.351260Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:35:17.351260Z digest=sha256:b3f67b625102bb784958a2323be50f56bd8f8648cb2ccfe9a6585c4da0635d99

Observation 1c958d98-78d3-44a9-96d2-4462074a7ef2 · outbound

This paper cites Automated Evaluation of Large Vision-Language Models on Self-driving Corner Cases.

WiseAD: Knowledge Augmented End-to-End Autonomous Driving with Vision-Language Model Automated Evaluation of Large Vision-Language Models on Self-driving Corner Cases

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-11T16:35:17.356576Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:35:17.356576Z digest=sha256:167885a344e79d84dbbb825085f26f804dcb3695680046cdd481f8f1881ac433

Observation 48cf9d2c-4176-4955-b600-f7948d885e7b · outbound

This paper cites Visual instruction tuning.

WiseAD: Knowledge Augmented End-to-End Autonomous Driving with Vision-Language Model Visual instruction tuning

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T16:35:17.878682Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T16:35:17.359217Z digest=sha256:d131d8869febad95968413937670486bb58de8355eb480870aa150ad99b46117

Observation 9878622d-d1ae-4b80-aeed-a592d766066b · outbound

This paper cites Video swin transformer.

WiseAD: Knowledge Augmented End-to-End Autonomous Driving with Vision-Language Model Video swin transformer

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-11T16:35:17.361897Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:35:17.361897Z digest=sha256:4fed8a2d469d79ca68a1bf72bfa3192030e4858ae087f94477a0fbf5ec03eb6f

Observation dad24fca-80b4-48a4-9967-3190f0f76469 · outbound

This paper cites Decoupled Weight Decay Regularization.

WiseAD: Knowledge Augmented End-to-End Autonomous Driving with Vision-Language Model Decoupled Weight Decay Regularization

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-11T16:35:17.364737Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:35:17.364737Z digest=sha256:6ae9fb82e2621e50722a63d42062e709620e7fdfaeb8f42918028572876a5a93

Observation 97edd19a-0225-4144-b8be-64f6aa9998fc · outbound

This paper cites DeepSeek-VL: Towards Real-World Vision-Language Understanding.

WiseAD: Knowledge Augmented End-to-End Autonomous Driving with Vision-Language Model DeepSeek-VL: Towards Real-World Vision-Language Understanding

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-11T16:35:17.367328Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:35:17.367328Z digest=sha256:566aede47071842402bab9f8a1621630d423b2c549f72e62f805b320391c8f1a

Observation d30b1989-2616-4541-87b4-e9bb8ec5e882 · outbound

This paper cites Drama: Joint risk localization and captioning in driving.

WiseAD: Knowledge Augmented End-to-End Autonomous Driving with Vision-Language Model Drama: Joint risk localization and captioning in driving

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T16:35:17.868792Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T16:35:17.370078Z digest=sha256:783635bc254fef99b61904b618b8d417ca7e383960dc4903ac2ec0b804b95a13

Observation c0e05998-dc84-4fe9-ac0b-44c3f0979833 · outbound

This paper cites LingoQA: Visual Question Answering for Autonomous Driving.

WiseAD: Knowledge Augmented End-to-End Autonomous Driving with Vision-Language Model LingoQA: Visual Question Answering for Autonomous Driving

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-11T16:35:17.372660Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:35:17.372660Z digest=sha256:ee95f2918bfc2a1c4606ac63f6c67ab6d5b9234f47b4942421d80dc4c34d8021

Observation 410d1120-7585-4125-b222-e01e03ec20b1 · outbound

This paper cites Bleu: a method for automatic evaluation of machine translation.

WiseAD: Knowledge Augmented End-to-End Autonomous Driving with Vision-Language Model Bleu: a method for automatic evaluation of machine translation

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-11T16:35:17.375387Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:35:17.375387Z digest=sha256:647c4a163ce4eaeca80f57d73291aaede35d651d27e8b251d37e3395f159e113

Observation 2fe4b077-1607-4556-bd86-6d15329f0360 · outbound

This paper cites Multi- modal fusion transformer for end-to-end autonomous driving.

WiseAD: Knowledge Augmented End-to-End Autonomous Driving with Vision-Language Model Multi- modal fusion transformer for end-to-end autonomous driving

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T16:35:17.859030Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T16:35:17.377992Z digest=sha256:3ec1b794ddce49f74857e0924db8a26cbc6be9a93c5c2c3236775a2753d0e7ef

Observation d5caaf17-ff3b-4c2a-a8d1-6b34d3194ad5 · outbound

This paper cites Nuscenes-qa: A multi-modal visual question answering benchmark for autonomous driving scenario.

WiseAD: Knowledge Augmented End-to-End Autonomous Driving with Vision-Language Model Nuscenes-qa: A multi-modal visual question answering benchmark for autonomous driving scenario

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T16:35:17.853042Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T16:35:17.380478Z digest=sha256:fb4b2fe1deae2dd6cf9b261f8d3d7a675ce080f6c15f1c0c728fff8fc81ee76a

Observation 411fabf2-b312-41ff-a21d-218551321fca · outbound

This paper cites Learning transferable visual models from natural language supervision.

WiseAD: Knowledge Augmented End-to-End Autonomous Driving with Vision-Language Model Learning transferable visual models from natural language supervision

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T16:35:17.846819Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T16:35:17.383143Z digest=sha256:d7634b71f81d032f333523f41be61bbbc84a4c0fce257459d0ad196beda1673c

Observation 1ef144fc-ab28-4bd3-9ba5-6e6d4c461b76 · outbound

This paper cites Toward driving scene understanding: A dataset for learning driver behavior and causal reasoning.

WiseAD: Knowledge Augmented End-to-End Autonomous Driving with Vision-Language Model Toward driving scene understanding: A dataset for learning driver behavior and causal reasoning

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T16:35:17.840567Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T16:35:17.385727Z digest=sha256:8744e4f95b4ff17505cecba6ead16a51a4f8d4e08f31b41e7fcff34fb9d678a7

Observation 924f91fb-7e58-4504-8af0-8d21f30181a2 · outbound

This paper cites Safety-enhanced autonomous driving using inter- pretable sensor fusion transformer.

WiseAD: Knowledge Augmented End-to-End Autonomous Driving with Vision-Language Model Safety-enhanced autonomous driving using inter- pretable sensor fusion transformer

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T16:35:17.834026Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T16:35:17.388233Z digest=sha256:1e333cda606dcef22e54edf3853a7b37bf4c98bfcee2436989fca2e68eb361ff

Observation 1654e167-bfe9-473a-accb-303b88908a5e · outbound

This paper cites Reasonnet: End-to-end driv- ing with temporal and global reasoning.

WiseAD: Knowledge Augmented End-to-End Autonomous Driving with Vision-Language Model Reasonnet: End-to-end driv- ing with temporal and global reasoning

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T16:35:17.826577Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T16:35:17.390816Z digest=sha256:a3cb303d588d405a1dde13475f0ff6b4e4797b84156c4f85451c79692884f443

Observation 25cce3c1-47ff-40e9-9406-661a45aa435c · outbound

This paper cites Lmdrive: Closed-loop end-to-end driving with large language models.

WiseAD: Knowledge Augmented End-to-End Autonomous Driving with Vision-Language Model Lmdrive: Closed-loop end-to-end driving with large language models

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T16:35:17.818905Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T16:35:17.456133Z digest=sha256:d7d8c6fc63be56a6278fde1ebb34336229184cf207b23417e47f603bcfca6420

Observation ca4a2359-d481-466f-a887-409df766c2a0 · outbound

This paper cites DriveLM: Driving with Graph Visual Question Answering.

WiseAD: Knowledge Augmented End-to-End Autonomous Driving with Vision-Language Model DriveLM: Driving with Graph Visual Question Answering

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-11T16:35:17.458991Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:35:17.458991Z digest=sha256:8d932226c104aa7221084077ad3566368b69562b2301871bfd81404c0118c052

Observation f3846abe-56ed-47db-9b7c-0c5a91bc6a02 · outbound

This paper cites Scalability in perception for autonomous driving: Waymo open dataset.

WiseAD: Knowledge Augmented End-to-End Autonomous Driving with Vision-Language Model Scalability in perception for autonomous driving: Waymo open dataset

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T16:35:17.811052Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T16:35:17.461725Z digest=sha256:f8fe7ad3b19d5f60b504ba886a4f1fa72d2f04ee77f52b3aae2640a2fb3b3296

Observation 263d49a2-2de9-4a9f-a153-fb4471654e88 · outbound

This paper cites DriveVLM: The Convergence of Autonomous Driving and Large Vision-Language Models.

WiseAD: Knowledge Augmented End-to-End Autonomous Driving with Vision-Language Model DriveVLM: The Convergence of Autonomous Driving and Large Vision-Language Models

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-11T16:35:17.464331Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:35:17.464331Z digest=sha256:3916f516355e3b79a136bae1345adb07ff4a5adf3cd28c4d3df4d0a652b85bc4

Observation 8d3ff5ea-8413-49a9-9358-a88a34a9c8a0 · outbound

This paper cites Llama: Open and efficient foundation language models.

WiseAD: Knowledge Augmented End-to-End Autonomous Driving with Vision-Language Model Llama: Open and efficient foundation language models

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T16:35:17.802362Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T16:35:17.467073Z digest=sha256:0df623005c1587d4dd58f814618ae6488799ece58654d29c591c8059b63227a1

Observation de9efed9-54b6-4f24-ba66-716930cec493 · outbound

This paper cites Cider: Consensus-based image description evalu- ation.

WiseAD: Knowledge Augmented End-to-End Autonomous Driving with Vision-Language Model Cider: Consensus-based image description evalu- ation

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T16:35:17.793871Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T16:35:17.469206Z digest=sha256:ec3befeb849a081fda10f9ab195ce6754b3c5501398564530d3203e7f2e97209

Observation 7b44f744-6236-46e4-a677-9ec13d34fa39 · outbound

This paper cites Drivemlm: Aligning multi-modal large language models with behavioral planning states for autonomous driving.

WiseAD: Knowledge Augmented End-to-End Autonomous Driving with Vision-Language Model Drivemlm: Aligning multi-modal large language models with behavioral planning states for autonomous driving

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-11T16:35:17.471390Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:35:17.471390Z digest=sha256:c3e33128ba11e39c8b7fe27c7ba2332603e3344f2fba1ab632df2ea8f6a8aa43

Observation aec87cd5-522e-4aa9-b222-b557e5dfb834 · outbound

This paper cites Drivegpt4: Interpretable end-to-end autonomous driving via large language model.

WiseAD: Knowledge Augmented End-to-End Autonomous Driving with Vision-Language Model Drivegpt4: Interpretable end-to-end autonomous driving via large language model

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T16:35:17.786459Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T16:35:17.473637Z digest=sha256:a7b2e3339aeaa8cd982c15260b23b06fcaf2c8dc90ba0f875f3d8a82cfde2cdb

Observation 4de2d74b-e40c-4096-8cd4-d42eef7e8b65 · outbound

This paper cites Rag-driver: Gen- eralisable driving explanations with retrieval-augmented in- context learning in multi-modal large language model.

WiseAD: Knowledge Augmented End-to-End Autonomous Driving with Vision-Language Model Rag-driver: Gen- eralisable driving explanations with retrieval-augmented in- context learning in multi-modal large language model

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-11T16:35:17.476047Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:35:17.476047Z digest=sha256:5cb79c8feca308b5e23ea0ca35e7799390e308a78ef21dfe73a870351d0d42d2

Observation 04433515-9eb9-498a-a7b5-6ab8884ea5c5 · outbound

This paper cites End-to-end urban driving by imitating a re- inforcement learning coach.

WiseAD: Knowledge Augmented End-to-End Autonomous Driving with Vision-Language Model End-to-end urban driving by imitating a re- inforcement learning coach

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T16:35:17.778873Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T16:35:17.477994Z digest=sha256:407d109bfe7ce08ddbc331719406abf3af92b0736e9e1b5aa7dbb8ad8ea68a00

Observation a43f6623-837b-4426-bd67-702236fdcbc4 · outbound

This paper cites Embodied Understanding of Driving Scenarios.

WiseAD: Knowledge Augmented End-to-End Autonomous Driving with Vision-Language Model Embodied Understanding of Driving Scenarios

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-11T16:35:17.480171Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:35:17.480171Z digest=sha256:134c49d76ac8623c7c037a886851e86a1f79837e425bd4824fc07749aec8cff6

Pith citing papers

Observation b31184b9-f863-4bba-9aa7-5523647c443b · inbound

Chain-of-Thought for Autonomous Driving: A Comprehensive Survey and Future Prospects cites this paper.

Chain-of-Thought for Autonomous Driving: A Comprehensive Survey and Future Prospects WiseAD: Knowledge Augmented End-to-End Autonomous Driving with Vision-Language Model

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T14:02:54.573103Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:02:54.573103Z digest=sha256:db67f554a4d7f5715b17055dd9759f020511829ff0599198b6f75aee9dccb73a

Observation b81e7337-aced-4e00-a817-1454e2c7d709 · inbound

AutoVLA: A Vision-Language-Action Model for End-to-End Autonomous Driving with Adaptive Reasoning and Reinforcement Fine-Tuning cites this paper.

AutoVLA: A Vision-Language-Action Model for End-to-End Autonomous Driving with Adaptive Reasoning and Reinforcement Fine-Tuning WiseAD: Knowledge Augmented End-to-End Autonomous Driving with Vision-Language Model

Reference 81

Resolution
verified exact
arxiv_id, observed 2026-05-14T21:46:44.143405Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-14T21:46:43.955825Z digest=sha256:9a8b92c4352ab557c8898d57b8bbffe33e617604e8c7cf7e82323d63a9542d46

Observation 9892e3a9-571d-4b26-9671-3812f91da7bb · inbound

DVGT-2: Vision-Geometry-Action Model for Autonomous Driving at Scale cites this paper.

DVGT-2: Vision-Geometry-Action Model for Autonomous Driving at Scale WiseAD: Knowledge Augmented End-to-End Autonomous Driving with Vision-Language Model

Reference 84

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T23:03:25.066904Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-13T22:59:57.744015Z digest=sha256:3f62d1632076e1e2a37845066919d6f3d5ed4fd47dde543d7e47103dfe317192

Observation 0daa3f72-1ea6-4a43-ae00-ea9382f21466 · inbound

The Blind Spot of Adaptation: Quantifying and Mitigating Forgetting in Fine-tuned Driving Models cites this paper.

The Blind Spot of Adaptation: Quantifying and Mitigating Forgetting in Fine-tuned Driving Models WiseAD: Knowledge Augmented End-to-End Autonomous Driving with Vision-Language Model

Reference 57

Resolution
verified exact
arxiv_id, observed 2026-05-10T22:15:50.119001Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-10T20:04:46.144856Z digest=sha256:3929059d49dfebae81cac8f08994e8029508b0c298c0156294429972259c6cbf

Observation 665c548c-f329-45b4-a727-abbe34529c3d · inbound

Towards Safe Mobility: A Unified Transportation Foundation Model enabled by Open-Ended Vision-Language Dataset cites this paper.

Towards Safe Mobility: A Unified Transportation Foundation Model enabled by Open-Ended Vision-Language Dataset WiseAD: Knowledge Augmented End-to-End Autonomous Driving with Vision-Language Model

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-11T19:06:11.211754Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-08T12:37:17.646552Z digest=sha256:c80a3aa9eefdc8ebdc16344b812c5ee08631905693845ff9e3fc9d3a23dd09ab

Observation ea951d77-167c-4181-8498-6d1dcb74d744 · inbound

DeepSight: Long-Horizon World Modeling via Latent States Prediction for End-to-End Autonomous Driving cites this paper.

DeepSight: Long-Horizon World Modeling via Latent States Prediction for End-to-End Autonomous Driving WiseAD: Knowledge Augmented End-to-End Autonomous Driving with Vision-Language Model

Reference 31

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T06:31:26.091928Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-05-12T04:13:37.421188Z digest=sha256:9a14f0f54c6195391212dcbcca3d0e1adb2d2c02c144b5777d16c8410587266b

Observation 3de3271e-acb4-42bf-b7e0-4e54a35120b7 · inbound

VL-DPO: Vision-Language-Guided Finetuning for Preference-Aligned Autonomous Driving cites this paper.

VL-DPO: Vision-Language-Guided Finetuning for Preference-Aligned Autonomous Driving WiseAD: Knowledge Augmented End-to-End Autonomous Driving with Vision-Language Model

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-05-20T05:43:06.036435Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-20T05:39:04.567741Z digest=sha256:3944aabebe8b61a8c8c1eb8edfccb254aaa3721498adba41410c6bcfa5dc16c4

Observation 56d2d32b-c206-4989-b4bf-c83d9d0ce7d1 · inbound

What to Test Next: Interpretable Coverage Gap Discovery in Driving VLMs cites this paper.

What to Test Next: Interpretable Coverage Gap Discovery in Driving VLMs WiseAD: Knowledge Augmented End-to-End Autonomous Driving with Vision-Language Model

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-07-01T22:16:16.230998Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-06-28T15:36:25.219607Z digest=sha256:26c94b2cafc47451d2fe633ca685489983f47950441407bc9538fd508e60396b

Observation c3e7aca4-d184-4b78-ac52-2ce33e246204 · inbound

VLGA: Vision-Language-Geometry-Action Models for Autonomous Driving cites this paper.

VLGA: Vision-Language-Geometry-Action Models for Autonomous Driving WiseAD: Knowledge Augmented End-to-End Autonomous Driving with Vision-Language Model

Reference 52

Resolution
verified exact
arxiv_id, observed 2026-07-03T11:08:03.456655Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-06-27T09:40:51.407286Z digest=sha256:79fac534ae7dcd3d952be809dd3a72606a128d1da708d16928087a45908fd3d2

Observation b2937be0-0f76-486b-a91e-476d575d6336 · inbound

GeoWorldAD: Geometry World Action Model for Autonomous Driving cites this paper.

GeoWorldAD: Geometry World Action Model for Autonomous Driving WiseAD: Knowledge Augmented End-to-End Autonomous Driving with Vision-Language Model

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-01T17:46:55.844970Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T17:46:55.844970Z digest=sha256:588c8cd181422a6e6956ce2fc1e41a58273ed8dd798ab52e7a53fd141da1021c

Observation ccac55f8-5bb8-43f1-bf0f-378109b7f667 · inbound

Deferred Exposure of Future Trajectories for Verifiable Reasoning in Autonomous Driving VLMs cites this paper.

Deferred Exposure of Future Trajectories for Verifiable Reasoning in Autonomous Driving VLMs WiseAD: Knowledge Augmented End-to-End Autonomous Driving with Vision-Language Model

Reference 79

Resolution
unresolved
no resolver link, observed 2026-08-07T00:13:57.120069Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:13:57.120069Z digest=sha256:ed76493a6b20202946332df661c5ae0eecb25c4b8c56e360ba59b3932ec1fb42