Pith. sign in

Paper Citation Record · LEDGER

Are Vision LLMs Road-Ready? A Comprehensive Benchmark for Safety-Critical Driving Video Understanding

As of 20 August 2026, this Paper Citation Record lists 55 of 55 outbound references and 2 inbound Pith citation observations for arXiv:2504.14526.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2504.14526 v1

Coverage vector

measured 55 of 55 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-16T11:50:21.321263Z

measured 57 of 57 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T06:02:13.157665Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-18T12:46:24.143801Z

Reference resolution

55 of 55 outbound references displayed

  • verified exact0
  • verified fuzzy3
  • unresolved51
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 2ea296ee-e7ac-4d61-8643-4d619357ae92 · outbound

This paper cites an unresolved cited work.

Are Vision LLMs Road-Ready? A Comprehensive Benchmark for Safety-Critical Driving Video Understanding Unresolved cited work

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-16T11:50:21.062326Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:50:21.062326Z digest=sha256:265240de20d5afca4f75f27dc3fd6bd8bbc1a98c5cb987d338cd758223dabed4

Observation f66d3cf6-0bc1-4116-af42-9b2a149aae8d · outbound

This paper cites an unresolved cited work.

Are Vision LLMs Road-Ready? A Comprehensive Benchmark for Safety-Critical Driving Video Understanding Unresolved cited work

Reference 2

Resolution
unresolved
raw_fallback, observed 2026-08-16T11:50:22.118913Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-16T11:50:21.072515Z digest=sha256:c6a21811205bc5517099ac79be4ed0d51f5969383b18f637298af156bb214e92

Observation bcf0b43e-e3a6-4df4-9799-e14cb1c8e6f7 · outbound

This paper cites Premise Order Matters in Reasoning with Large Language Models.

Are Vision LLMs Road-Ready? A Comprehensive Benchmark for Safety-Critical Driving Video Understanding Premise Order Matters in Reasoning with Large Language Models

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-16T11:50:21.077215Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:50:21.077215Z digest=sha256:64e5ee0ec4231fa2fcb7708bc5dd625166a3bb55b4bc3c33b61160b1a8a269bb

Observation ee6872ff-2327-451f-81ac-5eb17e0a6661 · outbound

This paper cites Talk2Car: Taking Control of Your Self-Driving Car.

Are Vision LLMs Road-Ready? A Comprehensive Benchmark for Safety-Critical Driving Video Understanding Talk2Car: Taking Control of Your Self-Driving Car

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-16T11:50:21.081969Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:50:21.081969Z digest=sha256:1aa895717d7e49ca2d8e6755c87913e9274a4f2819f7ee3b6d9ef5d102237387

Observation 7387aec3-7d21-4970-96cd-9d1749843adc · outbound

This paper cites Video-MME: The First-Ever Comprehensive Evaluation Benchmark of Multi-modal LLMs in Video Analysis.

Are Vision LLMs Road-Ready? A Comprehensive Benchmark for Safety-Critical Driving Video Understanding Video-MME: The First-Ever Comprehensive Evaluation Benchmark of Multi-modal LLMs in Video Analysis

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-16T11:50:21.086946Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:50:21.086946Z digest=sha256:496dacee5c0cf78fb2683e65d6e0a7092037fe5514e95af488532239bd0bb90e

Observation b036f884-c800-4ff1-bb27-317a9569528f · outbound

This paper cites an unresolved cited work.

Are Vision LLMs Road-Ready? A Comprehensive Benchmark for Safety-Critical Driving Video Understanding Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-08-16T11:50:22.103995Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-16T11:50:21.092257Z digest=sha256:ce67d44825fb8c3933c4e71d121738be9881dcdbddd5f0d50a9a2ead06a165da

Observation 7765deb5-3c29-4c7c-8101-45f8ece9636b · outbound

This paper cites SURDS: Benchmarking Spatial Understanding and Reasoning in Driving Scenarios with Vision Language Models.

Are Vision LLMs Road-Ready? A Comprehensive Benchmark for Safety-Critical Driving Video Understanding SURDS: Benchmarking Spatial Understanding and Reasoning in Driving Scenarios with Vision Language Models

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-16T11:50:21.097737Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:50:21.097737Z digest=sha256:4af0f8b6b2569e58d8f8127fad617f796905a4931486d1eefbfd0719a0232cf3

Observation 9b4aa964-c050-48f1-96ad-77371f23de0a · outbound

This paper cites DME-Driver: Integrating Human Decision Logic and 3D Scene Perception in Autonomous Driving.

Are Vision LLMs Road-Ready? A Comprehensive Benchmark for Safety-Critical Driving Video Understanding DME-Driver: Integrating Human Decision Logic and 3D Scene Perception in Autonomous Driving

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-16T11:50:21.102980Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:50:21.102980Z digest=sha256:3aecff477a569254a4f2437af3833c7253d823091545d3dbf97b9206cc5f0cc1

Observation 8d008a8a-f91b-4862-92a6-6dcbcc9d813d · outbound

This paper cites an unresolved cited work.

Are Vision LLMs Road-Ready? A Comprehensive Benchmark for Safety-Critical Driving Video Understanding Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-08-16T11:50:22.087520Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-16T11:50:21.107966Z digest=sha256:9859fde4234390629fe670a8933e0da8869e40387a4541ec5020b93d7b7ada33

Observation c51e2dcc-bb6a-40de-9771-7d4d15820722 · outbound

This paper cites GPT-4o System Card.

Are Vision LLMs Road-Ready? A Comprehensive Benchmark for Safety-Critical Driving Video Understanding GPT-4o System Card

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-16T11:50:21.112624Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:50:21.112624Z digest=sha256:63e2809c4a4e32271a3d24f291923064f260f3f84a71876d76884709823dce0d

Observation cb26d1d4-f317-4f39-a4d2-43a1dea79e3f · outbound

This paper cites an unresolved cited work.

Are Vision LLMs Road-Ready? A Comprehensive Benchmark for Safety-Critical Driving Video Understanding Unresolved cited work

Reference 11

Resolution
unresolved
raw_fallback, observed 2026-08-16T11:50:22.071361Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-16T11:50:21.117347Z digest=sha256:7026b0d2f3a670f957d25b0bf7761500eef7697b9ea975004a234ec8321633dc

Observation 461ae905-99e7-44bc-953c-14cf6fa76900 · outbound

This paper cites Challenges and Applications of Large Language Models.

Are Vision LLMs Road-Ready? A Comprehensive Benchmark for Safety-Critical Driving Video Understanding Challenges and Applications of Large Language Models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-16T11:50:21.122481Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:50:21.122481Z digest=sha256:e2f230685e40287c9b668d09ceb6789d6c5bcc076ff25d3376ca4e06663edc21

Observation 453fb0e5-0ac4-4c38-84a0-fe51e867bfa2 · outbound

This paper cites an unresolved cited work.

Are Vision LLMs Road-Ready? A Comprehensive Benchmark for Safety-Critical Driving Video Understanding Unresolved cited work

Reference 13

Resolution
unresolved
raw_fallback, observed 2026-08-16T11:50:22.053425Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-16T11:50:21.127534Z digest=sha256:3d50e381092515ed0cb2798e76cbc9358748e476c64bbfd31291494bab07b3e7

Observation 112366af-6e59-481b-a275-53564d48e8cf · outbound

This paper cites LLaVA-OneVision: Easy Visual Task Transfer.

Are Vision LLMs Road-Ready? A Comprehensive Benchmark for Safety-Critical Driving Video Understanding LLaVA-OneVision: Easy Visual Task Transfer

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-16T11:50:21.136424Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:50:21.136424Z digest=sha256:6d80c0d6186d6a7cb0dd70ace47fbd6a53f0ca3962f0f91c97cfd79bd93eb9c1

Observation 242d2273-e193-4161-99ac-059bca802ec9 · outbound

This paper cites an unresolved cited work.

Are Vision LLMs Road-Ready? A Comprehensive Benchmark for Safety-Critical Driving Video Understanding Unresolved cited work

Reference 15

Resolution
unresolved
raw_fallback, observed 2026-08-16T11:50:22.019082Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-16T11:50:21.140669Z digest=sha256:7a01c64eae15c5de74bd2baf9c5a69719f082a26a443e7d6c31678feeade0dc0

Observation 7b36ac52-d44a-4276-9c10-2f291f8fdab5 · outbound

This paper cites an unresolved cited work.

Are Vision LLMs Road-Ready? A Comprehensive Benchmark for Safety-Critical Driving Video Understanding Unresolved cited work

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-16T11:50:21.144962Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:50:21.144962Z digest=sha256:aa4a9c57410946e24e14a34fad38d8bc25b0de4ba404024113fc2c3650060460

Observation 51db16c5-eadf-46d1-8e36-6468982aceea · outbound

This paper cites Video-LLaVA: Learning United Visual Representation by Alignment Before Projection.

Are Vision LLMs Road-Ready? A Comprehensive Benchmark for Safety-Critical Driving Video Understanding Video-LLaVA: Learning United Visual Representation by Alignment Before Projection

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-16T11:50:21.149521Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:50:21.149521Z digest=sha256:d8f180a63d21f575660d82dcba5363233e92ee9b574a642ce9b5b4bca2a3f7c8

Observation df2953b1-b1b7-424a-9dd1-1bd03f30e32f · outbound

This paper cites an unresolved cited work.

Are Vision LLMs Road-Ready? A Comprehensive Benchmark for Safety-Critical Driving Video Understanding Unresolved cited work

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-16T11:50:21.154031Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:50:21.154031Z digest=sha256:567e181cf9cd56d18ac12335f8159ef1a661d39e3e1dd9eda66d3cc89daa6c78

Observation 39200787-f35b-4cc3-a6a8-c66f4a622fe6 · outbound

This paper cites an unresolved cited work.

Are Vision LLMs Road-Ready? A Comprehensive Benchmark for Safety-Critical Driving Video Understanding Unresolved cited work

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-16T11:50:21.158613Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:50:21.158613Z digest=sha256:9aa56db008c306b9a9c575719899033285e17cd167ef1cb63c1f943f0396acf1

Observation 9a17b74c-994b-49cb-9b0e-27cd0f9d8ba9 · outbound

This paper cites an unresolved cited work.

Are Vision LLMs Road-Ready? A Comprehensive Benchmark for Safety-Critical Driving Video Understanding Unresolved cited work

Reference 20

Resolution
unresolved
raw_fallback, observed 2026-08-16T11:50:21.967414Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-16T11:50:21.162970Z digest=sha256:cfcd19c8a896d972565723b2d3ab67dcb5a8f63e87340f9cb4659605562aecba

Observation d11362cc-da97-42be-b3a1-cfed466fab6b · outbound

This paper cites an unresolved cited work.

Are Vision LLMs Road-Ready? A Comprehensive Benchmark for Safety-Critical Driving Video Understanding Unresolved cited work

Reference 21

Resolution
unresolved
raw_fallback, observed 2026-08-16T11:50:21.950871Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-16T11:50:21.167517Z digest=sha256:b59af6660cd809eff88aac9aa1338e2985c96cb7b711b4bec16ed97075db17a8

Observation d12daf5d-6636-4d67-8691-1631ebd45611 · outbound

This paper cites an unresolved cited work.

Are Vision LLMs Road-Ready? A Comprehensive Benchmark for Safety-Critical Driving Video Understanding Unresolved cited work

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-16T11:50:21.177509Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:50:21.177509Z digest=sha256:c523f8332d83bd3160503d759c7ddf4a5a56c09f6b503cabc20ee9a9aa3d13dd

Observation 7608bcac-a179-45b7-8ba5-c05786e71591 · outbound

This paper cites an unresolved cited work.

Are Vision LLMs Road-Ready? A Comprehensive Benchmark for Safety-Critical Driving Video Understanding Unresolved cited work

Reference 23

Resolution
unresolved
raw_fallback, observed 2026-08-16T11:50:21.924023Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-16T11:50:21.188090Z digest=sha256:45b7e9db5bc352b686cdddcc1527e7d4bc75b2b0f6aa292ec15da74393acd4c8

Observation 739891be-773f-41cc-ad25-0980f8766019 · outbound

This paper cites GPT-Driver: Learning to Drive with GPT.

Are Vision LLMs Road-Ready? A Comprehensive Benchmark for Safety-Critical Driving Video Understanding GPT-Driver: Learning to Drive with GPT

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-16T11:50:21.192672Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:50:21.192672Z digest=sha256:21f74b213e2caee22fc28fb8731efb2b3538d45b62dd2f6bc44d1e48c1e449f4

Observation 8fd4f3a1-1fb5-4524-8983-0a9288234a31 · outbound

This paper cites an unresolved cited work.

Are Vision LLMs Road-Ready? A Comprehensive Benchmark for Safety-Critical Driving Video Understanding Unresolved cited work

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-16T11:50:21.197489Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:50:21.197489Z digest=sha256:6a5e39bad7f6036b451556309ffaf9ace3536a153537f0d233c2e5e3ac234ff6

Observation a195190f-d5c3-46e6-98cc-bf574df4c644 · outbound

This paper cites an unresolved cited work.

Are Vision LLMs Road-Ready? A Comprehensive Benchmark for Safety-Critical Driving Video Understanding Unresolved cited work

Reference 26

Resolution
unresolved
raw_fallback, observed 2026-08-16T11:50:21.896628Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-16T11:50:21.202293Z digest=sha256:1c02455d026c79290b323fa2fb968f4f17379c7d4c60fa0dadb80edccb301b53

Observation f14e36e6-bda0-4f69-92de-85bcf54ee430 · outbound

This paper cites an unresolved cited work.

Are Vision LLMs Road-Ready? A Comprehensive Benchmark for Safety-Critical Driving Video Understanding Unresolved cited work

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-16T11:50:21.207515Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:50:21.207515Z digest=sha256:ce69a0af68adb1c5741af44ecf47b57eb8dd3b052f6ecd9400d7a123ce718752

Observation fc02c19b-75cb-4f11-92d7-876c0066a1c4 · outbound

This paper cites an unresolved cited work.

Are Vision LLMs Road-Ready? A Comprehensive Benchmark for Safety-Critical Driving Video Understanding Unresolved cited work

Reference 28

Resolution
unresolved
raw_fallback, observed 2026-08-16T11:50:21.859479Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-16T11:50:21.217638Z digest=sha256:b2532158ad327a392afc7512f117617debc9d9850bf0a60a3d7c4e303fb76d08

Observation 29ad3f00-ec4f-497a-b214-eff5ad8407e8 · outbound

This paper cites ScVLM: Enhancing Vision-Language Model for Safety-Critical Event Understanding.

Are Vision LLMs Road-Ready? A Comprehensive Benchmark for Safety-Critical Driving Video Understanding ScVLM: Enhancing Vision-Language Model for Safety-Critical Event Understanding

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-16T11:50:21.222523Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:50:21.222523Z digest=sha256:2676e84f406835cacbee228ec5fa64bf721b234efeadfc4af7416468e9df6c96

Observation bd6c9737-588b-4cae-8388-394704b0a154 · outbound

This paper cites an unresolved cited work.

Are Vision LLMs Road-Ready? A Comprehensive Benchmark for Safety-Critical Driving Video Understanding Unresolved cited work

Reference 30

Resolution
unresolved
raw_fallback, observed 2026-08-16T11:50:21.843920Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-16T11:50:21.228194Z digest=sha256:acb0cf3962c33bcda8ff7df0831f153f8f8343240d17ef81bc562704798d84e3

Observation 12ceaa1e-73b0-4313-a09f-4f7d33227a33 · outbound

This paper cites an unresolved cited work.

Are Vision LLMs Road-Ready? A Comprehensive Benchmark for Safety-Critical Driving Video Understanding Unresolved cited work

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-16T11:50:21.237732Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:50:21.237732Z digest=sha256:a9b3576f68358efae29bad92b1f0bb9a50a6d7c58943a0223ccc9c0e79a5d3bb

Observation 3aacccda-5d94-4f2a-bed8-26154584aa6f · outbound

This paper cites In Proceedings of the AAAI Conference on Artificial Intelligence, Vol.

Are Vision LLMs Road-Ready? A Comprehensive Benchmark for Safety-Critical Driving Video Understanding In Proceedings of the AAAI Conference on Artificial Intelligence, Vol

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-16T11:50:21.212433Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:50:21.212433Z digest=sha256:59eee62f6ab6779be8910275f2c54731ec77093f12aedf3d324ae9cbcd7b5426

Observation 423a0837-c79e-4003-8feb-d5ac56dbe78d · outbound

This paper cites Gemini: A Family of Highly Capable Multimodal Models.

Are Vision LLMs Road-Ready? A Comprehensive Benchmark for Safety-Critical Driving Video Understanding Gemini: A Family of Highly Capable Multimodal Models

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-16T11:50:21.251129Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:50:21.251129Z digest=sha256:5fe98e4b5688cfe79eac7c346e269f2cf34b50a3671da3ab2fa7ba0671cb2633

Observation 258f011c-3b70-4bba-bafa-36cbb26730c1 · outbound

This paper cites an unresolved cited work.

Are Vision LLMs Road-Ready? A Comprehensive Benchmark for Safety-Critical Driving Video Understanding Unresolved cited work

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-16T11:50:21.255572Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:50:21.255572Z digest=sha256:059450f937a5284da9779096045f3109f4a99dcbad3c6b7cb20ac3fa73e36e42

Observation 728aef94-c7ab-4896-8883-d0fa7af15d70 · outbound

This paper cites DriveVLM: The Convergence of Autonomous Driving and Large Vision-Language Models.

Are Vision LLMs Road-Ready? A Comprehensive Benchmark for Safety-Critical Driving Video Understanding DriveVLM: The Convergence of Autonomous Driving and Large Vision-Language Models

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-16T11:50:21.260002Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:50:21.260002Z digest=sha256:811b696c68c988d47b449c4f6ca92ddc75a0b60c231f4becffbe2193b9d75c28

Observation 20b8710b-cea6-451b-a6c8-c96d8e532a0b · outbound

This paper cites In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) Workshops.

Are Vision LLMs Road-Ready? A Comprehensive Benchmark for Safety-Critical Driving Video Understanding In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) Workshops

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:50:21.827598Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-16T11:50:21.233010Z digest=sha256:23623deae6dc2c7494b6f74329c367cb461295d12106a568efb3274dc04a9b8d

Observation 1cbe1b67-25de-460f-af97-7f50edcb12c3 · outbound

This paper cites Are VLMs Ready for Autonomous Driving? An Empirical Study from the Reliability, Data, and Metric Perspectives.

Are Vision LLMs Road-Ready? A Comprehensive Benchmark for Safety-Critical Driving Video Understanding Are VLMs Ready for Autonomous Driving? An Empirical Study from the Reliability, Data, and Metric Perspectives

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-16T11:50:21.269233Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:50:21.269233Z digest=sha256:e2a88dd98646f7b80389af90100fda818bf5b558cb5bf33d054f5fa5375fb1b1

Observation 399d1e6a-dfbb-4879-9811-c444d98fb393 · outbound

This paper cites an unresolved cited work.

Are Vision LLMs Road-Ready? A Comprehensive Benchmark for Safety-Critical Driving Video Understanding Unresolved cited work

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-16T11:50:21.242312Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:50:21.242312Z digest=sha256:6710c60a9644693b109a666b7f1e3bb7530eec52c933e1a74ac9f1f210c38307

Observation 46bfa54d-d676-4a71-a0c9-a9efd171da38 · outbound

This paper cites an unresolved cited work.

Are Vision LLMs Road-Ready? A Comprehensive Benchmark for Safety-Critical Driving Video Understanding Unresolved cited work

Reference 39

Resolution
unresolved
raw_fallback, observed 2026-08-16T11:50:21.773409Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-16T11:50:21.278909Z digest=sha256:ab7d9a821df20dd17ade93f5ad2122faad2410fd94b23ea4eb1db058604f9f7e

Observation 1bf9c03c-0801-4237-8fc6-69acd905e9d6 · outbound

This paper cites an unresolved cited work.

Are Vision LLMs Road-Ready? A Comprehensive Benchmark for Safety-Critical Driving Video Understanding Unresolved cited work

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-16T11:50:21.283329Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:50:21.283329Z digest=sha256:d8bc580b4589b6ac06b141b779e42c519df152e1de05cb7c3d8561d4114e045f

Observation 9b208ea8-f334-479d-9498-9519c5278e62 · outbound

This paper cites MiniCPM-V: A GPT-4V Level MLLM on Your Phone.

Are Vision LLMs Road-Ready? A Comprehensive Benchmark for Safety-Critical Driving Video Understanding MiniCPM-V: A GPT-4V Level MLLM on Your Phone

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-16T11:50:21.287862Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:50:21.287862Z digest=sha256:4a21d41ab212eaa2060e30843989a6839c4d1ddabfd8d29f9fc7406a701d89bb

Observation d0debff8-feb8-4294-9e58-493c5523d9dc · outbound

This paper cites an unresolved cited work.

Are Vision LLMs Road-Ready? A Comprehensive Benchmark for Safety-Critical Driving Video Understanding Unresolved cited work

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-16T11:50:21.292850Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:50:21.292850Z digest=sha256:4d5a9f35a01c8efa7a930904337921261117d37b480b1fd341fc405ab4c0204f

Observation 6bb1e275-80c8-43ad-801e-930cece18633 · outbound

This paper cites Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution.

Are Vision LLMs Road-Ready? A Comprehensive Benchmark for Safety-Critical Driving Video Understanding Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-16T11:50:21.264821Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:50:21.264821Z digest=sha256:4f0f9d5552dab5d5d04de77db89e957f64a0f798bbf039b62b9fa825876ff4b3

Observation aec8d437-956a-4041-b98e-51004c149129 · outbound

This paper cites MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models.

Are Vision LLMs Road-Ready? A Comprehensive Benchmark for Safety-Critical Driving Video Understanding MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-16T11:50:21.302312Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:50:21.302312Z digest=sha256:1c231e4ad64b6fa68537a68405f229a3b83ea6b0848515cb9feb9c4ef505ce30

Observation b2e7bac9-354d-4cf2-8c53-c9b8a05aee9f · outbound

This paper cites PLLaVA : Parameter-free LLaVA Extension from Images to Videos for Video Dense Captioning.

Are Vision LLMs Road-Ready? A Comprehensive Benchmark for Safety-Critical Driving Video Understanding PLLaVA : Parameter-free LLaVA Extension from Images to Videos for Video Dense Captioning

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-16T11:50:21.274138Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:50:21.274138Z digest=sha256:a8efb9a635a765e09b657500925840824250ac7622a992aeccec44ba163088b7

Observation 84fd58ba-c22a-49e6-ba04-a9808b099ff5 · outbound

This paper cites an unresolved cited work.

Are Vision LLMs Road-Ready? A Comprehensive Benchmark for Safety-Critical Driving Video Understanding Unresolved cited work

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-16T11:50:21.297447Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:50:21.297447Z digest=sha256:239cd909545cc578143a0510c17ac3adddb300f10d7d0d79989628ba782adb44

Observation 11d46e1e-3005-4cab-8092-a4b3dc2c9724 · outbound

This paper cites Snowy",.

Are Vision LLMs Road-Ready? A Comprehensive Benchmark for Safety-Critical Driving Video Understanding Snowy",

Reference 52

Resolution
malformed identifier
raw_fallback, observed 2026-08-16T11:50:21.728095Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-16T11:50:21.307277Z digest=sha256:ae011f97d3560c8dab132ec6ad5e688ce207433a29488383b7a6265fdd6728a9

Observation baf99355-0365-4945-b44d-489a25c5b63e · outbound

This paper cites an unresolved cited work.

Are Vision LLMs Road-Ready? A Comprehensive Benchmark for Safety-Critical Driving Video Understanding Unresolved cited work

Reference 53

Resolution
unresolved
raw_fallback, observed 2026-08-16T11:50:21.712740Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-16T11:50:21.312234Z digest=sha256:62b277a8bea8edae300eac6925641e5f623030e39736dadce3e6465c875e7dcb

Observation 10fc2175-c211-420e-b5c6-0af16a96c153 · outbound

This paper cites an unresolved cited work.

Are Vision LLMs Road-Ready? A Comprehensive Benchmark for Safety-Critical Driving Video Understanding Unresolved cited work

Reference 54

Resolution
unresolved
raw_fallback, observed 2026-08-16T11:50:21.697986Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-16T11:50:21.316612Z digest=sha256:2c7948bc08026ffe1aff7dae2d8554c081cb541c19dd69efab0908a63da50384

Observation 6ee2577f-b130-49c3-b2cb-7d0cc741fddb · outbound

This paper cites Based on your observations, select the best option that accurately addresses the question.

Are Vision LLMs Road-Ready? A Comprehensive Benchmark for Safety-Critical Driving Video Understanding Based on your observations, select the best option that accurately addresses the question

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:50:21.682265Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-16T11:50:21.321263Z digest=sha256:649a5442b3139f3e06af7b99747813bd7e7db0e7554913e64fffa5ff1e78d0ee

Observation df19bcca-4080-42d7-8b51-51b383fcd735 · outbound

This paper cites InProceedings of the European conference on computer vision (ECCV).

Are Vision LLMs Road-Ready? A Comprehensive Benchmark for Safety-Critical Driving Video Understanding InProceedings of the European conference on computer vision (ECCV)

Reference 2018

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:50:22.035710Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-16T11:50:21.132208Z digest=sha256:f1766811621a50e96f2eeb869c2564f47f0881da88e3b7e19a9f9dbde08a9b8a

Observation b60d960a-8742-41da-878b-aa7f32220c49 · outbound

This paper cites In Proceedings of the IEEE/CVF conference on computer vision and pattern recognition.

Are Vision LLMs Road-Ready? A Comprehensive Benchmark for Safety-Critical Driving Video Understanding In Proceedings of the IEEE/CVF conference on computer vision and pattern recognition

Reference 2020

Resolution
unresolved
no resolver link, observed 2026-08-16T11:50:21.246875Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:50:21.246875Z digest=sha256:a363260a744bd9d0afdf9fcacc78cbd641519d0d12555ede2a96d85befd334ee

Observation 62d7753c-bb2d-4f1e-8a2d-1a0b7258c692 · outbound

This paper cites Advances in neural information processing systems 35 (2022), 23716–23736.

Are Vision LLMs Road-Ready? A Comprehensive Benchmark for Safety-Critical Driving Video Understanding Advances in neural information processing systems 35 (2022), 23716–23736

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-16T11:50:21.067584Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:50:21.067584Z digest=sha256:7f16cd90541c5bc9a8fd2662aff4bb4ba7312ea43e5f47f9960b85dc326d78e9

Observation 2ec1f9d4-8afe-49fc-8a9a-fd260a551f50 · outbound

This paper cites Video-ChatGPT: Towards Detailed Video Understanding via Large Vision and Language Models.

Are Vision LLMs Road-Ready? A Comprehensive Benchmark for Safety-Critical Driving Video Understanding Video-ChatGPT: Towards Detailed Video Understanding via Large Vision and Language Models

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-16T11:50:21.182984Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:50:21.182984Z digest=sha256:e1af4c9f842fc236c3471c7fa57d2458e6edbf770a20dafaa5bfb1232b01d9ec

Observation 627a32e6-d007-46c2-8500-04642a749afd · outbound

This paper cites Can LVLMs Obtain a Driver's License? A Benchmark Towards Reliable AGI for Autonomous Driving.

Are Vision LLMs Road-Ready? A Comprehensive Benchmark for Safety-Critical Driving Video Understanding Can LVLMs Obtain a Driver's License? A Benchmark Towards Reliable AGI for Autonomous Driving

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-16T11:50:21.172408Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:50:21.172408Z digest=sha256:53492732916eb299f32e1ec1ed4ea0cd21e6087247f621a7f7016f9d9565cb83

Pith citing papers

Observation 292ac349-af25-448f-a8da-beacad958f26 · inbound

STSBench: A Spatio-temporal Scenario Benchmark for Multi-modal Large Language Models in Autonomous Driving cites this paper.

STSBench: A Spatio-temporal Scenario Benchmark for Multi-modal Large Language Models in Autonomous Driving Are Vision LLMs Road-Ready? A Comprehensive Benchmark for Safety-Critical Driving Video Understanding

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-07T06:02:13.157665Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T06:02:13.157665Z digest=sha256:2b0e0077cc39312a3c4180e4777dc827853b5b0d3442e8f407dfb42cf601040f

Observation d046a268-4aa1-4bf1-9be0-2dd04892bb94 · inbound

NuRisk: A Visual Question Answering Dataset for Agent-Level Risk Assessment in Autonomous Driving cites this paper.

NuRisk: A Visual Question Answering Dataset for Agent-Level Risk Assessment in Autonomous Driving Are Vision LLMs Road-Ready? A Comprehensive Benchmark for Safety-Critical Driving Video Understanding

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-05-18T12:46:24.147180Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-18T12:44:24.574082Z digest=sha256:bae33186dc805d81252b9b4da06c4fc252b827f410f04d6138865179b4eba038