Pith. sign in

Paper Citation Record · LEDGER

Are Vision LLMs Road-Ready? A Comprehensive Benchmark for Safety-Critical Driving Video Understanding

As of 17 August 2026, this Paper Citation Record lists 55 of 55 outbound references and 2 inbound Pith citation observations for arXiv:2504.14526.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2504.14526 v1

Coverage vector

measured 55 of 55 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-16T11:50:21.321263Z

measured 57 of 57 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T06:02:13.157665Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-18T12:46:24.143801Z

Reference resolution

55 of 55 outbound references displayed

  • verified exact0
  • verified fuzzy3
  • unresolved51
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 2ea296ee-e7ac-4d61-8643-4d619357ae92 · outbound

This paper cites an unresolved cited work.

Are Vision LLMs Road-Ready? A Comprehensive Benchmark for Safety-Critical Driving Video Understanding Unresolved cited work

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-16T11:50:21.062326Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:50:21.062326Z digest=sha256:e7ceb77ea967ee273251839962facba6fc96d1315ba401c98ddde4287a1bf319

Observation f66d3cf6-0bc1-4116-af42-9b2a149aae8d · outbound

This paper cites an unresolved cited work.

Are Vision LLMs Road-Ready? A Comprehensive Benchmark for Safety-Critical Driving Video Understanding Unresolved cited work

Reference 2

Resolution
unresolved
raw_fallback, observed 2026-08-16T11:50:22.118913Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T11:50:21.072515Z digest=sha256:4ba8a343eb45d6df5050a0e7aab6fef826e059528ea14ae64d3af04f035382a9

Observation bcf0b43e-e3a6-4df4-9799-e14cb1c8e6f7 · outbound

This paper cites Premise Order Matters in Reasoning with Large Language Models.

Are Vision LLMs Road-Ready? A Comprehensive Benchmark for Safety-Critical Driving Video Understanding Premise Order Matters in Reasoning with Large Language Models

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-16T11:50:21.077215Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:50:21.077215Z digest=sha256:34cc650b0238699144ec2fba199055feb8ed18754f0068448572eb976b21f022

Observation ee6872ff-2327-451f-81ac-5eb17e0a6661 · outbound

This paper cites Talk2Car: Taking Control of Your Self-Driving Car.

Are Vision LLMs Road-Ready? A Comprehensive Benchmark for Safety-Critical Driving Video Understanding Talk2Car: Taking Control of Your Self-Driving Car

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-16T11:50:21.081969Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:50:21.081969Z digest=sha256:5f7e8529b7735576b355941b91ae14982bc0c846cb0ab62df1dccf2447da38f4

Observation 7387aec3-7d21-4970-96cd-9d1749843adc · outbound

This paper cites Video-MME: The First-Ever Comprehensive Evaluation Benchmark of Multi-modal LLMs in Video Analysis.

Are Vision LLMs Road-Ready? A Comprehensive Benchmark for Safety-Critical Driving Video Understanding Video-MME: The First-Ever Comprehensive Evaluation Benchmark of Multi-modal LLMs in Video Analysis

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-16T11:50:21.086946Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:50:21.086946Z digest=sha256:1df36d7f9215e2cf5657402e672dc667da68a5b1330b093cd3ce89b3d5a1400a

Observation b036f884-c800-4ff1-bb27-317a9569528f · outbound

This paper cites an unresolved cited work.

Are Vision LLMs Road-Ready? A Comprehensive Benchmark for Safety-Critical Driving Video Understanding Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-08-16T11:50:22.103995Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T11:50:21.092257Z digest=sha256:19603ad247974e79d0f6afe79fcfd778771c43b7ebe1ce6e3e8b4f824891bd16

Observation 7765deb5-3c29-4c7c-8101-45f8ece9636b · outbound

This paper cites SURDS: Benchmarking Spatial Understanding and Reasoning in Driving Scenarios with Vision Language Models.

Are Vision LLMs Road-Ready? A Comprehensive Benchmark for Safety-Critical Driving Video Understanding SURDS: Benchmarking Spatial Understanding and Reasoning in Driving Scenarios with Vision Language Models

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-16T11:50:21.097737Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:50:21.097737Z digest=sha256:95fa9baefe520f5484daf14ac2db4136b224daf532e7bc05b4082ebc4d50b276

Observation 9b4aa964-c050-48f1-96ad-77371f23de0a · outbound

This paper cites DME-Driver: Integrating Human Decision Logic and 3D Scene Perception in Autonomous Driving.

Are Vision LLMs Road-Ready? A Comprehensive Benchmark for Safety-Critical Driving Video Understanding DME-Driver: Integrating Human Decision Logic and 3D Scene Perception in Autonomous Driving

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-16T11:50:21.102980Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:50:21.102980Z digest=sha256:2054b97c6da447c7d9f2837aaad2a6771714c93c3817edbd368609798971487a

Observation 8d008a8a-f91b-4862-92a6-6dcbcc9d813d · outbound

This paper cites an unresolved cited work.

Are Vision LLMs Road-Ready? A Comprehensive Benchmark for Safety-Critical Driving Video Understanding Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-08-16T11:50:22.087520Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T11:50:21.107966Z digest=sha256:4ab50981fe5d6ced896afa5cd5417ce27865fbf213a2ddd8fd3af720fa532235

Observation c51e2dcc-bb6a-40de-9771-7d4d15820722 · outbound

This paper cites GPT-4o System Card.

Are Vision LLMs Road-Ready? A Comprehensive Benchmark for Safety-Critical Driving Video Understanding GPT-4o System Card

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-16T11:50:21.112624Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:50:21.112624Z digest=sha256:620299b77e7e90986502adf81abf7a34a25fae95a06ac1832c4e463337ca1191

Observation cb26d1d4-f317-4f39-a4d2-43a1dea79e3f · outbound

This paper cites an unresolved cited work.

Are Vision LLMs Road-Ready? A Comprehensive Benchmark for Safety-Critical Driving Video Understanding Unresolved cited work

Reference 11

Resolution
unresolved
raw_fallback, observed 2026-08-16T11:50:22.071361Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T11:50:21.117347Z digest=sha256:e56575b2539f8c4aa615d853f187237e6fb84aec1c8e038a0f8aa446201e4794

Observation 461ae905-99e7-44bc-953c-14cf6fa76900 · outbound

This paper cites Challenges and Applications of Large Language Models.

Are Vision LLMs Road-Ready? A Comprehensive Benchmark for Safety-Critical Driving Video Understanding Challenges and Applications of Large Language Models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-16T11:50:21.122481Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:50:21.122481Z digest=sha256:e91d13fac7265ac4b8f3f35253cb108c2ad495e0f4d2c2c898f94fa94178af0e

Observation 453fb0e5-0ac4-4c38-84a0-fe51e867bfa2 · outbound

This paper cites an unresolved cited work.

Are Vision LLMs Road-Ready? A Comprehensive Benchmark for Safety-Critical Driving Video Understanding Unresolved cited work

Reference 13

Resolution
unresolved
raw_fallback, observed 2026-08-16T11:50:22.053425Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T11:50:21.127534Z digest=sha256:931f655bec3067fbf4711c10b21d6e284606c2de063c3a72b52c5d1610c40d81

Observation 112366af-6e59-481b-a275-53564d48e8cf · outbound

This paper cites LLaVA-OneVision: Easy Visual Task Transfer.

Are Vision LLMs Road-Ready? A Comprehensive Benchmark for Safety-Critical Driving Video Understanding LLaVA-OneVision: Easy Visual Task Transfer

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-16T11:50:21.136424Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:50:21.136424Z digest=sha256:9785d40606c2f67537a3eebc6a254fb1020d260c35a6e2cde695426a15307b5c

Observation 242d2273-e193-4161-99ac-059bca802ec9 · outbound

This paper cites an unresolved cited work.

Are Vision LLMs Road-Ready? A Comprehensive Benchmark for Safety-Critical Driving Video Understanding Unresolved cited work

Reference 15

Resolution
unresolved
raw_fallback, observed 2026-08-16T11:50:22.019082Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T11:50:21.140669Z digest=sha256:c8018446722c6b9a8010c67b387eddec1bb3c544d8bdc95840a1f2071833ff52

Observation 7b36ac52-d44a-4276-9c10-2f291f8fdab5 · outbound

This paper cites an unresolved cited work.

Are Vision LLMs Road-Ready? A Comprehensive Benchmark for Safety-Critical Driving Video Understanding Unresolved cited work

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-16T11:50:21.144962Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:50:21.144962Z digest=sha256:5a0fbd8a6f1ad69cf3e2843a6faff181507d88b994453a542f455180d602652e

Observation 51db16c5-eadf-46d1-8e36-6468982aceea · outbound

This paper cites Video-LLaVA: Learning United Visual Representation by Alignment Before Projection.

Are Vision LLMs Road-Ready? A Comprehensive Benchmark for Safety-Critical Driving Video Understanding Video-LLaVA: Learning United Visual Representation by Alignment Before Projection

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-16T11:50:21.149521Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:50:21.149521Z digest=sha256:5bede1f174b0abe2afc2eaa24ce25a3a461742ead2adb365b22d94658ab50507

Observation df2953b1-b1b7-424a-9dd1-1bd03f30e32f · outbound

This paper cites an unresolved cited work.

Are Vision LLMs Road-Ready? A Comprehensive Benchmark for Safety-Critical Driving Video Understanding Unresolved cited work

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-16T11:50:21.154031Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:50:21.154031Z digest=sha256:735856fbc1f139d9b85f41d8e592dd2ea708458cdd3ca556af7b19cf0f23a1fa

Observation 39200787-f35b-4cc3-a6a8-c66f4a622fe6 · outbound

This paper cites an unresolved cited work.

Are Vision LLMs Road-Ready? A Comprehensive Benchmark for Safety-Critical Driving Video Understanding Unresolved cited work

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-16T11:50:21.158613Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:50:21.158613Z digest=sha256:461bc041cde874f5bd86f870071b386053863f03abeaf52db27e4ede82cbdfef

Observation 9a17b74c-994b-49cb-9b0e-27cd0f9d8ba9 · outbound

This paper cites an unresolved cited work.

Are Vision LLMs Road-Ready? A Comprehensive Benchmark for Safety-Critical Driving Video Understanding Unresolved cited work

Reference 20

Resolution
unresolved
raw_fallback, observed 2026-08-16T11:50:21.967414Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T11:50:21.162970Z digest=sha256:a433bbccb51b1d6004385f20295e4de32622a6b25ad61a32a1b30e205a92620f

Observation d11362cc-da97-42be-b3a1-cfed466fab6b · outbound

This paper cites an unresolved cited work.

Are Vision LLMs Road-Ready? A Comprehensive Benchmark for Safety-Critical Driving Video Understanding Unresolved cited work

Reference 21

Resolution
unresolved
raw_fallback, observed 2026-08-16T11:50:21.950871Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T11:50:21.167517Z digest=sha256:6f219773f4d3b404df0c7758fe1c35c7ea1e9145999f25d7cede2a658da7cde6

Observation d12daf5d-6636-4d67-8691-1631ebd45611 · outbound

This paper cites an unresolved cited work.

Are Vision LLMs Road-Ready? A Comprehensive Benchmark for Safety-Critical Driving Video Understanding Unresolved cited work

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-16T11:50:21.177509Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:50:21.177509Z digest=sha256:fa206ed311d7703f45fcb509baba0cd0f6c0661f28b41e9889047b5d3ea29a65

Observation 7608bcac-a179-45b7-8ba5-c05786e71591 · outbound

This paper cites an unresolved cited work.

Are Vision LLMs Road-Ready? A Comprehensive Benchmark for Safety-Critical Driving Video Understanding Unresolved cited work

Reference 23

Resolution
unresolved
raw_fallback, observed 2026-08-16T11:50:21.924023Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T11:50:21.188090Z digest=sha256:7ca07622fb1f867f93dc616b5d60ecd62d4e9d45916362f5912e38aea89da806

Observation 739891be-773f-41cc-ad25-0980f8766019 · outbound

This paper cites GPT-Driver: Learning to Drive with GPT.

Are Vision LLMs Road-Ready? A Comprehensive Benchmark for Safety-Critical Driving Video Understanding GPT-Driver: Learning to Drive with GPT

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-16T11:50:21.192672Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:50:21.192672Z digest=sha256:eb3dc4a1d8194ee81409b737483666e2f43271c21fc64d04f0fd2ab3aaa48e36

Observation 8fd4f3a1-1fb5-4524-8983-0a9288234a31 · outbound

This paper cites an unresolved cited work.

Are Vision LLMs Road-Ready? A Comprehensive Benchmark for Safety-Critical Driving Video Understanding Unresolved cited work

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-16T11:50:21.197489Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:50:21.197489Z digest=sha256:7606ce76e51b613987499bbb2deb823f9e262fdc3f64b36c60cec300a4d3787f

Observation a195190f-d5c3-46e6-98cc-bf574df4c644 · outbound

This paper cites an unresolved cited work.

Are Vision LLMs Road-Ready? A Comprehensive Benchmark for Safety-Critical Driving Video Understanding Unresolved cited work

Reference 26

Resolution
unresolved
raw_fallback, observed 2026-08-16T11:50:21.896628Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T11:50:21.202293Z digest=sha256:a89469c3fa96b85e97739c84bac794743c7cedce461877bae6b838693caa4ed2

Observation f14e36e6-bda0-4f69-92de-85bcf54ee430 · outbound

This paper cites an unresolved cited work.

Are Vision LLMs Road-Ready? A Comprehensive Benchmark for Safety-Critical Driving Video Understanding Unresolved cited work

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-16T11:50:21.207515Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:50:21.207515Z digest=sha256:da8f51f1713b66a561a07f624b938f7a708e18bdaf785236b44072be77c9b5ed

Observation fc02c19b-75cb-4f11-92d7-876c0066a1c4 · outbound

This paper cites an unresolved cited work.

Are Vision LLMs Road-Ready? A Comprehensive Benchmark for Safety-Critical Driving Video Understanding Unresolved cited work

Reference 28

Resolution
unresolved
raw_fallback, observed 2026-08-16T11:50:21.859479Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T11:50:21.217638Z digest=sha256:f458e296b0cc73de0daf218052dbc0835302244ac3680e90778fccbb4fb2d54a

Observation 29ad3f00-ec4f-497a-b214-eff5ad8407e8 · outbound

This paper cites ScVLM: Enhancing Vision-Language Model for Safety-Critical Event Understanding.

Are Vision LLMs Road-Ready? A Comprehensive Benchmark for Safety-Critical Driving Video Understanding ScVLM: Enhancing Vision-Language Model for Safety-Critical Event Understanding

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-16T11:50:21.222523Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:50:21.222523Z digest=sha256:6f9f3b0237bdab999e60f342f2a01d60f70f618a8522de1573a1342792d471cc

Observation bd6c9737-588b-4cae-8388-394704b0a154 · outbound

This paper cites an unresolved cited work.

Are Vision LLMs Road-Ready? A Comprehensive Benchmark for Safety-Critical Driving Video Understanding Unresolved cited work

Reference 30

Resolution
unresolved
raw_fallback, observed 2026-08-16T11:50:21.843920Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T11:50:21.228194Z digest=sha256:935dcf0f3e59b707df9b852ac221f682f275b66e36b07c6334fbedb70ba1fc0f

Observation 12ceaa1e-73b0-4313-a09f-4f7d33227a33 · outbound

This paper cites an unresolved cited work.

Are Vision LLMs Road-Ready? A Comprehensive Benchmark for Safety-Critical Driving Video Understanding Unresolved cited work

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-16T11:50:21.237732Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:50:21.237732Z digest=sha256:a1ae0bef0b42c4d945633ce07dbb0943ad45f329a2faccab26c4ce66b94bf995

Observation 3aacccda-5d94-4f2a-bed8-26154584aa6f · outbound

This paper cites In Proceedings of the AAAI Conference on Artificial Intelligence, Vol.

Are Vision LLMs Road-Ready? A Comprehensive Benchmark for Safety-Critical Driving Video Understanding In Proceedings of the AAAI Conference on Artificial Intelligence, Vol

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-16T11:50:21.212433Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:50:21.212433Z digest=sha256:b670143f77403c592f575e2f2b6ffd87e2cb76ff6e4d5701bf5ba041795a4668

Observation 423a0837-c79e-4003-8feb-d5ac56dbe78d · outbound

This paper cites Gemini: A Family of Highly Capable Multimodal Models.

Are Vision LLMs Road-Ready? A Comprehensive Benchmark for Safety-Critical Driving Video Understanding Gemini: A Family of Highly Capable Multimodal Models

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-16T11:50:21.251129Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:50:21.251129Z digest=sha256:0da841ec588b5af62c4cf0c1e19cf88e65a81aff94ca9744cc0d6ae802ecc954

Observation 258f011c-3b70-4bba-bafa-36cbb26730c1 · outbound

This paper cites an unresolved cited work.

Are Vision LLMs Road-Ready? A Comprehensive Benchmark for Safety-Critical Driving Video Understanding Unresolved cited work

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-16T11:50:21.255572Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:50:21.255572Z digest=sha256:e76ee2652d202fa97bac0f0a9b0e78c3bb0f19967f3fa3833cf16028185f477c

Observation 728aef94-c7ab-4896-8883-d0fa7af15d70 · outbound

This paper cites DriveVLM: The Convergence of Autonomous Driving and Large Vision-Language Models.

Are Vision LLMs Road-Ready? A Comprehensive Benchmark for Safety-Critical Driving Video Understanding DriveVLM: The Convergence of Autonomous Driving and Large Vision-Language Models

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-16T11:50:21.260002Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:50:21.260002Z digest=sha256:481b7f82401aa319c0ba0ca28e10e89bd458e2ea172a20d9384ee24eb9a8e136

Observation 20b8710b-cea6-451b-a6c8-c96d8e532a0b · outbound

This paper cites In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) Workshops.

Are Vision LLMs Road-Ready? A Comprehensive Benchmark for Safety-Critical Driving Video Understanding In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) Workshops

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:50:21.827598Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T11:50:21.233010Z digest=sha256:15e74cceec4ac17d03e7869338fbc6bc6e5851aee05d84db30dd3c98ea98c688

Observation 1cbe1b67-25de-460f-af97-7f50edcb12c3 · outbound

This paper cites Are VLMs Ready for Autonomous Driving? An Empirical Study from the Reliability, Data, and Metric Perspectives.

Are Vision LLMs Road-Ready? A Comprehensive Benchmark for Safety-Critical Driving Video Understanding Are VLMs Ready for Autonomous Driving? An Empirical Study from the Reliability, Data, and Metric Perspectives

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-16T11:50:21.269233Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:50:21.269233Z digest=sha256:f703326fb7e920237915914ac9b3cbecee9d4816f6af95c28ddb484d87de70f4

Observation 399d1e6a-dfbb-4879-9811-c444d98fb393 · outbound

This paper cites an unresolved cited work.

Are Vision LLMs Road-Ready? A Comprehensive Benchmark for Safety-Critical Driving Video Understanding Unresolved cited work

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-16T11:50:21.242312Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:50:21.242312Z digest=sha256:7cb20751fba77cba5347eefdba8d410a3817b6bf0bf71e16a530649febbf6728

Observation 46bfa54d-d676-4a71-a0c9-a9efd171da38 · outbound

This paper cites an unresolved cited work.

Are Vision LLMs Road-Ready? A Comprehensive Benchmark for Safety-Critical Driving Video Understanding Unresolved cited work

Reference 39

Resolution
unresolved
raw_fallback, observed 2026-08-16T11:50:21.773409Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T11:50:21.278909Z digest=sha256:3d81e1a2818077fa1d580599e2c19812f5e6831a026bae8232a890ff55e071f9

Observation 1bf9c03c-0801-4237-8fc6-69acd905e9d6 · outbound

This paper cites an unresolved cited work.

Are Vision LLMs Road-Ready? A Comprehensive Benchmark for Safety-Critical Driving Video Understanding Unresolved cited work

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-16T11:50:21.283329Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:50:21.283329Z digest=sha256:6fc791a6e63a9b5f1c70a24fdc7917f5176802ef7691130c96779f97c0cfa6dd

Observation 9b208ea8-f334-479d-9498-9519c5278e62 · outbound

This paper cites MiniCPM-V: A GPT-4V Level MLLM on Your Phone.

Are Vision LLMs Road-Ready? A Comprehensive Benchmark for Safety-Critical Driving Video Understanding MiniCPM-V: A GPT-4V Level MLLM on Your Phone

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-16T11:50:21.287862Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:50:21.287862Z digest=sha256:8d8875ee1f9443b2e2ae1b17a4e8c1b755ecd600ea436cff63b06e66375847ec

Observation d0debff8-feb8-4294-9e58-493c5523d9dc · outbound

This paper cites an unresolved cited work.

Are Vision LLMs Road-Ready? A Comprehensive Benchmark for Safety-Critical Driving Video Understanding Unresolved cited work

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-16T11:50:21.292850Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:50:21.292850Z digest=sha256:ddd60b96ebbde38ba94daee3fc0ac746ed85f00aa1c7e0ccdd455a2e1a483494

Observation 6bb1e275-80c8-43ad-801e-930cece18633 · outbound

This paper cites Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution.

Are Vision LLMs Road-Ready? A Comprehensive Benchmark for Safety-Critical Driving Video Understanding Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-16T11:50:21.264821Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:50:21.264821Z digest=sha256:a31b8b586911b3ab0f550064419daab3d621046b1cfdc2b9e5d998457170ccf4

Observation aec8d437-956a-4041-b98e-51004c149129 · outbound

This paper cites MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models.

Are Vision LLMs Road-Ready? A Comprehensive Benchmark for Safety-Critical Driving Video Understanding MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-16T11:50:21.302312Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:50:21.302312Z digest=sha256:31d4f00fcd506eddd5161dbd9b5df640b31e7c7b0046d3da57934b185e0e2429

Observation b2e7bac9-354d-4cf2-8c53-c9b8a05aee9f · outbound

This paper cites PLLaVA : Parameter-free LLaVA Extension from Images to Videos for Video Dense Captioning.

Are Vision LLMs Road-Ready? A Comprehensive Benchmark for Safety-Critical Driving Video Understanding PLLaVA : Parameter-free LLaVA Extension from Images to Videos for Video Dense Captioning

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-16T11:50:21.274138Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:50:21.274138Z digest=sha256:1c0502fe02ee6005760369c266d7bb41f5b1444088399da0648d4ca76b7256dd

Observation 84fd58ba-c22a-49e6-ba04-a9808b099ff5 · outbound

This paper cites an unresolved cited work.

Are Vision LLMs Road-Ready? A Comprehensive Benchmark for Safety-Critical Driving Video Understanding Unresolved cited work

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-16T11:50:21.297447Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:50:21.297447Z digest=sha256:10004a2abf81de338a638471e75aef5e7bea09d2b88ce00422592ab4cce1fb36

Observation 11d46e1e-3005-4cab-8092-a4b3dc2c9724 · outbound

This paper cites Snowy",.

Are Vision LLMs Road-Ready? A Comprehensive Benchmark for Safety-Critical Driving Video Understanding Snowy",

Reference 52

Resolution
malformed identifier
raw_fallback, observed 2026-08-16T11:50:21.728095Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T11:50:21.307277Z digest=sha256:2d198e67d0ca311671277715cb4228f308e55d1fff683d1594875788b1579e7c

Observation baf99355-0365-4945-b44d-489a25c5b63e · outbound

This paper cites an unresolved cited work.

Are Vision LLMs Road-Ready? A Comprehensive Benchmark for Safety-Critical Driving Video Understanding Unresolved cited work

Reference 53

Resolution
unresolved
raw_fallback, observed 2026-08-16T11:50:21.712740Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T11:50:21.312234Z digest=sha256:bc69234136c43d868903637b8be085b695f8ebb10765440d0bc8a4e82ef4f7d4

Observation 10fc2175-c211-420e-b5c6-0af16a96c153 · outbound

This paper cites an unresolved cited work.

Are Vision LLMs Road-Ready? A Comprehensive Benchmark for Safety-Critical Driving Video Understanding Unresolved cited work

Reference 54

Resolution
unresolved
raw_fallback, observed 2026-08-16T11:50:21.697986Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T11:50:21.316612Z digest=sha256:65d373b8af3f337c6be0eab560836ed11e4f0b7fbd31ce9996bda3817aaac454

Observation 6ee2577f-b130-49c3-b2cb-7d0cc741fddb · outbound

This paper cites Based on your observations, select the best option that accurately addresses the question.

Are Vision LLMs Road-Ready? A Comprehensive Benchmark for Safety-Critical Driving Video Understanding Based on your observations, select the best option that accurately addresses the question

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:50:21.682265Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T11:50:21.321263Z digest=sha256:7e75a532d1eb721d3b6b2ee14b8137eb9ab9ffbd89ec2930ea7ea518391c98a7

Observation df19bcca-4080-42d7-8b51-51b383fcd735 · outbound

This paper cites InProceedings of the European conference on computer vision (ECCV).

Are Vision LLMs Road-Ready? A Comprehensive Benchmark for Safety-Critical Driving Video Understanding InProceedings of the European conference on computer vision (ECCV)

Reference 2018

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:50:22.035710Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T11:50:21.132208Z digest=sha256:49d1199fa255c437dce3c396920ccae231d07c8bdcb75fc0fd2aa096a4be6974

Observation b60d960a-8742-41da-878b-aa7f32220c49 · outbound

This paper cites In Proceedings of the IEEE/CVF conference on computer vision and pattern recognition.

Are Vision LLMs Road-Ready? A Comprehensive Benchmark for Safety-Critical Driving Video Understanding In Proceedings of the IEEE/CVF conference on computer vision and pattern recognition

Reference 2020

Resolution
unresolved
no resolver link, observed 2026-08-16T11:50:21.246875Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:50:21.246875Z digest=sha256:a6f6ba7deba1d79bc72f102d08186b55e616411400756fd91116f4e8843784c3

Observation 62d7753c-bb2d-4f1e-8a2d-1a0b7258c692 · outbound

This paper cites Advances in neural information processing systems 35 (2022), 23716–23736.

Are Vision LLMs Road-Ready? A Comprehensive Benchmark for Safety-Critical Driving Video Understanding Advances in neural information processing systems 35 (2022), 23716–23736

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-16T11:50:21.067584Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:50:21.067584Z digest=sha256:153555de34a36162195869db2843e89077333d39a88d31aad9b1fc5a0c3b180a

Observation 2ec1f9d4-8afe-49fc-8a9a-fd260a551f50 · outbound

This paper cites Video-ChatGPT: Towards Detailed Video Understanding via Large Vision and Language Models.

Are Vision LLMs Road-Ready? A Comprehensive Benchmark for Safety-Critical Driving Video Understanding Video-ChatGPT: Towards Detailed Video Understanding via Large Vision and Language Models

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-16T11:50:21.182984Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:50:21.182984Z digest=sha256:13c3869306f3a9a7bc8800171a04ef03c0bdba889bdd47d0811c4e5b34644493

Observation 627a32e6-d007-46c2-8500-04642a749afd · outbound

This paper cites Can LVLMs Obtain a Driver's License? A Benchmark Towards Reliable AGI for Autonomous Driving.

Are Vision LLMs Road-Ready? A Comprehensive Benchmark for Safety-Critical Driving Video Understanding Can LVLMs Obtain a Driver's License? A Benchmark Towards Reliable AGI for Autonomous Driving

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-16T11:50:21.172408Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:50:21.172408Z digest=sha256:10b1ced9c9c41527784b562160244c86e3965ede6ee6de9fdf9e2cc4cbd8d633

Pith citing papers

Observation 292ac349-af25-448f-a8da-beacad958f26 · inbound

STSBench: A Spatio-temporal Scenario Benchmark for Multi-modal Large Language Models in Autonomous Driving cites this paper.

STSBench: A Spatio-temporal Scenario Benchmark for Multi-modal Large Language Models in Autonomous Driving Are Vision LLMs Road-Ready? A Comprehensive Benchmark for Safety-Critical Driving Video Understanding

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-07T06:02:13.157665Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T06:02:13.157665Z digest=sha256:bcc4a2fe8f46cd42c934e96a318b8fff7a2ac4ef0a6bc73e0559b449b0d3f38b

Observation d046a268-4aa1-4bf1-9be0-2dd04892bb94 · inbound

NuRisk: A Visual Question Answering Dataset for Agent-Level Risk Assessment in Autonomous Driving cites this paper.

NuRisk: A Visual Question Answering Dataset for Agent-Level Risk Assessment in Autonomous Driving Are Vision LLMs Road-Ready? A Comprehensive Benchmark for Safety-Critical Driving Video Understanding

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-05-18T12:46:24.147180Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-18T12:44:24.574082Z digest=sha256:0a20bd06cf88c79996ede976197c4b22e6b8bce7683da7ee126f0b32ce8ace58