Pith. sign in

Paper Citation Record · LEDGER

CarLLaVA: Vision language models for camera-only closed-loop driving

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 18 inbound Pith citation observations for arXiv:2406.10165.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2406.10165 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 18 of 18 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 18 of 18 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:24:07.325451Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T16:48:39.563623Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 6756f4cd-a63c-44c5-91cd-c115d8e8f6f7 · inbound

ORION: A Holistic End-to-End Autonomous Driving Framework by Vision-Language Instructed Action Generation cites this paper.

ORION: A Holistic End-to-End Autonomous Driving Framework by Vision-Language Instructed Action Generation CarLLaVA: Vision language models for camera-only closed-loop driving

Reference 46

Resolution
verified exact
arxiv_id, observed 2026-05-17T08:10:39.392246Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-17T08:10:39.240168Z digest=sha256:2b61816575086fdfcd6ff182325ce2581bf2695b88352933b4cbb7a03f17ae6b

Observation 3b29bd27-3d8c-4f05-8e06-16e447ee5871 · inbound

Generative AI for Autonomous Driving: A Review cites this paper.

Generative AI for Autonomous Driving: A Review CarLLaVA: Vision language models for camera-only closed-loop driving

Reference 253

Resolution
unresolved
no resolver link, observed 2026-08-07T15:24:07.325451Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:24:07.325451Z digest=sha256:00a751ab3f794541566e036a2bab23899d974c29992ed87e26c5cdf45485ea8b

Observation 72e4307d-33d0-4567-ae10-668bb23074c8 · inbound

Generalized Trajectory Scoring for End-to-end Multimodal Planning cites this paper.

Generalized Trajectory Scoring for End-to-end Multimodal Planning CarLLaVA: Vision language models for camera-only closed-loop driving

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T05:59:18.570814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:59:18.570814Z digest=sha256:debb793e6eacaf315b77b20f56ce9fe6b4fa07ff454e07afbc0e62a134b802f1

Observation db19ec5a-0383-4d55-927e-91ddafaea5de · inbound

ETA: Efficiency through Thinking Ahead, A Dual Approach to Self-Driving with Large Models cites this paper.

ETA: Efficiency through Thinking Ahead, A Dual Approach to Self-Driving with Large Models CarLLaVA: Vision language models for camera-only closed-loop driving

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T05:32:32.675801Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:32:32.675801Z digest=sha256:41a4ee1dc784c5575029ac72ce045c865a87d9c6309bdc096d268bb98b232a07

Observation dfcea971-7b05-491e-ad11-895c84303b51 · inbound

AutoVLA: A Vision-Language-Action Model for End-to-End Autonomous Driving with Adaptive Reasoning and Reinforcement Fine-Tuning cites this paper.

AutoVLA: A Vision-Language-Action Model for End-to-End Autonomous Driving with Adaptive Reasoning and Reinforcement Fine-Tuning CarLLaVA: Vision language models for camera-only closed-loop driving

Reference 41

Resolution
verified exact
arxiv_id, observed 2026-05-14T21:46:44.177559Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-14T21:46:43.955825Z digest=sha256:8ece316697fafed4a71e563c5ab2b9ac8d39dc33b30104e079c16df9cc5db0ce

Observation 10e4463e-a45e-4816-8eb5-f213050b1545 · inbound

LaViPlan : Language-Guided Visual Path Planning with RLVR cites this paper.

LaViPlan : Language-Guided Visual Path Planning with RLVR CarLLaVA: Vision language models for camera-only closed-loop driving

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T16:40:04.319609Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:40:04.319609Z digest=sha256:478792658ecf95d9bb4a8f2a1d226bfca52a5ec28bb35cafaf6a8e9d04859ef2

Observation 9ec1907e-0dfe-4b4b-88c5-64f011d8ef7b · inbound

DriveQA: Passing the Driving Knowledge Test cites this paper.

DriveQA: Passing the Driving Knowledge Test CarLLaVA: Vision language models for camera-only closed-loop driving

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-05T13:58:10.219174Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T13:58:10.219174Z digest=sha256:2b7dc3e6abb2dc14104f51fff038905cb19bcc2fff9f8234d947efd974409b3a

Observation 5152c52a-e923-4061-bc27-7bf10ca16f53 · inbound

CogDriver: Integrating Cognitive Inertia for Temporally Coherent Planning in Autonomous Driving cites this paper.

CogDriver: Integrating Cognitive Inertia for Temporally Coherent Planning in Autonomous Driving CarLLaVA: Vision language models for camera-only closed-loop driving

Reference 32

Resolution
verified exact
arxiv_id, observed 2026-05-18T20:11:50.808741Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-18T20:10:43.416488Z digest=sha256:ed50bd414480251f9e74a759d1752df0ca05d64176a16c40ea9fa2b398ef683e

Observation 390ac7d7-d185-435a-bca2-de21f12fd849 · inbound

DriveVLA-W0: World Models Amplify Data Scaling Law in Autonomous Driving cites this paper.

DriveVLA-W0: World Models Amplify Data Scaling Law in Autonomous Driving CarLLaVA: Vision language models for camera-only closed-loop driving

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-05-17T06:48:01.087290Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-17T06:48:00.943591Z digest=sha256:2c22938cec0dd9c38ab3c09bfe337577a534e3cf50f71cbe6cd3b2300a5c5941

Observation 37984482-9bc7-4476-898c-5b83ac98484d · inbound

A Review of Learning-Based Motion Planning: Toward a Data-Driven Optimal Control Approach cites this paper.

A Review of Learning-Based Motion Planning: Toward a Data-Driven Optimal Control Approach CarLLaVA: Vision language models for camera-only closed-loop driving

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-03T16:52:55.423521Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:52:55.423521Z digest=sha256:71f0a4310d7a819dd19de856534a3821f0920f7db6789106eab33d155bc52888

Observation dbbaca0c-d952-4cf0-9b47-521a441d1a54 · inbound

AlignDrive: Aligned Lateral-Longitudinal Planning for End-to-End Autonomous Driving cites this paper.

AlignDrive: Aligned Lateral-Longitudinal Planning for End-to-End Autonomous Driving CarLLaVA: Vision language models for camera-only closed-loop driving

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-05-16T18:41:11.162098Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-16T18:39:49.416119Z digest=sha256:85e6b1752363e1569a780effc990a586814171935406e97f8a47ccd4c564e82a

Observation 71fddd0e-e487-4fac-85eb-8e36dc2e83ba · inbound

AlignDrive: Aligned Lateral-Longitudinal Planning for End-to-End Autonomous Driving cites this paper.

AlignDrive: Aligned Lateral-Longitudinal Planning for End-to-End Autonomous Driving CarLLaVA: Vision language models for camera-only closed-loop driving

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-03T12:48:35.811131Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T12:48:35.811131Z digest=sha256:54ac9fa15c0366ffcd29e81ff116c61a9f1ed6a243338d584f2170f4eabe8782

Observation 989695e6-72a6-4d0a-8ec7-4a54304ae37a · inbound

TaCarla: A comprehensive benchmarking dataset for end-to-end autonomous driving cites this paper.

TaCarla: A comprehensive benchmarking dataset for end-to-end autonomous driving CarLLaVA: Vision language models for camera-only closed-loop driving

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-02T20:21:31.075670Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T20:21:31.075670Z digest=sha256:95f8411b5ebcc392d824ce6a90e6f800bfa058df49260fd6a9d864e891558520

Observation 89353404-a066-42c3-8092-3df05a1a6272 · inbound

DVGT-2: Vision-Geometry-Action Model for Autonomous Driving at Scale cites this paper.

DVGT-2: Vision-Geometry-Action Model for Autonomous Driving at Scale CarLLaVA: Vision language models for camera-only closed-loop driving

Reference 59

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T23:03:25.120064Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-13T22:59:57.744015Z digest=sha256:a87658e8497a89c662d4cff21347eb6efbcbefaba0e9f2493140f0556447edb9

Observation 0c567676-f9eb-4692-ae88-df2d94edf05a · inbound

DeepSight: Long-Horizon World Modeling via Latent States Prediction for End-to-End Autonomous Driving cites this paper.

DeepSight: Long-Horizon World Modeling via Latent States Prediction for End-to-End Autonomous Driving CarLLaVA: Vision language models for camera-only closed-loop driving

Reference 53

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T06:31:25.869965Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-12T04:13:37.421188Z digest=sha256:1be5aad235cbd58d8a47d3e6ac5b642c1140654ae5eaab183637ebfbb4e6a12c

Observation 091db9d8-e3fb-4eb6-b07a-abe8e2cbf683 · inbound

LVDrive: Latent Visual Representation Enhanced Vision-Language-Action Autonomous Driving Model cites this paper.

LVDrive: Latent Visual Representation Enhanced Vision-Language-Action Autonomous Driving Model CarLLaVA: Vision language models for camera-only closed-loop driving

Reference 41

Resolution
verified exact
arxiv_id, observed 2026-05-22T07:41:15.037981Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-22T07:37:11.292270Z digest=sha256:31300c2c104de1169c30b444f055f100cea6bcc89648884c41a449d11550e855

Observation f7076c21-e939-4102-aeda-032225d7d17d · inbound

VLGA: Vision-Language-Geometry-Action Models for Autonomous Driving cites this paper.

VLGA: Vision-Language-Geometry-Action Models for Autonomous Driving CarLLaVA: Vision language models for camera-only closed-loop driving

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-07-03T11:08:03.416062Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-27T09:40:51.407286Z digest=sha256:f0bfa8f3eef375f3b9a0c5d57061b3597e3d9a2be66313f0b67664d3ae5e0f9d

Observation 4843126b-cb6e-4163-8a59-290a4e8d47e7 · inbound

Teaching Vision-Language-Action Models What to See and Where to Look cites this paper.

Teaching Vision-Language-Action Models What to See and Where to Look CarLLaVA: Vision language models for camera-only closed-loop driving

Reference 44

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T16:48:39.564964Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-07-03T16:42:13.520913Z digest=sha256:8ed67b8385f763a2685178dd5b9e5e21bdce31b963a7071348d82f4fe88c1543