Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-11T20:17:47.749250Z
Paper Citation Record · LEDGER
As of 17 August 2026, this Paper Citation Record lists 28 of 28 outbound references and 4 inbound Pith citation observations for arXiv:2412.05893.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-11T20:17:47.749250Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-16T00:00:17.006999Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-06-30T13:44:41.209962Z
28 of 28 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 9e71f20a-1c88-43c6-a8a8-9e03a2307219 · outbound
doScenes: An Autonomous Driving Dataset with Natural Language Instruction for Human Interaction and Vision-Language Navigation Looking-in and looking-out of a vehicle: Computer-vision-based enhanced vehicle safety,
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation fb6d8cc1-0852-4ebd-a802-33df7c62b4c4 · outbound
doScenes: An Autonomous Driving Dataset with Natural Language Instruction for Human Interaction and Vision-Language Navigation Causal diagrams for empirical research,
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8ddbe2b7-dd1a-4403-97fc-786c9586d908 · outbound
doScenes: An Autonomous Driving Dataset with Natural Language Instruction for Human Interaction and Vision-Language Navigation Natsgd: A dataset with speech, gestures, and demonstrations for robot learning in natural human-robot interaction,
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation f78aad19-6767-47ab-b032-07b85801418d · outbound
doScenes: An Autonomous Driving Dataset with Natural Language Instruction for Human Interaction and Vision-Language Navigation Bridgedata v2: A dataset for robot learning at scale,
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 208a8f1f-e07a-4ad0-90d7-e47eaca658b2 · outbound
doScenes: An Autonomous Driving Dataset with Natural Language Instruction for Human Interaction and Vision-Language Navigation Handmethat: Human-robot communication in physical and social environments,
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 8dc1a5e3-1cf9-456c-8eca-f633520ddb47 · outbound
doScenes: An Autonomous Driving Dataset with Natural Language Instruction for Human Interaction and Vision-Language Navigation nuscenes: A multimodal dataset for autonomous driving,
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation d61e66c7-747b-44df-b324-33f6732f3558 · outbound
doScenes: An Autonomous Driving Dataset with Natural Language Instruction for Human Interaction and Vision-Language Navigation Nuscenes-qa: A multi-modal visual question answering benchmark for autonomous driving scenario,
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 59d2b4a1-0430-47ca-b003-98e8a1b74ac4 · outbound
doScenes: An Autonomous Driving Dataset with Natural Language Instruction for Human Interaction and Vision-Language Navigation Nuscenes-mqa: Integrated evaluation of captions and qa for autonomous driving datasets using markup annotations,
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation ceebec12-22f9-40af-8ec0-21e3f73ed6a0 · outbound
doScenes: An Autonomous Driving Dataset with Natural Language Instruction for Human Interaction and Vision-Language Navigation Rank2tell: A multimodal driving dataset for joint importance ranking and reasoning,
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fd143e34-8bd7-47b3-9880-d6e0b1f3046a · outbound
doScenes: An Autonomous Driving Dataset with Natural Language Instruction for Human Interaction and Vision-Language Navigation GPT-Driver: Learning to Drive with GPT
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 429f3cee-5cb9-4742-8431-45217236a94b · outbound
doScenes: An Autonomous Driving Dataset with Natural Language Instruction for Human Interaction and Vision-Language Navigation Drivemlm: Aligning multi-modal large language models with behavioral planning states for autonomous driving,
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5481b524-6019-46a4-b535-4d7ef531a481 · outbound
doScenes: An Autonomous Driving Dataset with Natural Language Instruction for Human Interaction and Vision-Language Navigation Lmdrive: Closed-loop end-to-end driving with large language models,
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation b37bea59-7248-430e-8355-eb6f817f3f88 · outbound
doScenes: An Autonomous Driving Dataset with Natural Language Instruction for Human Interaction and Vision-Language Navigation Drivegpt4: Interpretable end-to-end autonomous driving via large language model,
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 52a0155d-f0ef-416a-9de8-6fa8a8877ad8 · outbound
doScenes: An Autonomous Driving Dataset with Natural Language Instruction for Human Interaction and Vision-Language Navigation Drama: Joint risk localization and captioning in driving,
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation d1bbb3ac-d33f-446e-b0d5-0e99919b911c · outbound
doScenes: An Autonomous Driving Dataset with Natural Language Instruction for Human Interaction and Vision-Language Navigation Towards Explainable, Safe Autonomous Driving with Language Embeddings for Novelty Identification and Active Learning: Framework and Experimental Analysis with Real-World Data Sets
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cae4fcc8-9365-47fe-b758-b4e48c0d7ec6 · outbound
doScenes: An Autonomous Driving Dataset with Natural Language Instruction for Human Interaction and Vision-Language Navigation Autonomous vehicles that alert humans to take-over controls: Modeling with real-world data,
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 4f4e354d-8cab-43c8-9952-586c624e2926 · outbound
doScenes: An Autonomous Driving Dataset with Natural Language Instruction for Human Interaction and Vision-Language Navigation Safe control transitions: Machine vision based observable readiness index and data-driven takeover time prediction,
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation ae5aff59-c9ee-4da2-89a2-d44c6744c37b · outbound
doScenes: An Autonomous Driving Dataset with Natural Language Instruction for Human Interaction and Vision-Language Navigation Predicting Take-over Time for Autonomous Driving with Real-World Data: Robust Data Augmentation, Models, and Evaluation
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 87032ec4-0de6-42c4-8414-021e309234ab · outbound
doScenes: An Autonomous Driving Dataset with Natural Language Instruction for Human Interaction and Vision-Language Navigation NaVILA: Legged Robot Vision-Language-Action Model for Navigation
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 121abb2c-1508-457f-b9ac-ddfc0330ce10 · outbound
doScenes: An Autonomous Driving Dataset with Natural Language Instruction for Human Interaction and Vision-Language Navigation Vision-and-language navigation: Interpreting visually-grounded navigation instructions in real environments,
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation fa4e0cd9-0f35-4208-a9be-25385e215704 · outbound
doScenes: An Autonomous Driving Dataset with Natural Language Instruction for Human Interaction and Vision-Language Navigation Towards learning a generalist model for embodied navigation,
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 82843183-42c9-4690-9bdf-17478fa71b0a · outbound
doScenes: An Autonomous Driving Dataset with Natural Language Instruction for Human Interaction and Vision-Language Navigation Lm-nav: Robotic navigation with large pre-trained models of language, vision, and action,
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 0e9708ab-e70d-403f-bcdb-86f0366a59a4 · outbound
doScenes: An Autonomous Driving Dataset with Natural Language Instruction for Human Interaction and Vision-Language Navigation Grounding language to natural human-robot interaction in robot navigation tasks,
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 75edb90a-9305-4559-97c6-2b63e08cacf6 · outbound
doScenes: An Autonomous Driving Dataset with Natural Language Instruction for Human Interaction and Vision-Language Navigation Safe navigation with human instructions in complex scenes,
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 00a1ef50-6b58-4e04-862e-6cc1173b4d8e · outbound
doScenes: An Autonomous Driving Dataset with Natural Language Instruction for Human Interaction and Vision-Language Navigation Multimodal trajectory prediction conditioned on lane-graph traversals,
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation a50a6262-c3e0-48fc-a428-208815f64282 · outbound
doScenes: An Autonomous Driving Dataset with Natural Language Instruction for Human Interaction and Vision-Language Navigation Thomas: Trajectory heatmap output with learned multi-agent sam- pling,
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 225b3fde-5ad8-4ac8-a355-ce86e3a8ab70 · outbound
doScenes: An Autonomous Driving Dataset with Natural Language Instruction for Human Interaction and Vision-Language Navigation Trajectory prediction in autonomous driving with a lane heading auxiliary loss,
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation bb20e531-d271-491c-8aea-515a6f177026 · outbound
doScenes: An Autonomous Driving Dataset with Natural Language Instruction for Human Interaction and Vision-Language Navigation Spatialrgpt: Grounded spatial reasoning in vision-language models,
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 605b6ada-bced-410b-ab1d-d95e1c184e27 · inbound
Automated Data Curation Using GPS & NLP to Generate Instruction-Action Pairs for Autonomous Vehicle Vision-Language Navigation Datasets doScenes: An Autonomous Driving Dataset with Natural Language Instruction for Human Interaction and Vision-Language Navigation
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4eaaf0ee-dd26-492f-a0b8-1012ac20ae43 · inbound
Generative AI for Autonomous Driving: Frontiers and Opportunities doScenes: An Autonomous Driving Dataset with Natural Language Instruction for Human Interaction and Vision-Language Navigation
Reference 152
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ed211282-ce36-4975-b8ce-3f8a2e67abbc · inbound
Looking and Listening Inside and Outside: Multimodal Artificial Intelligence Systems for Driver Safety Assessment and Intelligent Vehicle Decision-Making doScenes: An Autonomous Driving Dataset with Natural Language Instruction for Human Interaction and Vision-Language Navigation
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation a258f878-1f5f-43f3-b257-10773882147e · inbound
NudgeVAD: Language-Nudged End-to-End Driving via FiLM Residuals doScenes: An Autonomous Driving Dataset with Natural Language Instruction for Human Interaction and Vision-Language Navigation
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.