Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T11:23:14.558335Z
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 27 of 27 outbound references and 0 inbound Pith citation observations for arXiv:2506.02615.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T11:23:14.558335Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
27 of 27 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation b0cbb2dd-c1e8-4c7c-8218-59c0e5f81b4c · outbound
Hierarchical Question-Answering for Driving Scene Understanding Using Vision-Language Models Spherical transformer for lidar-based 3d recognition,
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation db2325a5-b133-499e-b29b-043b829e37ba · outbound
Hierarchical Question-Answering for Driving Scene Understanding Using Vision-Language Models Rea- son2drive: Towards interpretable and chain-based reasoning for au- tonomous driving,
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ebd7634f-7846-439d-a3ec-c488704bf2aa · outbound
Hierarchical Question-Answering for Driving Scene Understanding Using Vision-Language Models GPT-4V as Traffic Assistant: An In-depth Look at Vision Language Model on Complex Traffic Events
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b21e0480-4d1f-4f23-b446-a5dcb535547a · outbound
Hierarchical Question-Answering for Driving Scene Understanding Using Vision-Language Models DriveVLM: The convergence of autonomous driving and large vision-language models,
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a2ca9fb8-e7f4-4487-acec-05afb4933ac4 · outbound
Hierarchical Question-Answering for Driving Scene Understanding Using Vision-Language Models BLIP: Bootstrapping language- image pre-training for unified vision-language understanding and generation,
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6963edf6-5390-4462-8e3d-3c80c5aa53aa · outbound
Hierarchical Question-Answering for Driving Scene Understanding Using Vision-Language Models Clip2scene: Towards label-efficient 3d scene understanding by clip,
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 597f786b-fe0c-447d-9e0e-11ca408ee3c9 · outbound
Hierarchical Question-Answering for Driving Scene Understanding Using Vision-Language Models Jiang, A. Tagliasacchi, M. Pollefeys, and T. Funkhouser, “Openscene: 3d scene understanding with open vocab- ularies,
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e316f44c-cf65-42b6-8a72-d5771e19f43d · outbound
Hierarchical Question-Answering for Driving Scene Understanding Using Vision-Language Models Pla: Language-driven open-vocabulary 3d scene understanding,
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 8e082adb-b2e0-4caf-8991-74438cb51821 · outbound
Hierarchical Question-Answering for Driving Scene Understanding Using Vision-Language Models Learning transferable visual models from natural language supervision,
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0c7d06ac-d4f1-4a68-9d5a-24fd23cf89eb · outbound
Hierarchical Question-Answering for Driving Scene Understanding Using Vision-Language Models The traffic scene un- derstanding and prediction based on image captioning,
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 101ebe9c-9183-4f49-82aa-f422d7fb56a6 · outbound
Hierarchical Question-Answering for Driving Scene Understanding Using Vision-Language Models Delving into clip latent space for video anomaly recognition,
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 212ca42e-a757-42f2-ac76-e26492b69d59 · outbound
Hierarchical Question-Answering for Driving Scene Understanding Using Vision-Language Models Learning to prompt for vision-language models,
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 08761c9a-b025-4ba5-8f76-416a3fc6726a · outbound
Hierarchical Question-Answering for Driving Scene Understanding Using Vision-Language Models LiDAR-LLM: Exploring the Potential of Large Language Models for 3D LiDAR Understanding
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 43ab1c19-b7c8-48aa-aa60-4a48018b4346 · outbound
Hierarchical Question-Answering for Driving Scene Understanding Using Vision-Language Models Vectornet: Encoding hd maps and agent dynamics from vectorized representation,
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation bfe27062-e1d8-437a-a555-578abb92c5fb · outbound
Hierarchical Question-Answering for Driving Scene Understanding Using Vision-Language Models Reimagining an autonomous vehicle
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a0a3c807-ec43-42a4-8d44-2a30e46a5e91 · outbound
Hierarchical Question-Answering for Driving Scene Understanding Using Vision-Language Models Explainable artificial intelligence (xai): Concepts, taxonomies, opportunities and challenges toward responsible ai,
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 26ff000f-8917-4c67-bca6-31f7d4ad6930 · outbound
Hierarchical Question-Answering for Driving Scene Understanding Using Vision-Language Models From automation to autonomy and autonomous vehicles: Challenges and opportunities for human-computer interaction,
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation fd326759-a542-4496-b394-370e4317e8fb · outbound
Hierarchical Question-Answering for Driving Scene Understanding Using Vision-Language Models LanguageMPC: Large Language Models as Decision Makers for Autonomous Driving
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 009525c1-7108-4b0f-841b-2898fddf9f13 · outbound
Hierarchical Question-Answering for Driving Scene Understanding Using Vision-Language Models Lm-nav: Robotic navigation with large pre-trained models of language, vision, and action,
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 439f4f6a-e764-44cf-bd1c-228f1e3b279a · outbound
Hierarchical Question-Answering for Driving Scene Understanding Using Vision-Language Models DiLu: A Knowledge-Driven Approach to Autonomous Driving with Large Language Models
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c9dcfaf7-e65c-424a-92cf-24e977e57cb6 · outbound
Hierarchical Question-Answering for Driving Scene Understanding Using Vision-Language Models Driving with llms: Fusing object- level vector modality for explainable autonomous driving,
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a06e06f1-6544-4233-b29f-5e7294f6bfdf · outbound
Hierarchical Question-Answering for Driving Scene Understanding Using Vision-Language Models GPT-Driver: Learning to Drive with GPT
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9bc840b0-96eb-4faf-bba9-7f50d9c1cd97 · outbound
Hierarchical Question-Answering for Driving Scene Understanding Using Vision-Language Models Drivegpt4: Interpretable end-to-end autonomous driving via large language model,
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f106eb8c-399d-4736-be2e-3d6bc7a9dd8d · outbound
Hierarchical Question-Answering for Driving Scene Understanding Using Vision-Language Models RAG-Driver: Generalisable Driving Explanations with Retrieval-Augmented In-Context Learning in Multi-Modal Large Lan- guage Model,
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9962cfb5-f5b4-4242-9d78-35efb8d775fc · outbound
Hierarchical Question-Answering for Driving Scene Understanding Using Vision-Language Models DriveMLM: Aligning Multi-Modal Large Language Models with Behavioral Planning States for Autonomous Driving,
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ecc36b9e-663c-4f8c-9e43-53724ef1cebb · outbound
Hierarchical Question-Answering for Driving Scene Understanding Using Vision-Language Models Lmdrive: Closed-loop end-to-end driving with large language models,
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0de58213-75e5-4885-9290-f651cc754242 · outbound
Hierarchical Question-Answering for Driving Scene Understanding Using Vision-Language Models Lingoqa: Visual question answering for au- tonomous driving,
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
No inbound Pith citation observations are available.