Pith. sign in

Paper Citation Record · LEDGER

Multimodal Framework for Explainable Autonomous Driving: Integrating Video, Sensor, and Textual Data for Enhanced Decision-Making and Transparency

As of 7 August 2026, this Paper Citation Record lists 21 of 21 outbound references and 0 inbound Pith citation observations for arXiv:2507.07938.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.07938 v1

Coverage vector

measured 21 of 21 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T18:33:26.603654Z

measured 21 of 21 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-06T06:34:29.942622+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

21 of 21 outbound references displayed

  • verified exact0
  • verified fuzzy2
  • unresolved13
  • parse uncertain0
  • malformed identifier2
  • metadata mismatch4

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 20539d96-2825-4042-a2bb-3fb6e3d6b48d · outbound

This paper cites A survey of autonomous driving: Common practices and emerging technolo- gies,.

Multimodal Framework for Explainable Autonomous Driving: Integrating Video, Sensor, and Textual Data for Enhanced Decision-Making and Transparency A survey of autonomous driving: Common practices and emerging technolo- gies,

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T18:33:24.489572Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:33:24.489572Z digest=sha256:90338d3ea1fed567001a76ff970faa769d6cb294c4bfe933188eeef9ff4fc636

Observation 8520d665-ff66-47ee-8dbd-4e72f7cb95e7 · outbound

This paper cites ClusT3: Information Invariant Test-Time Training.

Multimodal Framework for Explainable Autonomous Driving: Integrating Video, Sensor, and Textual Data for Enhanced Decision-Making and Transparency ClusT3: Information Invariant Test-Time Training

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T18:33:24.607371Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:33:24.607371Z digest=sha256:5db11d3f055a9973272049d142cbe0dd24fbd29f5d3dd467fcfa09de9dd3297c

Observation f82fddac-1ecb-415b-9d97-1fe8922bc95b · outbound

This paper cites Planning-oriented autonomous driving,.

Multimodal Framework for Explainable Autonomous Driving: Integrating Video, Sensor, and Textual Data for Enhanced Decision-Making and Transparency Planning-oriented autonomous driving,

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T18:33:24.769558Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:33:24.769558Z digest=sha256:aa083b164c196b5896951b5b386b74a889cb4448cc88d04bc85162f451ff9020

Observation 6c117f6c-590a-4b39-85ec-4e8a6c0c984d · outbound

This paper cites Deep multi-modal object detection and semantic segmentation for au- tonomous driving,.

Multimodal Framework for Explainable Autonomous Driving: Integrating Video, Sensor, and Textual Data for Enhanced Decision-Making and Transparency Deep multi-modal object detection and semantic segmentation for au- tonomous driving,

Reference 4

Resolution
metadata mismatch
raw_fallback, observed 2026-08-06T18:33:27.877104Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T18:33:24.896397Z digest=sha256:5f0259f1ef634add5f1717fa626be1ab6310929bcd8865bbf29e69a7c95ad922

Observation 7e43c368-c3a3-400b-b3cd-23fdfd13c3fc · outbound

This paper cites Ex- plainable artificial intelligence (XAI),.

Multimodal Framework for Explainable Autonomous Driving: Integrating Video, Sensor, and Textual Data for Enhanced Decision-Making and Transparency Ex- plainable artificial intelligence (XAI),

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T18:33:25.053146Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:33:25.053146Z digest=sha256:f8fae14ec73ff4cf72bdfebb6764f0ba62581a6155ef676b98ec616f99b82ec7

Observation 2a9fe5ed-4484-4588-930e-de14a4d03178 · outbound

This paper cites Why did the AI make that decision?,.

Multimodal Framework for Explainable Autonomous Driving: Integrating Video, Sensor, and Textual Data for Enhanced Decision-Making and Transparency Why did the AI make that decision?,

Reference 6

Resolution
metadata mismatch
raw_fallback, observed 2026-08-06T18:33:27.596507Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T18:33:25.148590Z digest=sha256:b4e928c2b423670e67deac1ab91208ebcc5323c5324f750fe6fb53e2d6aec961

Observation 71590303-b49f-48f7-9bed-69c9d9e00b4d · outbound

This paper cites Ex- plainable artificial intelligence (XAI),.

Multimodal Framework for Explainable Autonomous Driving: Integrating Video, Sensor, and Textual Data for Enhanced Decision-Making and Transparency Ex- plainable artificial intelligence (XAI),

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T18:33:25.251827Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:33:25.251827Z digest=sha256:0e2ec6d6386bcdad9395581609dc5a8aeed6867f1fb424a1ae9b6b7672c0ea68

Observation 45f9e0cd-7f03-410d-b6de-9df6a1259202 · outbound

This paper cites Interpretable autonomous driving: A sur- vey of recent advances,.

Multimodal Framework for Explainable Autonomous Driving: Integrating Video, Sensor, and Textual Data for Enhanced Decision-Making and Transparency Interpretable autonomous driving: A sur- vey of recent advances,

Reference 8

Resolution
malformed identifier
no resolver link, observed 2026-08-06T18:33:25.417228Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:33:25.417228Z digest=sha256:e44bb01cc7688de9164f9328fa55761c9bfc71c74f045aaa020e47ad720f04a6

Observation 63ae09ef-8a4d-4c8b-af86-25bb78cb87e8 · outbound

This paper cites Attention- based multimodal framework for au- tonomous driving,.

Multimodal Framework for Explainable Autonomous Driving: Integrating Video, Sensor, and Textual Data for Enhanced Decision-Making and Transparency Attention- based multimodal framework for au- tonomous driving,

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T18:33:25.485181Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:33:25.485181Z digest=sha256:cc4822059be4cdfd5300c281c3d3e1ac6539358446dc7e8ac065f1f5c1b85c97

Observation bbbe183c-1677-42ea-b03b-7709cd130195 · outbound

This paper cites DeepDriving: Learning affordance for direct perception in autonomous driving,.

Multimodal Framework for Explainable Autonomous Driving: Integrating Video, Sensor, and Textual Data for Enhanced Decision-Making and Transparency DeepDriving: Learning affordance for direct perception in autonomous driving,

Reference 10

Resolution
malformed identifier
no resolver link, observed 2026-08-06T18:33:25.565184Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:33:25.565184Z digest=sha256:9bd3fdfc3e214329cf9f34bf426f7b8f3989117cdf082200ab332358cf671a4d

Observation fdfd001c-ed76-4d3c-8e29-360eeecfa416 · outbound

This paper cites CARLA: An Open Urban Driving Simulator.

Multimodal Framework for Explainable Autonomous Driving: Integrating Video, Sensor, and Textual Data for Enhanced Decision-Making and Transparency CARLA: An Open Urban Driving Simulator

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T18:33:25.619982Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:33:25.619982Z digest=sha256:6c301a40268a5421a0f199e95e6dbbd80b3392fb0aa2246d405c2c02bc0a041b

Observation 13d5428e-6ba5-44f3-9677-ad9af44217c8 · outbound

This paper cites VideoMAE: Masked Autoencoders are Data-Efficient Learners for Self-Supervised Video Pre-Training.

Multimodal Framework for Explainable Autonomous Driving: Integrating Video, Sensor, and Textual Data for Enhanced Decision-Making and Transparency VideoMAE: Masked Autoencoders are Data-Efficient Learners for Self-Supervised Video Pre-Training

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T18:33:25.677404Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:33:25.677404Z digest=sha256:452aa9cdda57c29912fa0be5c941163009ae02ba75a2258099c2942368c2e8bc

Observation 62b49ade-ec11-4b84-97cf-ba2a095ad09f · outbound

This paper cites BERT: Pre-training of deep bidi- rectional transformers for language under- standing,.

Multimodal Framework for Explainable Autonomous Driving: Integrating Video, Sensor, and Textual Data for Enhanced Decision-Making and Transparency BERT: Pre-training of deep bidi- rectional transformers for language under- standing,

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T18:33:25.751970Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:33:25.751970Z digest=sha256:6a0f49bca67652ef0c208cfb64fa1bff9facc9feb06566de9e39936d4ae15041

Observation 07b0436b-7af7-4f2d-b463-bff89e3a65bc · outbound

This paper cites nuScenes: A mul- timodal dataset for autonomous driving,.

Multimodal Framework for Explainable Autonomous Driving: Integrating Video, Sensor, and Textual Data for Enhanced Decision-Making and Transparency nuScenes: A mul- timodal dataset for autonomous driving,

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T18:33:25.938197Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:33:25.938197Z digest=sha256:4245433a8a97a1722ced22ba52d4d21d8848a03ea703184db4a2590206960761

Observation d66212bc-28ca-4069-88b0-d919a9919170 · outbound

This paper cites BART: Denoising sequence- to-sequence pre-training for natural lan- guage generation, translation, and com- prehension,.

Multimodal Framework for Explainable Autonomous Driving: Integrating Video, Sensor, and Textual Data for Enhanced Decision-Making and Transparency BART: Denoising sequence- to-sequence pre-training for natural lan- guage generation, translation, and com- prehension,

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T18:33:26.045042Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:33:26.045042Z digest=sha256:4e869a64ac577eeb24d35cb6fb3b25fa3481615f6f79df05865ef4d6a11fdaad

Observation 5f648b4d-208e-469f-bbe9-b5e6a4e582f2 · outbound

This paper cites Language-augmented Bird’s-eye View Maps for autonomous driving,.

Multimodal Framework for Explainable Autonomous Driving: Integrating Video, Sensor, and Textual Data for Enhanced Decision-Making and Transparency Language-augmented Bird’s-eye View Maps for autonomous driving,

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T18:33:26.182936Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:33:26.182936Z digest=sha256:9fd38e3932adebec99a7a9e5dab1a55a39ccd46a42bd4d25488dec74ed501acc

Observation 059c3613-b99e-4c49-a31e-bcee2249e905 · outbound

This paper cites Vi- sual question answering and natural lan- guage explanations for autonomous driv- ing,.

Multimodal Framework for Explainable Autonomous Driving: Integrating Video, Sensor, and Textual Data for Enhanced Decision-Making and Transparency Vi- sual question answering and natural lan- guage explanations for autonomous driv- ing,

Reference 18

Resolution
metadata mismatch
raw_fallback, observed 2026-08-06T18:33:27.076617Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T18:33:26.293828Z digest=sha256:ff8a3a8c0892c443ddad4b09691f867fc66ef5e56ccf59836d3e67af68dcabdb

Observation 0ad41744-5391-44a0-8385-d63e9f736a87 · outbound

This paper cites Antagonising explanation and revealing bias directly through sequencing and multimodal inference.

Multimodal Framework for Explainable Autonomous Driving: Integrating Video, Sensor, and Textual Data for Enhanced Decision-Making and Transparency Antagonising explanation and revealing bias directly through sequencing and multimodal inference

Reference 19

Resolution
metadata mismatch
local_arxiv, observed 2026-08-06T18:33:26.829568Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T18:33:26.341130Z digest=sha256:ebd137ea7257f07462a1112510a8eb20523b8297367b397a5a305aca3efdd268

Observation b7c750e8-da9a-4a73-90ee-405e352b7312 · outbound

This paper cites an unresolved cited work.

Multimodal Framework for Explainable Autonomous Driving: Integrating Video, Sensor, and Textual Data for Enhanced Decision-Making and Transparency Unresolved cited work

Reference 20

Resolution
unresolved
raw_fallback, observed 2026-08-06T18:33:28.597555Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T18:33:26.410084Z digest=sha256:bbd5124a2ec7e061177d453ea32fba312a2b73b7a724b8db358321824039de28

Observation 6f7ac32a-f619-4442-8762-8e0c5f5a6ff2 · outbound

This paper cites ROUGE: A package for automatic evaluation of summaries,.

Multimodal Framework for Explainable Autonomous Driving: Integrating Video, Sensor, and Textual Data for Enhanced Decision-Making and Transparency ROUGE: A package for automatic evaluation of summaries,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:33:28.484960Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T18:33:26.474538Z digest=sha256:a3e3f5620f7036f7d3feb50fc9a1c67a1201f7901b0ef76116673e36c406dc56

Observation 1a763c5f-6605-4540-8378-1459416a7ed9 · outbound

This paper cites METEOR: An automatic metric for MT evaluation with im- proved correlation with human judgments,.

Multimodal Framework for Explainable Autonomous Driving: Integrating Video, Sensor, and Textual Data for Enhanced Decision-Making and Transparency METEOR: An automatic metric for MT evaluation with im- proved correlation with human judgments,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:33:28.231684Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T18:33:26.603654Z digest=sha256:32cbd044cbeee69176285f63af813bcd6943d4f5ece70110376b9b1a1da062dd

Pith citing papers

No inbound Pith citation observations are available.