Pith. sign in

Paper Citation Record · LEDGER

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving

As of 6 August 2026, this Paper Citation Record lists 82 of 82 outbound references and 0 inbound Pith citation observations for arXiv:2512.04733.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2512.04733 v3

Coverage vector

measured 82 of 82 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-03T18:37:03.723052Z

measured 82 of 82 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-06T06:34:29.942622+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

82 of 82 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved82
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation fda3d3f5-307a-4752-a66d-449ce8e0b3ad · outbound

This paper cites Qwen2.5-VL Technical Report.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving Qwen2.5-VL Technical Report

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:01.908488Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:01.908488Z digest=sha256:b21f6397241e9b812581481dd61af66832514dbcf21a11592235c09ef2661ff3

Observation 3930f420-ec89-4bba-a6db-9978dd834442 · outbound

This paper cites Spatial memory: how egocentric and allocentric combine.Trends in cognitive sciences, 10(12):551–557, 2006.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving Spatial memory: how egocentric and allocentric combine.Trends in cognitive sciences, 10(12):551–557, 2006

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:02.005335Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:02.005335Z digest=sha256:d78fd430fb49b56a5d0204acbc643823256525a35f007442f7a5ea8797b965f4

Observation 3f5eb0eb-53fa-4a73-88f9-2e715c8bba8d · outbound

This paper cites nuscenes: A multi- modal dataset for autonomous driving.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving nuscenes: A multi- modal dataset for autonomous driving

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:02.109624Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:02.109624Z digest=sha256:38fb16524b2346718b89fa6c18f0d78b9e1cbb449022768bb117c35831439216

Observation 78bd2a0e-ab73-42b3-9a4e-9f2fc4e7ce2e · outbound

This paper cites Ground- ing commands for autonomous vehicles via layer fusion with region-specific dynamic layer attention.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving Ground- ing commands for autonomous vehicles via layer fusion with region-specific dynamic layer attention

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:02.219643Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:02.219643Z digest=sha256:04e0868842700801907420b3eb1fdc7f85c1541e0559154420584b177a0e7005

Observation bf81f2cd-49be-4c04-a5b5-91d480aac5ee · outbound

This paper cites Eeg-based emotion recognition for road accidents in a simulated driving environment.Biomedical signal processing and control, 87:105411, 2024.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving Eeg-based emotion recognition for road accidents in a simulated driving environment.Biomedical signal processing and control, 87:105411, 2024

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:02.279160Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:02.279160Z digest=sha256:86343ae9afe6fdfb7d809632db151fc9977dff409da2f8687f47aaed82dc7ec9

Observation f3119527-444e-4e19-93df-4b76e6b6b1e7 · outbound

This paper cites End-to-end autonomous driving: Challenges and frontiers.IEEE Transactions on Pattern Analysis and Machine Intelligence, 2024.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving End-to-end autonomous driving: Challenges and frontiers.IEEE Transactions on Pattern Analysis and Machine Intelligence, 2024

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:02.437211Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:02.437211Z digest=sha256:2fd97f289eca32ee8fad37f0e8b832a6f0f437157a4411690600240fd9132982

Observation e97273e4-b293-4266-ad01-20db6a5b6e70 · outbound

This paper cites Emotion-aware design in automobiles: Embracing technology advancements to enhance human-vehicle interaction.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving Emotion-aware design in automobiles: Embracing technology advancements to enhance human-vehicle interaction

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:02.603358Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:02.603358Z digest=sha256:a0f017b3372d478ebda14aae3efa5b329b5a64b9bc8c499be4bca525bb004b28

Observation 8fead26e-8a64-41b6-a547-2108a5ded26c · outbound

This paper cites Emotion recognition in human-computer interaction.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving Emotion recognition in human-computer interaction

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:02.702226Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:02.702226Z digest=sha256:9fbebfd363580e0daf721635a5f5cc672145d547c5999670ed7be366f7a0507f

Observation cc90a849-5a7d-4dd2-b404-d27223fd6a0e · outbound

This paper cites A survey on multimodal large language models for autonomous driving.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving A survey on multimodal large language models for autonomous driving

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:02.763541Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:02.763541Z digest=sha256:c7c7de956c9148ea7ccb52d841dd9b68fc792e8dd384e1349fba91db95b75fb7

Observation 019e1247-76ed-4e72-a0e7-ff8b3c932dd0 · outbound

This paper cites Com- mands for autonomous vehicles by progressively stacking visual-linguistic representations.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving Com- mands for autonomous vehicles by progressively stacking visual-linguistic representations

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:02.862493Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:02.862493Z digest=sha256:4a8189b40dcedbb38704f8fa713639b5d6c6acf6a569a70cb65f43144764a837

Observation 61afb2d4-bf7f-4fd5-8f93-9350630c4f55 · outbound

This paper cites GoEmotions: A Dataset of Fine-Grained Emotions.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving GoEmotions: A Dataset of Fine-Grained Emotions

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:02.966502Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:02.966502Z digest=sha256:b8809326980bad718c2e7efb7f7764b93a0fd9623fa76a7b70c8010e0202f60d

Observation 0310f769-0b0b-42aa-92c6-fca76afdb989 · outbound

This paper cites Goal- gan: Multimodal trajectory prediction based on goal position estimation.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving Goal- gan: Multimodal trajectory prediction based on goal position estimation

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.084349Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.084349Z digest=sha256:ec946a8c3fc3e8259cc9fe519bec2435a7740c9d423bc9e26e030dc0db88a902

Observation 027ec76e-9250-40bb-8e5b-4ccc7bcbfcc3 · outbound

This paper cites Transvg: End-to-end visual ground- ing with transformers.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving Transvg: End-to-end visual ground- ing with transformers

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.190418Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.190418Z digest=sha256:8ead561b8f3bd50791e135424d2c739b237eb1c4c6dd5cb9b3cae3197f8bf928

Observation 7cf5a902-9b0c-4552-bd15-1edc167ef5fe · outbound

This paper cites Talk2Car: Taking Control of Your Self-Driving Car.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving Talk2Car: Taking Control of Your Self-Driving Car

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.277319Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.277319Z digest=sha256:1e57024b2f0d2c5efcdc72d270b67b396be2fac7edef5e742cd111803c4ac35c

Observation 48ec2828-0c7e-4caf-9bf0-c45a3fdf22e8 · outbound

This paper cites Talk2car: Predicting physical trajectories for natural language commands.Ieee Access, 10: 123809–123834, 2022.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving Talk2car: Predicting physical trajectories for natural language commands.Ieee Access, 10: 123809–123834, 2022

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.383485Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.383485Z digest=sha256:a3e923ebd71b5f93826f3211a21b39e8a83cab4140aeec72a137a0f3a273b154

Observation 6ee6a772-a3a3-4b65-bc8f-0dcce87a525b · outbound

This paper cites Bert: Pre-training of deep bidirectional transform- ers for language understanding.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving Bert: Pre-training of deep bidirectional transform- ers for language understanding

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.387154Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.387154Z digest=sha256:1a9b112cbbd098ee3892d374fc7b1d19178ebedea8fdc912e42a163b10e50df8

Observation e2bb7b4e-4696-41ed-9ac7-5d6e908d580f · outbound

This paper cites Multi-class emotion recognition within the valence-arousal-dominance space using eeg.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving Multi-class emotion recognition within the valence-arousal-dominance space using eeg

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.390464Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.390464Z digest=sha256:d031346336447dac0ba516bf0f8e5491da87bb2843798152aea69f4d0290c9db

Observation 5199d07d-9a41-4612-bf56-e1efe1898d34 · outbound

This paper cites Predicting physical world destina- tions for commands given to self-driving cars.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving Predicting physical world destina- tions for commands given to self-driving cars

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.393633Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.393633Z digest=sha256:9697d455c489dc7a815e607efb56c290326cec282b29d61ea9f3c2948524da85

Observation 6465536a-bedf-4af5-9dfe-15ca353abc0e · outbound

This paper cites Think before you drive: World model-inspired multimodal grounding for autonomous driving.arXiv preprint, 2025.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving Think before you drive: World model-inspired multimodal grounding for autonomous driving.arXiv preprint, 2025

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.397038Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.397038Z digest=sha256:5d7cc1d3377a02896cfe313c3c9956bd88bc5768ca77395776f050d82c6e06ee

Observation 7c553ee8-589a-4aec-80f7-c5a42bde8ed6 · outbound

This paper cites A formal basis for the heuristic determination of minimum cost paths.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving A formal basis for the heuristic determination of minimum cost paths

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.400130Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.400130Z digest=sha256:d39dca8cff4adc279ae16c5d5717d07929cd3a6580c5f9ca2438d24083ee9e63

Observation 28f7a811-872f-4ac2-8882-e298d864d099 · outbound

This paper cites Hu, Yelong Shen, Phillip Wallis, Zeyuan Allen- Zhu, Yuanzhi Li, Shean Wang, Lu Wang, and Weizhu Chen.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving Hu, Yelong Shen, Phillip Wallis, Zeyuan Allen- Zhu, Yuanzhi Li, Shean Wang, Lu Wang, and Weizhu Chen

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.403382Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.403382Z digest=sha256:d5d625352b58290dea9ba8d52ce063073d2d751d4b99a19b5020a555b168fd72

Observation 0d760349-5e9f-4050-94c3-ecfe12390898 · outbound

This paper cites Planning-oriented autonomous driving.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving Planning-oriented autonomous driving

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.406493Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.406493Z digest=sha256:a8a91f7b58366cf82ac787ed3e7734ff92027d6c36a6f564ef07e9ed4692091d

Observation a54a34c2-68c5-46c2-9b6b-f432381fbf07 · outbound

This paper cites Think twice before driving: Towards scalable decoders for end-to-end autonomous driving.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving Think twice before driving: Towards scalable decoders for end-to-end autonomous driving

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.409654Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.409654Z digest=sha256:b07f4db353f16b6b1eabae5c8290a260586a0d0fc8956429450aab11d4b2698b

Observation eec9da9d-0c22-4260-bd8e-0d1437df1798 · outbound

This paper cites Senna: Bridging Large Vision-Language Models and End-to-End Autonomous Driving.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving Senna: Bridging Large Vision-Language Models and End-to-End Autonomous Driving

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.412774Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.412774Z digest=sha256:ceddd6acf411a2d3aac6cf799d977eb3f9ea5be6273b50eeb1e6a23be7ca119f

Observation 92d6ba2c-4203-42ed-9441-008dca2c40cc · outbound

This paper cites A Survey on Vision-Language-Action Models for Autonomous Driving.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving A Survey on Vision-Language-Action Models for Autonomous Driving

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.416823Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.416823Z digest=sha256:cb5fd04af29f11f91819fd4c3718f88b288ad7d06bf3119293a1cb3c10f29acf

Observation 5b53327a-e621-4924-8a82-ebbf3472d915 · outbound

This paper cites Mdetr-modulated detection for end-to-end multi-modal understanding.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving Mdetr-modulated detection for end-to-end multi-modal understanding

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.420264Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.420264Z digest=sha256:12bf97c685d8b703a4b7e78f9b59005bf33de7f897533f2833aa99f06588c73c

Observation 4256fc99-65fb-4313-9b5a-6b539ee99f10 · outbound

This paper cites Detection of drivers’ anxiety invoked by driving situations using multimodal biosignals.Processes, 8 (2):155, 2020.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving Detection of drivers’ anxiety invoked by driving situations using multimodal biosignals.Processes, 8 (2):155, 2020

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.423509Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.423509Z digest=sha256:a7b3669715d1edc62120eb1fef7d8f8891c29d333798f42740493b9e148aaff9

Observation fb6447ce-1f1f-45a2-88ad-1a2061d42b28 · outbound

This paper cites Review and perspectives on human emotion for connected automated vehicles.Automotive Innovation, 7(1):4–44, 2024.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving Review and perspectives on human emotion for connected automated vehicles.Automotive Innovation, 7(1):4–44, 2024

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.426705Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.426705Z digest=sha256:366cdd807583a65dc4bbe6e98156cdc193981b22458f42f81876f8b8d1f3bd87

Observation 41e42b77-c10a-46ce-927d-6d54987d2368 · outbound

This paper cites DriveVLA-W0: World Models Amplify Data Scaling Law in Autonomous Driving.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving DriveVLA-W0: World Models Amplify Data Scaling Law in Autonomous Driving

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.430100Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.430100Z digest=sha256:ba8468c2907ebc2644279d8f80522b077b704f1b857d2b89fe96f56409583b7d

Observation abfb8096-6550-492a-8eb4-3cd4962bc0d3 · outbound

This paper cites Fine-grained evaluation of large vision-language mod- els in autonomous driving.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving Fine-grained evaluation of large vision-language mod- els in autonomous driving

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.433519Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.433519Z digest=sha256:dd10777dd71008cf5b41488c9bd79810a349926313f9a352837b3b730afd0fb8

Observation 592f17f5-0761-4c59-808a-a331ad8325ad · outbound

This paper cites ReCogDrive: A Reinforced Cognitive Framework for End-to-End Autonomous Driving.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving ReCogDrive: A Reinforced Cognitive Framework for End-to-End Autonomous Driving

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.436654Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.436654Z digest=sha256:7ffc4211ed03471e375a45fea27efc42383b04274cddfc0ac74257ac0d19ed4e

Observation 164db171-1c2c-4cef-98a4-e2e067982719 · outbound

This paper cites Steering the future: Redefining intelligent transportation systems with foundation models.Chain, 1(1):46–53, 2024.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving Steering the future: Redefining intelligent transportation systems with foundation models.Chain, 1(1):46–53, 2024

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.440641Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.440641Z digest=sha256:45b2bfd9a525ae5c18a5d8facb73f2395923eb336f16eb46a41b01230dd528fc

Observation 8ef1f9c2-1860-4c95-98c1-d937321234c6 · outbound

This paper cites Mamba-va: A mamba-based approach for continuous emotion recognition in valence-arousal space.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving Mamba-va: A mamba-based approach for continuous emotion recognition in valence-arousal space

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.444039Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.444039Z digest=sha256:b2faf100572747e1eeb7f5d5590c33e7cb24a47271aaecf6eff7c42d8ac0709d

Observation 4c9c1a28-1530-4ffb-aa5c-e9ad2595e409 · outbound

This paper cites Gpt- 4 enhanced multimodal grounding for autonomous driving: Leveraging cross-modal attention with large language models.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving Gpt- 4 enhanced multimodal grounding for autonomous driving: Leveraging cross-modal attention with large language models

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.447243Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.447243Z digest=sha256:8f3e42a4a5975488cdf16525dfd8a1aa7ef2ec121a7eefb79da3c8b528d39baa

Observation c6d9346f-f898-48e6-a30c-4257c43dfce9 · outbound

This paper cites Cot-drive: Efficient motion forecasting for autonomous driving with llms and chain-of-thought prompting.IEEE Transactions on Artificial Intelligence, 2025.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving Cot-drive: Efficient motion forecasting for autonomous driving with llms and chain-of-thought prompting.IEEE Transactions on Artificial Intelligence, 2025

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.450634Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.450634Z digest=sha256:8fdf673d041bfeec7f10e0b2938f1ea9aff4f327187fface9b73e945d8ede527

Observation a3ff138e-769d-4c4e-b876-1169c5e1149e · outbound

This paper cites Toward human-like trajectory predic- tion for autonomous driving: A behavior-centric approach.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving Toward human-like trajectory predic- tion for autonomous driving: A behavior-centric approach

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.453786Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.453786Z digest=sha256:38d5625ea819d0376b1ecac566988cc549e9bdbd0f8d15df1f6a0afe5c561774

Observation 60d2c7b5-60f5-452e-be8e-3e76368f64ae · outbound

This paper cites CogDriver: Integrating Cognitive Inertia for Temporally Coherent Planning in Autonomous Driving.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving CogDriver: Integrating Cognitive Inertia for Temporally Coherent Planning in Autonomous Driving

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.457803Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.457803Z digest=sha256:7e38128a6aeca4e0e9df940f04489364eab9557e3ce8b37bd7732141d935180e

Observation 59161342-bc84-4f12-a823-d1ec6ea125fe · outbound

This paper cites Grounding DINO: Marrying DINO with Grounded Pre-Training for Open-Set Object Detection.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving Grounding DINO: Marrying DINO with Grounded Pre-Training for Open-Set Object Detection

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.461246Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.461246Z digest=sha256:a445af2851cb2a3e8611cace9b4938da9aac4d08003e27cdbe26a877e6866b32

Observation 0fc9e669-8ccd-481f-95d9-960a95edf6c5 · outbound

This paper cites RoBERTa: A Robustly Optimized BERT Pretraining Approach.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving RoBERTa: A Robustly Optimized BERT Pretraining Approach

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.464784Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.464784Z digest=sha256:c185bc2ca1eaa4c2f7d6977f87fa5355ea4f57f0772d73f999980bffe7e4aa8f

Observation 2f771d60-f432-4a2a-8fec-9437ea2b5537 · outbound

This paper cites C4av: learning cross-modal representations from transformers.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving C4av: learning cross-modal representations from transformers

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.468365Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.468365Z digest=sha256:ee82a94857001b44c4881a202907dea4ac4233bbded1448b366d954adc6e855a

Observation d0054764-b0eb-4d7b-b6ab-4420c5b7d1aa · outbound

This paper cites From goals, waypoints & paths to long term human trajectory forecasting.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving From goals, waypoints & paths to long term human trajectory forecasting

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.471618Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.471618Z digest=sha256:4ba838ba4a42cfb92d653cf0e325465f21d747ed999972f88dfc403e341570f6

Observation a9b7a4fd-71aa-4af6-980b-00e31e5a78fe · outbound

This paper cites Person- ality correlates of driver stress.Personality and Individual Differences, 12(6):535–549, 1991.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving Person- ality correlates of driver stress.Personality and Individual Differences, 12(6):535–549, 1991

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.475111Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.475111Z digest=sha256:9e6fd10650f520e27250b470ff3e6d444637ad4e02dcd47e861083ada7122b31

Observation 05f04bb3-68ce-462f-aa69-0a66b0183311 · outbound

This paper cites Attngrounder: Talking to cars with attention.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving Attngrounder: Talking to cars with attention

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.478914Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.478914Z digest=sha256:5c5fc62dc0399823f64a8fdbf470d048955bb73fac1da41e074bfbb0929e2a4d

Observation 03cdd1aa-9a89-410a-8288-1e5f388a91bc · outbound

This paper cites Mohammad.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving Mohammad

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.482355Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.482355Z digest=sha256:e89516e33b8620085920b9b197d14f5d8b8661d2fe45e2d1cb80f700921c6584

Observation 49629f17-eb20-4156-b46b-60bee485f144 · outbound

This paper cites Driver emotion recognition with a hybrid attentional multimodal fusion framework.IEEE Transactions on Affective Computing, 14(4):2970–2981, 2023.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving Driver emotion recognition with a hybrid attentional multimodal fusion framework.IEEE Transactions on Affective Computing, 14(4):2970–2981, 2023

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.485750Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.485750Z digest=sha256:f1659929d0b7b990f3ff7463ec85c2190ba124851ddb3278a3f942862ecc5b6f

Observation c2ade741-bd71-42f6-9a97-e89e143b69c6 · outbound

This paper cites Steerable Adversarial Scenario Generation through Test-Time Preference Alignment.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving Steerable Adversarial Scenario Generation through Test-Time Preference Alignment

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.489316Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.489316Z digest=sha256:da4903c8ee6931c702a5c45dc46b6d11ed37c96e05fe8583d6732059adb260de

Observation ff74b0dc-11e7-4fb4-8cf7-ceefdba0af67 · outbound

This paper cites Vlp: Vision language planning for autonomous driving.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving Vlp: Vision language planning for autonomous driving

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.493431Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.493431Z digest=sha256:4139ad37421808f9c7814323e57dd8a984009a44f76379d0a11081eccc14329e

Observation 20c430bb-5f4c-4996-bb4f-c57b5bb94b18 · outbound

This paper cites Multi- modal fusion transformer for end-to-end autonomous driving.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving Multi- modal fusion transformer for end-to-end autonomous driving

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.496826Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.496826Z digest=sha256:07cabac642abc88d7df714dd9a3c5b844b96d09332ce89e92a0f4430f441b672

Observation 4e999f0f-b598-474e-a806-35ba6c449071 · outbound

This paper cites Direct prefer- ence optimization: Your language model is secretly a reward model.Advances in neural information processing systems, 36:53728–53741, 2023.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving Direct prefer- ence optimization: Your language model is secretly a reward model.Advances in neural information processing systems, 36:53728–53741, 2023

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.500524Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.500524Z digest=sha256:29c11faada31bbaaf41c89014601d105678546c68e6e89dec87e144929996a55

Observation 5b2652a7-24d8-4252-ad2a-ef83a9f1461b · outbound

This paper cites Simlingo: Vision-only closed-loop autonomous driving with language-action alignment.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving Simlingo: Vision-only closed-loop autonomous driving with language-action alignment

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.503797Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.503797Z digest=sha256:c7c2f051b219773e67681f5549ee26af279f4642c32ac4bd7dbd68dc0e1b6072

Observation 256c2928-f6fe-44d5-9df8-d866e39e3abc · outbound

This paper cites Cosine meets softmax: A tough-to-beat baseline for visual grounding.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving Cosine meets softmax: A tough-to-beat baseline for visual grounding

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.612115Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.612115Z digest=sha256:5717366619f53cda7888c615e44db1744104a9649caf9833ac36a1d5f75fb606

Observation abe76bdb-0c96-41cf-a1a9-4d095699d7ca · outbound

This paper cites DistilBERT, a distilled version of BERT: smaller, faster, cheaper and lighter.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving DistilBERT, a distilled version of BERT: smaller, faster, cheaper and lighter

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.616071Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.616071Z digest=sha256:93c86ccbbf604e1b2ae089f48901f9905adb28e857cec753482c23e928b0a482

Observation a842e73d-7073-4107-b7f5-8bc6be2131ab · outbound

This paper cites Vision-language-action models: Con- cepts, progress, applications and challenges.arXiv preprint arXiv:2505.04769, 2025.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving Vision-language-action models: Con- cepts, progress, applications and challenges.arXiv preprint arXiv:2505.04769, 2025

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.620338Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.620338Z digest=sha256:f490f5bc1d4f32ba81b143fd7292ea8dbfc1688913e3cfa19211ac9aae511ce4

Observation d67903a5-2881-420e-95f2-bdf2fe0ec06d · outbound

This paper cites Lmdrive: Closed-loop end-to-end driving with large language models.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving Lmdrive: Closed-loop end-to-end driving with large language models

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.623869Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.623869Z digest=sha256:3a4fb34aee21396a18a7d84fe0e9954bfa86df7be8e52444c778e47155c84078

Observation f6d03268-3e61-40e2-8949-3e88c917dd00 · outbound

This paper cites Passengers’ emotions recognition to improve social acceptance of autonomous driving vehicles.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving Passengers’ emotions recognition to improve social acceptance of autonomous driving vehicles

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.627230Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.627230Z digest=sha256:e94a7f5a65c4713e47f9b93ce6e151bc35f08369501668a0877609af05313afe

Observation fbf56c07-0f16-4392-b907-4f3182384d97 · outbound

This paper cites an unresolved cited work.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving Unresolved cited work

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.630523Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.630523Z digest=sha256:6bb8e3f629861ae90f6e0c00cb6a66dba78ffc01c25e302f974b6b594322199c

Observation d40fc4b8-12d8-4a05-b005-a8fb31a050b6 · outbound

This paper cites INTENT: Trajectory Prediction Framework with Intention-Guided Contrastive Clustering.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving INTENT: Trajectory Prediction Framework with Intention-Guided Contrastive Clustering

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.633960Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.633960Z digest=sha256:94ccdf7315bee17d5208fbb2d292ecb5a270c8853cd0e74dd9b122aace96240f

Observation d9d6a883-fd90-4a5b-9392-20ed204e4cf2 · outbound

This paper cites Itinera: Integrating spatial optimization with large language models for open-domain urban itinerary planning.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving Itinera: Integrating spatial optimization with large language models for open-domain urban itinerary planning

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.638292Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.638292Z digest=sha256:4b653b31a1d6e3d1c470e9377803600651f8167ee587594466ba0e4d6586664a

Observation 30554d7a-e70d-4187-85ce-ff7241a42a1f · outbound

This paper cites Sparkle: Mastering basic spatial ca- pabilities in vision language models elicits generalization to spatial reasoning.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving Sparkle: Mastering basic spatial ca- pabilities in vision language models elicits generalization to spatial reasoning

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.641894Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.641894Z digest=sha256:13d514827b0d33f2f9a5d8318f6ebe0c90d05be0a4d4fcf1382efb486c9875b2

Observation 10022ae3-0336-4784-9de6-11b413574c57 · outbound

This paper cites Qwen2.5: A party of foundation models, 2024.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving Qwen2.5: A party of foundation models, 2024

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.645153Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.645153Z digest=sha256:d210a42cb1a41fb1d879227a5457f9fe73947c12ed1024c4387a1317a4a8cc84

Observation 35b882d3-cab5-4e4e-b808-08a621e0b176 · outbound

This paper cites DriveVLM: The Convergence of Autonomous Driving and Large Vision-Language Models.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving DriveVLM: The Convergence of Autonomous Driving and Large Vision-Language Models

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.648353Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.648353Z digest=sha256:1b5ceb0dc4e7940e952401ded14cf6e5c5ad3cb1a59dc17250361da13eceacfe

Observation c5553fd1-8fd6-4f19-8c6c-fefd928b9fff · outbound

This paper cites Fcos3d: Fully convolutional one-stage monocular 3d object detection.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving Fcos3d: Fully convolutional one-stage monocular 3d object detection

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.652297Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.652297Z digest=sha256:15bce96bfae6a938692884abd9e6770ecf52b335799da91363d67cbe76501151

Observation 7ba7e06a-745a-49c2-993e-e5e6d01981db · outbound

This paper cites Norms of valence, arousal, and dominance for 13,915 english lemmas.Behavior research methods, 45(4):1191–1207, 2013.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving Norms of valence, arousal, and dominance for 13,915 english lemmas.Behavior research methods, 45(4):1191–1207, 2013

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.655497Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.655497Z digest=sha256:c3c748a6532b09a12a313b78f1bc54454141fd8a5c2532d9b4bb74da1fbc551c

Observation 56127d28-9105-43bf-af44-7341389924db · outbound

This paper cites Driver multi-task emotion recognition network based on multi-modal facial video analysis.Pattern Recognition, 161:111241, 2025.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving Driver multi-task emotion recognition network based on multi-modal facial video analysis.Pattern Recognition, 161:111241, 2025

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.658743Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.658743Z digest=sha256:ff72842a17346933cf8748b4944f5dfd8f7c85e4ff9562452bb74d6d68b74cae

Observation d491ecf8-b5e0-4cdb-bbd9-661ba2b095e0 · outbound

This paper cites On-road driver emotion recognition using facial expression.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving On-road driver emotion recognition using facial expression

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.662544Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.662544Z digest=sha256:713bd9ce5adfc3535cbb89de1a1b7d21a04d1a852af78f993d06e0695900f012

Observation a8df3031-c179-468b-a047-002ae3d4dda9 · outbound

This paper cites Openemma: Open-source multimodal model for end-to-end autonomous driving.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving Openemma: Open-source multimodal model for end-to-end autonomous driving

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.666081Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.666081Z digest=sha256:106cffe974fa519683839e51db9768b227f8b20a8d5bdfb38bd364c8b5a7374a

Observation e920f7bb-0f9f-47b3-b03a-e7966eac6164 · outbound

This paper cites Drivegpt4: Interpretable end-to-end autonomous driving via large language model.IEEE Robotics and Automation Letters,.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving Drivegpt4: Interpretable end-to-end autonomous driving via large language model.IEEE Robotics and Automation Letters,

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.669450Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.669450Z digest=sha256:0b6b0096bda1f60957346b9d7dd2f278d96cb3a56b8b0bbbce982177c24467ff

Observation 57368989-594c-4a39-92c1-b4a11fd4e1b7 · outbound

This paper cites Universal instance percep- tion as object discovery and retrieval.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving Universal instance percep- tion as object discovery and retrieval

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.673056Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.673056Z digest=sha256:b769da489bb997e6a2d52da6f448c07423a306dd2e9800ea08f3c5476a8782c9

Observation f248a8d4-07d8-47b3-8598-9cbcfd693439 · outbound

This paper cites Qwen3 Technical Report.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving Qwen3 Technical Report

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.676510Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.676510Z digest=sha256:c8227e174de72ade03aac4dadc4e90a740760a49bce89f7e97266c40fed334bf

Observation ac44f8ca-786f-4e25-8077-9458c741cf1c · outbound

This paper cites Improving visual grounding with visual- linguistic verification and iterative reasoning.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving Improving visual grounding with visual- linguistic verification and iterative reasoning

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.680731Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.680731Z digest=sha256:96dfa4362c8e343ad8054d74b312128be76f8040b0047130f19ae31635d4917f

Observation df25b92d-630e-429d-ab14-b1fb7eb07d31 · outbound

This paper cites Human-centric autonomous systems with llms for user command reasoning.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving Human-centric autonomous systems with llms for user command reasoning

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.683912Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.683912Z digest=sha256:c6f9ee92e88388af969ac2504b4717bbcf3673e1f76792101c07b9e8cc5365a0

Observation 9a493113-acd6-4c4d-9a47-74bd8c9ed12d · outbound

This paper cites DriveMoE: Mixture-of-Experts for Vision-Language-Action Model in End-to-End Autonomous Driving.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving DriveMoE: Mixture-of-Experts for Vision-Language-Action Model in End-to-End Autonomous Driving

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.687072Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.687072Z digest=sha256:19356a68d1f9b7be3abc8096200aa20000c5828c9273d4e2fc058c3fc8836676

Observation c879d654-0663-40b5-8dab-f2bc31a67832 · outbound

This paper cites FutureSightDrive: Thinking Visually with Spatio-Temporal CoT for Autonomous Driving.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving FutureSightDrive: Thinking Visually with Spatio-Temporal CoT for Autonomous Driving

Reference 73

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.691579Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.691579Z digest=sha256:0d40d6774545926dcff369831bcff55b0f177332496d40663d0db909799ca03b

Observation c8bcc00f-788b-4981-8cf3-7d8698cdcf5c · outbound

This paper cites Driver emotion recog- nition for intelligent vehicles: A survey.ACM Computing Surveys (CSUR), 53(3):1–30, 2020.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving Driver emotion recog- nition for intelligent vehicles: A survey.ACM Computing Surveys (CSUR), 53(3):1–30, 2020

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.695032Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.695032Z digest=sha256:36f4f052d81b377e64e9c4d39eaf0529077eff479e3b1895f340a646f97840eb

Observation 16d9c207-6e28-4112-8af1-63f59e52c853 · outbound

This paper cites A comprehensive review: Multisen- sory and cross-cultural approaches to driver emotion modula- tion in vehicle systems.Applied Sciences, 14(15):6819, 2024.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving A comprehensive review: Multisen- sory and cross-cultural approaches to driver emotion modula- tion in vehicle systems.Applied Sciences, 14(15):6819, 2024

Reference 75

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.698581Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.698581Z digest=sha256:0b8fa253e423b0372611c2018f64bb1891de77784a5dd9d2288c839799c8adb6

Observation bb611e13-b34c-4555-98e3-7ddbec524471 · outbound

This paper cites Qwen3 Embedding: Advancing Text Embedding and Reranking Through Foundation Models.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving Qwen3 Embedding: Advancing Text Embedding and Reranking Through Foundation Models

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.702348Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.702348Z digest=sha256:1ff360d513128291b0fc663fbbecdb070f849efa139922dc80d02f602b2331fa

Observation 856c9ebf-1b38-403c-a2ad-acd307f91a9b · outbound

This paper cites Where are you heading? dynamic trajectory prediction with expert goal examples.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving Where are you heading? dynamic trajectory prediction with expert goal examples

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.706019Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.706019Z digest=sha256:31bc7d356cf868588bf25a15c15f359eeb060902b375493bdbfebd5f1fddd465

Observation aa69a70c-2cb0-4265-b9f6-cef9f6ce925a · outbound

This paper cites A survey of autonomous driving from a deep learning perspective.ACM Computing Surveys, 57(10): 1–60, 2025.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving A survey of autonomous driving from a deep learning perspective.ACM Computing Surveys, 57(10): 1–60, 2025

Reference 78

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.709280Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.709280Z digest=sha256:0fa5cc69158653bdb80ac0bb93cd98d35c061b3cf81e7aae6345551a69923be9

Observation bc8fa26b-b8f8-42de-abaf-658eaddfb799 · outbound

This paper cites SWIFT:A Scalable lightWeight Infrastructure for Fine-Tuning.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving SWIFT:A Scalable lightWeight Infrastructure for Fine-Tuning

Reference 79

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.712885Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.712885Z digest=sha256:675c81e42434f39a5b595a2ac45f495b8cdeb5ee2ac276661229fb53f0845d6a

Observation cc32ee54-49b4-4926-87e2-6842f53a717b · outbound

This paper cites Opendrivevla: Towards end-to-end au- tonomous driving with large vision language action model.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving Opendrivevla: Towards end-to-end au- tonomous driving with large vision language action model

Reference 80

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.716364Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.716364Z digest=sha256:d84b4b03242eb5ad5ae051112e9dd43a599227ec5f41da094830bc285b8bd786

Observation f62ba2de-e9a0-453b-86ab-82f5d5704bc7 · outbound

This paper cites Autovla: A vision- language-action model for end-to-end autonomous driving with adaptive reasoning and reinforcement fine-tuning.Nips,.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving Autovla: A vision- language-action model for end-to-end autonomous driving with adaptive reasoning and reinforcement fine-tuning.Nips,

Reference 81

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.719804Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.719804Z digest=sha256:1aa2b1c2be8c842c4b029b6b4e907d5633de717fc8f3f0c7cb3f4d3de8da7a00

Observation 4ab534f4-0210-41bc-b28c-7bfe8b733aee · outbound

This paper cites Exclamation Boost.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving Exclamation Boost

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.723052Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.723052Z digest=sha256:cac519c187fa580ecdf1f83d1d57a035bb884f4799d1559281b11a7c53492870

Pith citing papers

No inbound Pith citation observations are available.