Pith. sign in

Paper Citation Record · LEDGER

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving

As of 9 August 2026, this Paper Citation Record lists 82 of 82 outbound references and 0 inbound Pith citation observations for arXiv:2512.04733.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2512.04733 v3

Coverage vector

measured 82 of 82 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-03T18:37:03.723052Z

measured 82 of 82 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

82 of 82 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved82
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation fda3d3f5-307a-4752-a66d-449ce8e0b3ad · outbound

This paper cites Qwen2.5-VL Technical Report.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving Qwen2.5-VL Technical Report

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:01.908488Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:01.908488Z digest=sha256:77f0a6a5a2bb64a4cb48b90ac48c429ddba08e34eb1941cc86affe410f559bf9

Observation 3930f420-ec89-4bba-a6db-9978dd834442 · outbound

This paper cites Spatial memory: how egocentric and allocentric combine.Trends in cognitive sciences, 10(12):551–557, 2006.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving Spatial memory: how egocentric and allocentric combine.Trends in cognitive sciences, 10(12):551–557, 2006

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:02.005335Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:02.005335Z digest=sha256:0c2333cd027f2fe1c6a833bea6373d7d92ec731d6adc0ee7c1da497fd89e39a5

Observation 3f5eb0eb-53fa-4a73-88f9-2e715c8bba8d · outbound

This paper cites nuscenes: A multi- modal dataset for autonomous driving.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving nuscenes: A multi- modal dataset for autonomous driving

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:02.109624Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:02.109624Z digest=sha256:5d8f0eade5f362aaaa6afeff0fa09c7b168fe96d756e8793e4e3186e56b1e7f0

Observation 78bd2a0e-ab73-42b3-9a4e-9f2fc4e7ce2e · outbound

This paper cites Ground- ing commands for autonomous vehicles via layer fusion with region-specific dynamic layer attention.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving Ground- ing commands for autonomous vehicles via layer fusion with region-specific dynamic layer attention

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:02.219643Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:02.219643Z digest=sha256:b0f7a8cd59279044003817a5e8118ab1dbbbe595da8ed68f7a39c28930e4ad60

Observation bf81f2cd-49be-4c04-a5b5-91d480aac5ee · outbound

This paper cites Eeg-based emotion recognition for road accidents in a simulated driving environment.Biomedical signal processing and control, 87:105411, 2024.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving Eeg-based emotion recognition for road accidents in a simulated driving environment.Biomedical signal processing and control, 87:105411, 2024

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:02.279160Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:02.279160Z digest=sha256:03f61116ca8a9891bfbfed9c10698c85837c07c5cc8deb840bb790e845ffc435

Observation f3119527-444e-4e19-93df-4b76e6b6b1e7 · outbound

This paper cites End-to-end autonomous driving: Challenges and frontiers.IEEE Transactions on Pattern Analysis and Machine Intelligence, 2024.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving End-to-end autonomous driving: Challenges and frontiers.IEEE Transactions on Pattern Analysis and Machine Intelligence, 2024

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:02.437211Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:02.437211Z digest=sha256:e8d93e7738139113daf0ffd6438785884c31555e1698dc400ba8364f74ef828f

Observation e97273e4-b293-4266-ad01-20db6a5b6e70 · outbound

This paper cites Emotion-aware design in automobiles: Embracing technology advancements to enhance human-vehicle interaction.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving Emotion-aware design in automobiles: Embracing technology advancements to enhance human-vehicle interaction

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:02.603358Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:02.603358Z digest=sha256:20b3564a63600e5ca4b0491fd71900fc1846713e0049135aade7b42d75a33dab

Observation 8fead26e-8a64-41b6-a547-2108a5ded26c · outbound

This paper cites Emotion recognition in human-computer interaction.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving Emotion recognition in human-computer interaction

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:02.702226Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:02.702226Z digest=sha256:f91496dcbb80643d3e7d65469bb2973a7186ccaa453b22c63db24c23920d12de

Observation cc90a849-5a7d-4dd2-b404-d27223fd6a0e · outbound

This paper cites A survey on multimodal large language models for autonomous driving.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving A survey on multimodal large language models for autonomous driving

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:02.763541Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:02.763541Z digest=sha256:4588c10ed8ce208bb2808acee03213278351b0dcf684d96f31cfad9ef0be3d38

Observation 019e1247-76ed-4e72-a0e7-ff8b3c932dd0 · outbound

This paper cites Com- mands for autonomous vehicles by progressively stacking visual-linguistic representations.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving Com- mands for autonomous vehicles by progressively stacking visual-linguistic representations

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:02.862493Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:02.862493Z digest=sha256:5ae6a96c3be77be50628a30844e85f140b3fbdd9f6ce5a4b9819dd18bad07c89

Observation 61afb2d4-bf7f-4fd5-8f93-9350630c4f55 · outbound

This paper cites GoEmotions: A Dataset of Fine-Grained Emotions.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving GoEmotions: A Dataset of Fine-Grained Emotions

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:02.966502Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:02.966502Z digest=sha256:6df1e5adb56bf9ee73e358f1de4667a456572023ffd95313d66ff3c778b18329

Observation 0310f769-0b0b-42aa-92c6-fca76afdb989 · outbound

This paper cites Goal- gan: Multimodal trajectory prediction based on goal position estimation.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving Goal- gan: Multimodal trajectory prediction based on goal position estimation

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.084349Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.084349Z digest=sha256:3d976f8498526b67ff9b55783c607f249022b4d3ba8e0bc7a684452ede8df69c

Observation 027ec76e-9250-40bb-8e5b-4ccc7bcbfcc3 · outbound

This paper cites Transvg: End-to-end visual ground- ing with transformers.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving Transvg: End-to-end visual ground- ing with transformers

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.190418Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.190418Z digest=sha256:7ef8e3f27e35c62afbf1e60aa7e0e1028f8de6c3849b753a0da8b27eb0458fa5

Observation 7cf5a902-9b0c-4552-bd15-1edc167ef5fe · outbound

This paper cites Talk2Car: Taking Control of Your Self-Driving Car.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving Talk2Car: Taking Control of Your Self-Driving Car

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.277319Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.277319Z digest=sha256:8d254c76be49254ac852f80b6b14f350055879364632bdeaad1f14f235ffd486

Observation 48ec2828-0c7e-4caf-9bf0-c45a3fdf22e8 · outbound

This paper cites Talk2car: Predicting physical trajectories for natural language commands.Ieee Access, 10: 123809–123834, 2022.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving Talk2car: Predicting physical trajectories for natural language commands.Ieee Access, 10: 123809–123834, 2022

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.383485Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.383485Z digest=sha256:171d4379c4be3fb568dc151e0e413579fb74ac0462f0b5e8ee05e3f43ea6937c

Observation 6ee6a772-a3a3-4b65-bc8f-0dcce87a525b · outbound

This paper cites Bert: Pre-training of deep bidirectional transform- ers for language understanding.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving Bert: Pre-training of deep bidirectional transform- ers for language understanding

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.387154Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.387154Z digest=sha256:261871c2e3aeb335041acca12a943ea737738910a3fce06b5d4afbef9ab4b420

Observation e2bb7b4e-4696-41ed-9ac7-5d6e908d580f · outbound

This paper cites Multi-class emotion recognition within the valence-arousal-dominance space using eeg.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving Multi-class emotion recognition within the valence-arousal-dominance space using eeg

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.390464Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.390464Z digest=sha256:d35d65c54704e3f4c9bfa02d9dde3f06c74eea78c5022702c8938fa12f45b428

Observation 5199d07d-9a41-4612-bf56-e1efe1898d34 · outbound

This paper cites Predicting physical world destina- tions for commands given to self-driving cars.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving Predicting physical world destina- tions for commands given to self-driving cars

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.393633Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.393633Z digest=sha256:905ecc2bc72fe3ba9ddaf45b49be1551d74e341d5f77f633a59eb508ac604329

Observation 6465536a-bedf-4af5-9dfe-15ca353abc0e · outbound

This paper cites Think before you drive: World model-inspired multimodal grounding for autonomous driving.arXiv preprint, 2025.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving Think before you drive: World model-inspired multimodal grounding for autonomous driving.arXiv preprint, 2025

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.397038Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.397038Z digest=sha256:710f9d83eadaf202413395c8dc0539effbb6e3e2a421b06f52585840ee56edf6

Observation 7c553ee8-589a-4aec-80f7-c5a42bde8ed6 · outbound

This paper cites A formal basis for the heuristic determination of minimum cost paths.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving A formal basis for the heuristic determination of minimum cost paths

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.400130Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.400130Z digest=sha256:9899127f1455530c81d5418992986e366001f5c28a4b789679e9814e3001845f

Observation 28f7a811-872f-4ac2-8882-e298d864d099 · outbound

This paper cites Hu, Yelong Shen, Phillip Wallis, Zeyuan Allen- Zhu, Yuanzhi Li, Shean Wang, Lu Wang, and Weizhu Chen.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving Hu, Yelong Shen, Phillip Wallis, Zeyuan Allen- Zhu, Yuanzhi Li, Shean Wang, Lu Wang, and Weizhu Chen

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.403382Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.403382Z digest=sha256:28653e217073fa272cf01c8a44cd1b66dd8281e549f03141147a8f73ee4a0575

Observation 0d760349-5e9f-4050-94c3-ecfe12390898 · outbound

This paper cites Planning-oriented autonomous driving.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving Planning-oriented autonomous driving

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.406493Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.406493Z digest=sha256:57cab61b20c292806a60c164c8b62291ae260e11aae45efc101b569751a2f9bd

Observation a54a34c2-68c5-46c2-9b6b-f432381fbf07 · outbound

This paper cites Think twice before driving: Towards scalable decoders for end-to-end autonomous driving.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving Think twice before driving: Towards scalable decoders for end-to-end autonomous driving

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.409654Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.409654Z digest=sha256:a3d0c9033c797e44b2c87cb404204089fff6048ed035786a4f3fba9fa150362d

Observation eec9da9d-0c22-4260-bd8e-0d1437df1798 · outbound

This paper cites Senna: Bridging Large Vision-Language Models and End-to-End Autonomous Driving.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving Senna: Bridging Large Vision-Language Models and End-to-End Autonomous Driving

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.412774Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.412774Z digest=sha256:f711ed4f9f3f6e1cd315273d154722f66f5fef8f4a93805775751f003dd60b5d

Observation 92d6ba2c-4203-42ed-9441-008dca2c40cc · outbound

This paper cites A Survey on Vision-Language-Action Models for Autonomous Driving.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving A Survey on Vision-Language-Action Models for Autonomous Driving

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.416823Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.416823Z digest=sha256:6b64683f0877df243cba89f73d3be93308bd6af33cf214a1dc53230c67f4543b

Observation 5b53327a-e621-4924-8a82-ebbf3472d915 · outbound

This paper cites Mdetr-modulated detection for end-to-end multi-modal understanding.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving Mdetr-modulated detection for end-to-end multi-modal understanding

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.420264Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.420264Z digest=sha256:65bee3b857ae7d61d5616a756e550f2fa21d7b82cc7cab4cf9338d3b74956d8c

Observation 4256fc99-65fb-4313-9b5a-6b539ee99f10 · outbound

This paper cites Detection of drivers’ anxiety invoked by driving situations using multimodal biosignals.Processes, 8 (2):155, 2020.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving Detection of drivers’ anxiety invoked by driving situations using multimodal biosignals.Processes, 8 (2):155, 2020

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.423509Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.423509Z digest=sha256:653624d4c5c2fabbdb1059515334a6a47854b2918e28eeb4806284a7c32f63ad

Observation fb6447ce-1f1f-45a2-88ad-1a2061d42b28 · outbound

This paper cites Review and perspectives on human emotion for connected automated vehicles.Automotive Innovation, 7(1):4–44, 2024.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving Review and perspectives on human emotion for connected automated vehicles.Automotive Innovation, 7(1):4–44, 2024

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.426705Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.426705Z digest=sha256:d92080669c2a7d62263c685438152224629d22396bd5522c4856bce11f2c3dda

Observation 41e42b77-c10a-46ce-927d-6d54987d2368 · outbound

This paper cites DriveVLA-W0: World Models Amplify Data Scaling Law in Autonomous Driving.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving DriveVLA-W0: World Models Amplify Data Scaling Law in Autonomous Driving

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.430100Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.430100Z digest=sha256:a626e2f5519ecddb7f7e33f17d559622867460198c034efbdefc1572da600f2e

Observation abfb8096-6550-492a-8eb4-3cd4962bc0d3 · outbound

This paper cites Fine-grained evaluation of large vision-language mod- els in autonomous driving.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving Fine-grained evaluation of large vision-language mod- els in autonomous driving

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.433519Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.433519Z digest=sha256:b507e436d1a81841dffb2938e64c597a823d1d65f049474b4607672206f278be

Observation 592f17f5-0761-4c59-808a-a331ad8325ad · outbound

This paper cites ReCogDrive: A Reinforced Cognitive Framework for End-to-End Autonomous Driving.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving ReCogDrive: A Reinforced Cognitive Framework for End-to-End Autonomous Driving

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.436654Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.436654Z digest=sha256:3ca372e266a8970d7de2385ff6425063dd8960e4e98f7abd37e1958d0594c88c

Observation 164db171-1c2c-4cef-98a4-e2e067982719 · outbound

This paper cites Steering the future: Redefining intelligent transportation systems with foundation models.Chain, 1(1):46–53, 2024.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving Steering the future: Redefining intelligent transportation systems with foundation models.Chain, 1(1):46–53, 2024

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.440641Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.440641Z digest=sha256:1aa6165ed0ca4d8c7c3080d58b6da9bb7682d467de3855780e8d15176f2a3775

Observation 8ef1f9c2-1860-4c95-98c1-d937321234c6 · outbound

This paper cites Mamba-va: A mamba-based approach for continuous emotion recognition in valence-arousal space.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving Mamba-va: A mamba-based approach for continuous emotion recognition in valence-arousal space

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.444039Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.444039Z digest=sha256:dd33c742c818ddc7bc1c78748c23b8708bc558fccb72cda1a977bb11c7a992da

Observation 4c9c1a28-1530-4ffb-aa5c-e9ad2595e409 · outbound

This paper cites Gpt- 4 enhanced multimodal grounding for autonomous driving: Leveraging cross-modal attention with large language models.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving Gpt- 4 enhanced multimodal grounding for autonomous driving: Leveraging cross-modal attention with large language models

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.447243Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.447243Z digest=sha256:271a87228f51138702291cd438099b25390de5acdb2c971fd25dfe14982b035d

Observation c6d9346f-f898-48e6-a30c-4257c43dfce9 · outbound

This paper cites Cot-drive: Efficient motion forecasting for autonomous driving with llms and chain-of-thought prompting.IEEE Transactions on Artificial Intelligence, 2025.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving Cot-drive: Efficient motion forecasting for autonomous driving with llms and chain-of-thought prompting.IEEE Transactions on Artificial Intelligence, 2025

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.450634Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.450634Z digest=sha256:0b05683fe1116ae7e9900f09ee74d6b45c7e947863a74332384c85d1a76ce65a

Observation a3ff138e-769d-4c4e-b876-1169c5e1149e · outbound

This paper cites Toward human-like trajectory predic- tion for autonomous driving: A behavior-centric approach.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving Toward human-like trajectory predic- tion for autonomous driving: A behavior-centric approach

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.453786Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.453786Z digest=sha256:e99ab8b3a25e10672220de3723526240d5a5ee87d7fb69d51c8c49f1d59d0a09

Observation 60d2c7b5-60f5-452e-be8e-3e76368f64ae · outbound

This paper cites CogDriver: Integrating Cognitive Inertia for Temporally Coherent Planning in Autonomous Driving.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving CogDriver: Integrating Cognitive Inertia for Temporally Coherent Planning in Autonomous Driving

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.457803Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.457803Z digest=sha256:1262e47713c66d9d9856cf8ece91c41517feb2c2172c7ca9b9a4aeae9a434f43

Observation 59161342-bc84-4f12-a823-d1ec6ea125fe · outbound

This paper cites Grounding DINO: Marrying DINO with Grounded Pre-Training for Open-Set Object Detection.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving Grounding DINO: Marrying DINO with Grounded Pre-Training for Open-Set Object Detection

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.461246Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.461246Z digest=sha256:a5b163979a718f3798dfe61c21c03b55ebf2d2dd526c2167427fb0297c6831c1

Observation 0fc9e669-8ccd-481f-95d9-960a95edf6c5 · outbound

This paper cites RoBERTa: A Robustly Optimized BERT Pretraining Approach.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving RoBERTa: A Robustly Optimized BERT Pretraining Approach

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.464784Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.464784Z digest=sha256:fa8c3d919724e1db108988a14d7b6fea59349ee3fc0f5f42fbc15ee8d818ab73

Observation 2f771d60-f432-4a2a-8fec-9437ea2b5537 · outbound

This paper cites C4av: learning cross-modal representations from transformers.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving C4av: learning cross-modal representations from transformers

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.468365Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.468365Z digest=sha256:e24b3e441b503fbbf913f0e5ecc5d2eb8adefa0f70775ec96dd61c36e389acf9

Observation d0054764-b0eb-4d7b-b6ab-4420c5b7d1aa · outbound

This paper cites From goals, waypoints & paths to long term human trajectory forecasting.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving From goals, waypoints & paths to long term human trajectory forecasting

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.471618Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.471618Z digest=sha256:dd4bd932c6189e3923fbf11f8e6e796051b99fc58e463cdde06b1a53afc1848d

Observation a9b7a4fd-71aa-4af6-980b-00e31e5a78fe · outbound

This paper cites Person- ality correlates of driver stress.Personality and Individual Differences, 12(6):535–549, 1991.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving Person- ality correlates of driver stress.Personality and Individual Differences, 12(6):535–549, 1991

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.475111Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.475111Z digest=sha256:bc201a39b685c74648f3cdc098c722393313b07913997cd803cb6480e1c47972

Observation 05f04bb3-68ce-462f-aa69-0a66b0183311 · outbound

This paper cites Attngrounder: Talking to cars with attention.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving Attngrounder: Talking to cars with attention

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.478914Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.478914Z digest=sha256:3a16bd18f1c9154ee2df33cc88ba6b5eb734f367ab56b5f8b78976c32676f7a3

Observation 03cdd1aa-9a89-410a-8288-1e5f388a91bc · outbound

This paper cites Mohammad.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving Mohammad

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.482355Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.482355Z digest=sha256:e20c4ea42c180edd78a06a6b624434a97e62a1e8e11bd694c9a5952768ce4ab5

Observation 49629f17-eb20-4156-b46b-60bee485f144 · outbound

This paper cites Driver emotion recognition with a hybrid attentional multimodal fusion framework.IEEE Transactions on Affective Computing, 14(4):2970–2981, 2023.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving Driver emotion recognition with a hybrid attentional multimodal fusion framework.IEEE Transactions on Affective Computing, 14(4):2970–2981, 2023

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.485750Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.485750Z digest=sha256:ddfedca6284b35736b0197f6b0582ae526a5b6d2e75c8c464f7ff8ca5742365c

Observation c2ade741-bd71-42f6-9a97-e89e143b69c6 · outbound

This paper cites Steerable Adversarial Scenario Generation through Test-Time Preference Alignment.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving Steerable Adversarial Scenario Generation through Test-Time Preference Alignment

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.489316Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.489316Z digest=sha256:5552400774c31cc71dbc6cc01d27f605464192d84ee21d7e2d89c3c16dfe1c82

Observation ff74b0dc-11e7-4fb4-8cf7-ceefdba0af67 · outbound

This paper cites Vlp: Vision language planning for autonomous driving.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving Vlp: Vision language planning for autonomous driving

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.493431Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.493431Z digest=sha256:153ef8893d311af4813346bb3996fe066fc03a9dd204d8556d905f186876a93c

Observation 20c430bb-5f4c-4996-bb4f-c57b5bb94b18 · outbound

This paper cites Multi- modal fusion transformer for end-to-end autonomous driving.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving Multi- modal fusion transformer for end-to-end autonomous driving

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.496826Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.496826Z digest=sha256:a888b3c42bb91a98c41df7b7099de144a31d65f602ff0977bdda63da644bf25d

Observation 4e999f0f-b598-474e-a806-35ba6c449071 · outbound

This paper cites Direct prefer- ence optimization: Your language model is secretly a reward model.Advances in neural information processing systems, 36:53728–53741, 2023.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving Direct prefer- ence optimization: Your language model is secretly a reward model.Advances in neural information processing systems, 36:53728–53741, 2023

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.500524Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.500524Z digest=sha256:b50cf9eb1e91fa23146f2448d6fa475019c3296d1df9ecc81d3f156473f76616

Observation 5b2652a7-24d8-4252-ad2a-ef83a9f1461b · outbound

This paper cites Simlingo: Vision-only closed-loop autonomous driving with language-action alignment.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving Simlingo: Vision-only closed-loop autonomous driving with language-action alignment

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.503797Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.503797Z digest=sha256:f13dd4214cff0a63fa70c82b5df2e02d9bd9810b5a125d7001af0cccf8712b07

Observation 256c2928-f6fe-44d5-9df8-d866e39e3abc · outbound

This paper cites Cosine meets softmax: A tough-to-beat baseline for visual grounding.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving Cosine meets softmax: A tough-to-beat baseline for visual grounding

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.612115Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.612115Z digest=sha256:3d597baa9840f0c846b059f73622e58f3bd2ddd0ff6cfff6d356081541965539

Observation abe76bdb-0c96-41cf-a1a9-4d095699d7ca · outbound

This paper cites DistilBERT, a distilled version of BERT: smaller, faster, cheaper and lighter.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving DistilBERT, a distilled version of BERT: smaller, faster, cheaper and lighter

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.616071Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.616071Z digest=sha256:5d194a397d0f060387db5decee15f8c35e6a579da8c6871fb6cd8dcef80146e0

Observation a842e73d-7073-4107-b7f5-8bc6be2131ab · outbound

This paper cites Vision-language-action models: Con- cepts, progress, applications and challenges.arXiv preprint arXiv:2505.04769, 2025.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving Vision-language-action models: Con- cepts, progress, applications and challenges.arXiv preprint arXiv:2505.04769, 2025

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.620338Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.620338Z digest=sha256:d0a863ce79295c68c928c839240dcfe008db2a2447efdd315e05cc7adfc4e957

Observation d67903a5-2881-420e-95f2-bdf2fe0ec06d · outbound

This paper cites Lmdrive: Closed-loop end-to-end driving with large language models.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving Lmdrive: Closed-loop end-to-end driving with large language models

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.623869Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.623869Z digest=sha256:81fcac4a58d58480ea11cedc1cfdae37a4cf98ae8a5866ad332453d8b460c7ba

Observation f6d03268-3e61-40e2-8949-3e88c917dd00 · outbound

This paper cites Passengers’ emotions recognition to improve social acceptance of autonomous driving vehicles.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving Passengers’ emotions recognition to improve social acceptance of autonomous driving vehicles

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.627230Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.627230Z digest=sha256:dc457d3a5be40f80dd7840083a4ef0077739627f7bd4d690c4932767491e51ab

Observation fbf56c07-0f16-4392-b907-4f3182384d97 · outbound

This paper cites an unresolved cited work.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving Unresolved cited work

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.630523Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.630523Z digest=sha256:6b5dae7af2c8fa5cbbbafff569ce453cae11c497f19a89426059c55bb6aa965b

Observation d40fc4b8-12d8-4a05-b005-a8fb31a050b6 · outbound

This paper cites INTENT: Trajectory Prediction Framework with Intention-Guided Contrastive Clustering.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving INTENT: Trajectory Prediction Framework with Intention-Guided Contrastive Clustering

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.633960Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.633960Z digest=sha256:8d35f3112768055c570f4421883829a00e1649c5ab9bea135ed15f3823eab69c

Observation d9d6a883-fd90-4a5b-9392-20ed204e4cf2 · outbound

This paper cites Itinera: Integrating spatial optimization with large language models for open-domain urban itinerary planning.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving Itinera: Integrating spatial optimization with large language models for open-domain urban itinerary planning

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.638292Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.638292Z digest=sha256:c111b6eb336d9a6119b60c3f7b093ed570f8ddf9c48d3123a9385bba8586160b

Observation 30554d7a-e70d-4187-85ce-ff7241a42a1f · outbound

This paper cites Sparkle: Mastering basic spatial ca- pabilities in vision language models elicits generalization to spatial reasoning.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving Sparkle: Mastering basic spatial ca- pabilities in vision language models elicits generalization to spatial reasoning

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.641894Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.641894Z digest=sha256:fc5ec7b668981a47bab5f3ed867b6bfaed7463394d2b736c0531575e1bc2da90

Observation 10022ae3-0336-4784-9de6-11b413574c57 · outbound

This paper cites Qwen2.5: A party of foundation models, 2024.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving Qwen2.5: A party of foundation models, 2024

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.645153Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.645153Z digest=sha256:e53060f3fd67e4fe5521325a13e0da3f236158b4e1566420337cf4c504d3bdfc

Observation 35b882d3-cab5-4e4e-b808-08a621e0b176 · outbound

This paper cites DriveVLM: The Convergence of Autonomous Driving and Large Vision-Language Models.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving DriveVLM: The Convergence of Autonomous Driving and Large Vision-Language Models

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.648353Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.648353Z digest=sha256:d573188aac961bfb7719859befd751d158db8d58d6f3e610aaa101569d2b2f4b

Observation c5553fd1-8fd6-4f19-8c6c-fefd928b9fff · outbound

This paper cites Fcos3d: Fully convolutional one-stage monocular 3d object detection.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving Fcos3d: Fully convolutional one-stage monocular 3d object detection

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.652297Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.652297Z digest=sha256:73cfc1a6decfff3f8e5b1a4b180409feeb45dd21a033c8644d7f9f77f11b9dc7

Observation 7ba7e06a-745a-49c2-993e-e5e6d01981db · outbound

This paper cites Norms of valence, arousal, and dominance for 13,915 english lemmas.Behavior research methods, 45(4):1191–1207, 2013.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving Norms of valence, arousal, and dominance for 13,915 english lemmas.Behavior research methods, 45(4):1191–1207, 2013

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.655497Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.655497Z digest=sha256:3babd0908bc0c9a37c8a66474eda29c44ec4069c7d2371606b5363586766a64c

Observation 56127d28-9105-43bf-af44-7341389924db · outbound

This paper cites Driver multi-task emotion recognition network based on multi-modal facial video analysis.Pattern Recognition, 161:111241, 2025.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving Driver multi-task emotion recognition network based on multi-modal facial video analysis.Pattern Recognition, 161:111241, 2025

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.658743Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.658743Z digest=sha256:d18509d6cba9453c5ab0a728c1cb892f2c304b3fa91493a5c677430733120a86

Observation d491ecf8-b5e0-4cdb-bbd9-661ba2b095e0 · outbound

This paper cites On-road driver emotion recognition using facial expression.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving On-road driver emotion recognition using facial expression

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.662544Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.662544Z digest=sha256:58dec9032c52f15d826a9cdc6cbe441bb31559a12203864e1b94405238f68b9c

Observation a8df3031-c179-468b-a047-002ae3d4dda9 · outbound

This paper cites Openemma: Open-source multimodal model for end-to-end autonomous driving.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving Openemma: Open-source multimodal model for end-to-end autonomous driving

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.666081Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.666081Z digest=sha256:0f148199642fb52ed0e57a98d32de5c9d46213dc265cc047762fdd52bc8db727

Observation e920f7bb-0f9f-47b3-b03a-e7966eac6164 · outbound

This paper cites Drivegpt4: Interpretable end-to-end autonomous driving via large language model.IEEE Robotics and Automation Letters,.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving Drivegpt4: Interpretable end-to-end autonomous driving via large language model.IEEE Robotics and Automation Letters,

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.669450Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.669450Z digest=sha256:58078e6e4b2845174dc627c732737f229545217e13961c055ccf997647217a87

Observation 57368989-594c-4a39-92c1-b4a11fd4e1b7 · outbound

This paper cites Universal instance percep- tion as object discovery and retrieval.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving Universal instance percep- tion as object discovery and retrieval

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.673056Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.673056Z digest=sha256:acfa2ddb5fe9f8f349fe948a3c8d5c67404ad2d58a8f6f947cbf84c572a6913a

Observation f248a8d4-07d8-47b3-8598-9cbcfd693439 · outbound

This paper cites Qwen3 Technical Report.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving Qwen3 Technical Report

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.676510Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.676510Z digest=sha256:7aa073899449c9b09106600006e3baea083cf1cbedbcaa547a3cfd2709ab7ed7

Observation ac44f8ca-786f-4e25-8077-9458c741cf1c · outbound

This paper cites Improving visual grounding with visual- linguistic verification and iterative reasoning.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving Improving visual grounding with visual- linguistic verification and iterative reasoning

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.680731Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.680731Z digest=sha256:186f0161fa8f8fa1695c6c8e098a17d79cee049ba5d47111b4f6c77f73713b59

Observation df25b92d-630e-429d-ab14-b1fb7eb07d31 · outbound

This paper cites Human-centric autonomous systems with llms for user command reasoning.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving Human-centric autonomous systems with llms for user command reasoning

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.683912Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.683912Z digest=sha256:8a9a0b6535a669b809a1d8d17dbbecb78ee8822a219f26b873154b6e6b718719

Observation 9a493113-acd6-4c4d-9a47-74bd8c9ed12d · outbound

This paper cites DriveMoE: Mixture-of-Experts for Vision-Language-Action Model in End-to-End Autonomous Driving.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving DriveMoE: Mixture-of-Experts for Vision-Language-Action Model in End-to-End Autonomous Driving

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.687072Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.687072Z digest=sha256:d1b09e4d62aa879689e07bb17d506f8f69320eb6b7676c53bab01aa9e85332fb

Observation c879d654-0663-40b5-8dab-f2bc31a67832 · outbound

This paper cites FutureSightDrive: Thinking Visually with Spatio-Temporal CoT for Autonomous Driving.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving FutureSightDrive: Thinking Visually with Spatio-Temporal CoT for Autonomous Driving

Reference 73

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.691579Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.691579Z digest=sha256:a7a5d96a142dd3e45b141de2c5934823ce69ac305af0ae982077c4c5be266dcd

Observation c8bcc00f-788b-4981-8cf3-7d8698cdcf5c · outbound

This paper cites Driver emotion recog- nition for intelligent vehicles: A survey.ACM Computing Surveys (CSUR), 53(3):1–30, 2020.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving Driver emotion recog- nition for intelligent vehicles: A survey.ACM Computing Surveys (CSUR), 53(3):1–30, 2020

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.695032Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.695032Z digest=sha256:f1f4b364349fab3adbbb0323a41cdc3a77cc0a077f69f24ddba287d2992cf543

Observation 16d9c207-6e28-4112-8af1-63f59e52c853 · outbound

This paper cites A comprehensive review: Multisen- sory and cross-cultural approaches to driver emotion modula- tion in vehicle systems.Applied Sciences, 14(15):6819, 2024.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving A comprehensive review: Multisen- sory and cross-cultural approaches to driver emotion modula- tion in vehicle systems.Applied Sciences, 14(15):6819, 2024

Reference 75

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.698581Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.698581Z digest=sha256:68338ae4b0b559d77aa78213c1281f830e536ec6d319ad5bec435855af644b5e

Observation bb611e13-b34c-4555-98e3-7ddbec524471 · outbound

This paper cites Qwen3 Embedding: Advancing Text Embedding and Reranking Through Foundation Models.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving Qwen3 Embedding: Advancing Text Embedding and Reranking Through Foundation Models

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.702348Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.702348Z digest=sha256:a4cb363f8c2977e522dd63b3b24857982a74682f62eac5ae8016eabe637ef10f

Observation 856c9ebf-1b38-403c-a2ad-acd307f91a9b · outbound

This paper cites Where are you heading? dynamic trajectory prediction with expert goal examples.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving Where are you heading? dynamic trajectory prediction with expert goal examples

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.706019Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.706019Z digest=sha256:adec072c4d2aa37657df73a98e511bea758583d886db54f87f11c20fc77391b5

Observation aa69a70c-2cb0-4265-b9f6-cef9f6ce925a · outbound

This paper cites A survey of autonomous driving from a deep learning perspective.ACM Computing Surveys, 57(10): 1–60, 2025.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving A survey of autonomous driving from a deep learning perspective.ACM Computing Surveys, 57(10): 1–60, 2025

Reference 78

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.709280Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.709280Z digest=sha256:5175ea5e7696cc51946500f45831f0a5e3e632c13f6b9d80b236e0c8d42b3f0e

Observation bc8fa26b-b8f8-42de-abaf-658eaddfb799 · outbound

This paper cites SWIFT:A Scalable lightWeight Infrastructure for Fine-Tuning.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving SWIFT:A Scalable lightWeight Infrastructure for Fine-Tuning

Reference 79

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.712885Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.712885Z digest=sha256:87cd3d9e699b7f82a11921149992892f178c975777dd2f8083a633bc2f07a981

Observation cc32ee54-49b4-4926-87e2-6842f53a717b · outbound

This paper cites Opendrivevla: Towards end-to-end au- tonomous driving with large vision language action model.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving Opendrivevla: Towards end-to-end au- tonomous driving with large vision language action model

Reference 80

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.716364Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.716364Z digest=sha256:17b5e61508bd9ef28b6cba735305cd6b4e6da9a3e206748985a74fd287973191

Observation f62ba2de-e9a0-453b-86ab-82f5d5704bc7 · outbound

This paper cites Autovla: A vision- language-action model for end-to-end autonomous driving with adaptive reasoning and reinforcement fine-tuning.Nips,.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving Autovla: A vision- language-action model for end-to-end autonomous driving with adaptive reasoning and reinforcement fine-tuning.Nips,

Reference 81

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.719804Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.719804Z digest=sha256:ae74b1dc5edccef9c195ecffc2b2fb218702cf39fb3035465fd051e03c10dbc8

Observation 4ab534f4-0210-41bc-b28c-7bfe8b733aee · outbound

This paper cites Exclamation Boost.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving Exclamation Boost

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.723052Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.723052Z digest=sha256:1cf12bbe26904ab96124bebcf0f2ee8e23b4a2906d5cb4a2e0140fed31fc0ef7

Pith citing papers

No inbound Pith citation observations are available.