Pith. sign in

Paper Citation Record · LEDGER

A Survey on Vision-Language-Action Models for Autonomous Driving

As of 15 August 2026, this Paper Citation Record lists 100 of 166 outbound references and 14 inbound Pith citation observations for arXiv:2506.24044.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.24044 v1

Coverage vector

measured 100 of 166 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T21:31:04.674245Z

measured 114 of 114 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00

measured 14 of 14 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T12:39:02.315303Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-05T11:41:02.543365Z

Reference resolution

100 of 166 outbound references displayed

  • verified exact1
  • verified fuzzy0
  • unresolved99
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 95972b4e-af7f-4426-b799-a6a82ba1c909 · outbound

This paper cites Flamingo: a visual language model for few-shot learn- ing.

A Survey on Vision-Language-Action Models for Autonomous Driving Flamingo: a visual language model for few-shot learn- ing

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.317875Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.317875Z digest=sha256:e4203713824423f481c5dbe8817a5fff5c781ef4898389f28b404349ee2d4ffa

Observation 9683c8e4-9fa6-4ba5-b4de-b343ce145803 · outbound

This paper cites An lstm net- work for highway trajectory prediction.

A Survey on Vision-Language-Action Models for Autonomous Driving An lstm net- work for highway trajectory prediction

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.321970Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.321970Z digest=sha256:876532a75d2d42979fbe6761bcbb0053e8e648d215b5bc532df6f3be1dae47ef

Observation 1d2f59fb-38cf-4392-b6ad-e788b6dbb978 · outbound

This paper cites VaViM and VaVAM: Autonomous Driving through Video Generative Modeling.

A Survey on Vision-Language-Action Models for Autonomous Driving VaViM and VaVAM: Autonomous Driving through Video Generative Modeling

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.325481Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.325481Z digest=sha256:d0f39f33975edc05d4fab8df75844da0c4745d76101bcdbad716a247c92f91a5

Observation 3be9daf3-30aa-47f2-8f53-03ea0dfa3524 · outbound

This paper cites $\pi_0$: A Vision-Language-Action Flow Model for General Robot Control.

A Survey on Vision-Language-Action Models for Autonomous Driving $\pi_0$: A Vision-Language-Action Flow Model for General Robot Control

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.329364Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.329364Z digest=sha256:7c856d55fffe74460cf8cda0f2aff4fab4f64d6d48292908381c4c31e4f7fee8

Observation a940b45e-a0dc-4940-be05-64a3d59804c0 · outbound

This paper cites Fine-grained affective processing capabilities emerging from large lan- guage models.

A Survey on Vision-Language-Action Models for Autonomous Driving Fine-grained affective processing capabilities emerging from large lan- guage models

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.333208Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.333208Z digest=sha256:ecbe42594cb5830aad62bc3989d8d95b0757cc2cfc073fa056e6d7e1abc129ba

Observation 0c255595-3398-4239-bed9-78afb5ad81ef · outbound

This paper cites Language models are few-shot learners.

A Survey on Vision-Language-Action Models for Autonomous Driving Language models are few-shot learners

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.336827Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.336827Z digest=sha256:3ee1798e8088793f510cf7a4c885d8f33a98ce8a7196e4035936355a0f62ea48

Observation eddffb62-ec8c-4c99-9782-f80d4cdc0146 · outbound

This paper cites nuscenes: A multimodal dataset for autonomous driving.

A Survey on Vision-Language-Action Models for Autonomous Driving nuscenes: A multimodal dataset for autonomous driving

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.341206Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.341206Z digest=sha256:e3e86a780f3dfbff32484daa6ef00589f685ea76a35e4d99d644eb1b51890273

Observation 7fa2c5dd-9bfd-4e66-9c0b-c682eeeca2fa · outbound

This paper cites nuscenes: A mul- timodal dataset for autonomous driving.

A Survey on Vision-Language-Action Models for Autonomous Driving nuscenes: A mul- timodal dataset for autonomous driving

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.344745Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.344745Z digest=sha256:06d5f2fbc3ff17d72d2dfe7314614c61ba654ef5e087e34288c0286000d56080

Observation 87ae5ca0-c84f-4946-9287-e7bf4952b552 · outbound

This paper cites Learning from all vehi- cles.

A Survey on Vision-Language-Action Models for Autonomous Driving Learning from all vehi- cles

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.349406Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.349406Z digest=sha256:c9ef08ae3c2895fb4b5b4f9343c3decf39a6552b8b9024dfa358b7d7c2bda51f

Observation 15618d1e-cae1-4536-a0a4-4ad30489a0fb · outbound

This paper cites Insight: Enhancing autonomous driving safety through vision-language models on context-aware hazard detection and edge case evaluation.

A Survey on Vision-Language-Action Models for Autonomous Driving Insight: Enhancing autonomous driving safety through vision-language models on context-aware hazard detection and edge case evaluation

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.352905Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.352905Z digest=sha256:67a285d0dca0da9fc3d65fec8f095dcc20d3a6e6f4290161a28ea0eabfd9163b

Observation 88023e48-d173-4a4c-a9d3-d4966c1b92f8 · outbound

This paper cites TS-VLM: Text-Guided SoftSort Pooling for Vision-Language Models in Multi-View Driving Reasoning.

A Survey on Vision-Language-Action Models for Autonomous Driving TS-VLM: Text-Guided SoftSort Pooling for Vision-Language Models in Multi-View Driving Reasoning

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.357239Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.357239Z digest=sha256:1c4ee17064736aa4bf919f57de6d600cab60c332838c9bd039b9f74b9aa6fea5

Observation 4e41fc49-42a9-42da-9ee8-3ce930c479e6 · outbound

This paper cites What data do we need for training an av motion planner? In 2021 IEEE Inter- national Conference on Robotics and Automation (ICRA) , pages 1066–1072.

A Survey on Vision-Language-Action Models for Autonomous Driving What data do we need for training an av motion planner? In 2021 IEEE Inter- national Conference on Robotics and Automation (ICRA) , pages 1066–1072

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.361431Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.361431Z digest=sha256:8c7facd27f113294370bdc4bf3170edf387241e66942cdb592dde2f5f3a77beb

Observation b5cd6e6b-7abf-4712-a9a2-783a5470fee6 · outbound

This paper cites End-to-end autonomous driving: Challenges and frontiers.

A Survey on Vision-Language-Action Models for Autonomous Driving End-to-end autonomous driving: Challenges and frontiers

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.364887Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.364887Z digest=sha256:b50fc8e58268342188c3378ae391deae3c09738181c5e0cb68fd3865e93f7550

Observation f48e424e-dd98-4372-b145-c8ce827bcbc7 · outbound

This paper cites VADv2: End-to-End Vectorized Autonomous Driving via Probabilistic Planning.

A Survey on Vision-Language-Action Models for Autonomous Driving VADv2: End-to-End Vectorized Autonomous Driving via Probabilistic Planning

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.369131Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.369131Z digest=sha256:29a06417c7bc7a4b48328a2e41d969c1be0725992cd71d45bfe8da66ad9c9aeb

Observation 1d70c450-28a6-4dcc-8af5-14c10c1389aa · outbound

This paper cites Asynchronous large language model en- hanced planner for autonomous driving.

A Survey on Vision-Language-Action Models for Autonomous Driving Asynchronous large language model en- hanced planner for autonomous driving

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.373472Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.373472Z digest=sha256:2e57890d7005b99c3fa738ce6cc430786350bf3f870916518d3dcbbc97c514cc

Observation 206166b7-1c33-4c8c-971f-57f8bf472cfb · outbound

This paper cites Ppad: Iterative interactions of prediction and planning for end-to-end autonomous driving.

A Survey on Vision-Language-Action Models for Autonomous Driving Ppad: Iterative interactions of prediction and planning for end-to-end autonomous driving

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.377211Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.377211Z digest=sha256:a9e0c61c880d402c38f2905f29bc50545ed381cd154b89775fd59755188dd532

Observation 981fc864-c778-4305-a2b6-26039998b808 · outbound

This paper cites CoVLA: Comprehensive vision-language-action dataset for au- tonomous driving.

A Survey on Vision-Language-Action Models for Autonomous Driving CoVLA: Comprehensive vision-language-action dataset for au- tonomous driving

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.381473Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.381473Z digest=sha256:92263aa746b10fe2261513574cfb56281697a4a3bc7bbaa16c376cd74a52d6f6

Observation ae9035c8-cebe-40f7-8752-3275413763fa · outbound

This paper cites Impromptu VLA: Open Weights and Open Data for Driving Vision-Language-Action Models.

A Survey on Vision-Language-Action Models for Autonomous Driving Impromptu VLA: Open Weights and Open Data for Driving Vision-Language-Action Models

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.384791Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.384791Z digest=sha256:1a64d2b554de7e2795ff74e20bd675f0fe31c75a5de920b7baa8fe7165b0a6ec

Observation f0a55e4d-3b65-4899-9f39-c0b41d12bc6b · outbound

This paper cites Neat: Neural attention fields for end-to-end autonomous driving.

A Survey on Vision-Language-Action Models for Autonomous Driving Neat: Neural attention fields for end-to-end autonomous driving

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.388971Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.388971Z digest=sha256:850ed756e11bfa0576520b1cea6c8c0fdab0bafa8ac960ca5e4ab4d9b0aa1345

Observation 16c070ea-5acb-463f-9fea-4626f7e83d4f · outbound

This paper cites Transfuser: Imita- tion with transformer-based sensor fusion for autonomous driving.

A Survey on Vision-Language-Action Models for Autonomous Driving Transfuser: Imita- tion with transformer-based sensor fusion for autonomous driving

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.392370Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.392370Z digest=sha256:ab39bbc76b5bd8868c03b44a58d8c5a0726b7f9b93586713f04c29d677625e1e

Observation 230a8c27-d8da-41f2-863a-2fe5460e25b3 · outbound

This paper cites Talk2bev: Language-enhanced bird’s- eye view maps for autonomous driving.

A Survey on Vision-Language-Action Models for Autonomous Driving Talk2bev: Language-enhanced bird’s- eye view maps for autonomous driving

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.395495Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.395495Z digest=sha256:1b56f927df3fed4d08e2200e3572ecca8acfa35e01ec44fcbe78b6b6bde04bb0

Observation d5233365-64c2-4c2b-9811-1693fd109563 · outbound

This paper cites A survey on multimodal large lan- guage models for autonomous driving.

A Survey on Vision-Language-Action Models for Autonomous Driving A survey on multimodal large lan- guage models for autonomous driving

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.398942Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.398942Z digest=sha256:57fc3b9ce8d6e1d18975b494ce83e0200457a5a2e549ed2228ec14c2c80b5cf4

Observation bb544cc9-b712-435f-b5f1-ee1b4e55acd0 · outbound

This paper cites Chain-of-Thought for Autonomous Driving: A Comprehensive Survey and Future Prospects.

A Survey on Vision-Language-Action Models for Autonomous Driving Chain-of-Thought for Autonomous Driving: A Comprehensive Survey and Future Prospects

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.402197Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.402197Z digest=sha256:f63846162d4183f62e241bb8a265bdb52f9e2692244affcaacf873964b455d8e

Observation 97a7f9ab-fabd-4d1e-ba7a-1558608947d7 · outbound

This paper cites Hint-AD: Holistically Aligned Interpretability in End-to-End Autonomous Driving.

A Survey on Vision-Language-Action Models for Autonomous Driving Hint-AD: Holistically Aligned Interpretability in End-to-End Autonomous Driving

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.405773Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.405773Z digest=sha256:c03802a7f6e3336da1472a91e300dbff292423d3cdcaf81a14341f21b4c4c8d2

Observation 400d1b16-2441-4279-a5c5-cb36963b63cf · outbound

This paper cites Dualad: Disentangling the dynamic and static world for end-to-end driving.

A Survey on Vision-Language-Action Models for Autonomous Driving Dualad: Disentangling the dynamic and static world for end-to-end driving

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.409296Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.409296Z digest=sha256:afaff73225a8745025ee041129ee13e867c649a4721294771acc305b177e7b49

Observation 043f54c0-3ae0-49c0-b9c5-106b8097b215 · outbound

This paper cites Carla: An open urban driv- ing simulator.

A Survey on Vision-Language-Action Models for Autonomous Driving Carla: An open urban driv- ing simulator

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.412510Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.412510Z digest=sha256:7e189c164511482a2fa4f22f66783a189b2fb6bf01dfdae051e26468605552e0

Observation 18721357-3ffb-480b-a858-92e7014b9fc1 · outbound

This paper cites On the road to portability: Compressing end-to-end motion planner for autonomous driving.

A Survey on Vision-Language-Action Models for Autonomous Driving On the road to portability: Compressing end-to-end motion planner for autonomous driving

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.415798Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.415798Z digest=sha256:998414215f756aad20e30ca1d8075459f3d204125bd70cf3f5714b2261ac779b

Observation bf68921b-35b3-4fb8-984d-61e1486323c8 · outbound

This paper cites Polarpoint-bev: Bird-eye- view perception in polar points for explainable end-to-end autonomous driving.

A Survey on Vision-Language-Action Models for Autonomous Driving Polarpoint-bev: Bird-eye- view perception in polar points for explainable end-to-end autonomous driving

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.418970Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.418970Z digest=sha256:a211fc659ae2fd60b89ffc50acdcbcd7055ec749f671de7376df0c5bcb973793

Observation c5da9684-8de4-4242-87fb-5b4afdb9c75e · outbound

This paper cites Drive like a human: Rethinking autonomous driving with large language models.

A Survey on Vision-Language-Action Models for Autonomous Driving Drive like a human: Rethinking autonomous driving with large language models

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.422092Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.422092Z digest=sha256:8c1f07987b560a9e6215caab72f11b0a98148bf080cd31279f5259ffb6ff8ff3

Observation 25067f8d-5942-446f-b9ee-06a64cc0550d · outbound

This paper cites ORION: A Holistic End-to-End Autonomous Driving Framework by Vision-Language Instructed Action Generation.

A Survey on Vision-Language-Action Models for Autonomous Driving ORION: A Holistic End-to-End Autonomous Driving Framework by Vision-Language Instructed Action Generation

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.429364Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.429364Z digest=sha256:b23f2747b02ea127fe43dbaa14f1512287a36f0bcd1fc43488dc702ff14e624d

Observation ff5b3305-7a88-4afd-9ab0-2fc46ae66821 · outbound

This paper cites A Survey for Foundation Models in Autonomous Driving.

A Survey on Vision-Language-Action Models for Autonomous Driving A Survey for Foundation Models in Autonomous Driving

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.432603Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.432603Z digest=sha256:05e05617587110ad761a560ca9b05f2b1e38c38b5656f35e21a1366961aa54b0

Observation e064c0f6-803e-497d-814f-5ec440ab0104 · outbound

This paper cites LangCoop: Collaborative Driving with Language.

A Survey on Vision-Language-Action Models for Autonomous Driving LangCoop: Collaborative Driving with Language

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.436251Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.436251Z digest=sha256:5ce2eb36602598bfbcd871ab0fbc33dd8a9d129a4aafafb90031ca71b95d9502

Observation a06dc6e1-527a-4ff7-84eb-a8eb02c12f7c · outbound

This paper cites A review of motion planning techniques for au- tomated vehicles.

A Survey on Vision-Language-Action Models for Autonomous Driving A review of motion planning techniques for au- tomated vehicles

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.439666Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.439666Z digest=sha256:03d8a6bf51ebfe84735cc967635f0b89c52d1080b55f8394ab8dad2ee2ab9919

Observation eb80006e-0204-4af8-9719-477ef2246b51 · outbound

This paper cites iPad: Iterative Proposal-centric End-to-End Autonomous Driving.

A Survey on Vision-Language-Action Models for Autonomous Driving iPad: Iterative Proposal-centric End-to-End Autonomous Driving

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.442877Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.442877Z digest=sha256:61f0d1742bc14d7055da32c6cc24c6d947682335e60dd8e1b8d7f9d75b96aca6

Observation f6b14606-f44c-4317-aff4-815a937b3744 · outbound

This paper cites End-to-End Autonomous Driving without Costly Modularization and 3D Manual Annotation.

A Survey on Vision-Language-Action Models for Autonomous Driving End-to-End Autonomous Driving without Costly Modularization and 3D Manual Annotation

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.446259Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.446259Z digest=sha256:4be37c9ef76ec881925f695e81a30732abb5dd2199b27dbf0df84cff0e795d30

Observation cbe35daa-af34-47e7-9f69-a5ff07a714ed · outbound

This paper cites SURDS: Benchmarking Spatial Understanding and Reasoning in Driving Scenarios with Vision Language Models.

A Survey on Vision-Language-Action Models for Autonomous Driving SURDS: Benchmarking Spatial Understanding and Reasoning in Driving Scenarios with Vision Language Models

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.449729Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.449729Z digest=sha256:74239edb211708a9dc12101d32627b937a3f2b942291ff08000c866653ff8439

Observation c8a4693a-d8ac-41fc-86b0-14116f3e8d94 · outbound

This paper cites Dme-driver: Integrating human decision logic and 3d scene perception in autonomous driving.

A Survey on Vision-Language-Action Models for Autonomous Driving Dme-driver: Integrating human decision logic and 3d scene perception in autonomous driving

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.453118Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.453118Z digest=sha256:53813a65be2d9d7fe100a7e97d4e256d4bd8daad96cb0bdae5820d496c6bf95d

Observation ed0d1377-4c9b-413f-b84e-4ff74c4aeca0 · outbound

This paper cites Driveaction: A benchmark for exploring human-like driving decisions in vla models.

A Survey on Vision-Language-Action Models for Autonomous Driving Driveaction: A benchmark for exploring human-like driving decisions in vla models

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.456365Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.456365Z digest=sha256:0b9ad03d7ebabf07cb955540681d3b8ecdbfbbca50d61bdd00700aec9f7615d3

Observation 625c298e-8009-4e1a-815c-aeb0cfe79849 · outbound

This paper cites Urban driving with conditional imitation learning.

A Survey on Vision-Language-Action Models for Autonomous Driving Urban driving with conditional imitation learning

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.459454Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.459454Z digest=sha256:793fde80a6ce32e684cc4e190f2596fd0b1b8a63fe23fbc92d64c7cce5779dab

Observation 49fdd285-272c-4511-a055-efdfe7f3aef7 · outbound

This paper cites DriveAgent: Multi-Agent Structured Reasoning with LLM and Multimodal Sensor Fusion for Autonomous Driving.

A Survey on Vision-Language-Action Models for Autonomous Driving DriveAgent: Multi-Agent Structured Reasoning with LLM and Multimodal Sensor Fusion for Autonomous Driving

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.462537Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.462537Z digest=sha256:10a0c1fe614680c9c37e46b886fb49aee91ac0efbbb3a8c8641d442e829eeeee

Observation 7222be9c-5737-4c26-b8fb-be9cc01874f8 · outbound

This paper cites Lora: Low-rank adaptation of large language models.

A Survey on Vision-Language-Action Models for Autonomous Driving Lora: Low-rank adaptation of large language models

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.465990Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.465990Z digest=sha256:abe286dd56af0c4f8d72dfe2208460ef0a52c960deb57ad49b4258aa11a5f07a

Observation 3b76c5c3-ee14-47a7-a98d-07ad99bca1d9 · outbound

This paper cites Planning-oriented autonomous driving.

A Survey on Vision-Language-Action Models for Autonomous Driving Planning-oriented autonomous driving

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.469144Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.469144Z digest=sha256:49d619854365934f55f2b44225f2a9c8ef82c0f848978792b6a5e7376dc533c0

Observation 25810685-f58b-4f16-a052-e0c8c25afd6c · outbound

This paper cites Planning-oriented autonomous driving.

A Survey on Vision-Language-Action Models for Autonomous Driving Planning-oriented autonomous driving

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.472273Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.472273Z digest=sha256:a64fe458b0f07c01461b5836266a42d98f71d5fe3a471b377180f85802c0d57b

Observation 3f37d0c4-e48d-4bbc-9443-1f08c6123840 · outbound

This paper cites A survey on trajectory-prediction methods for autonomous driving.

A Survey on Vision-Language-Action Models for Autonomous Driving A survey on trajectory-prediction methods for autonomous driving

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.475355Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.475355Z digest=sha256:5cd73f89ac7ab46c3f8b38f5c6535d2b92895c552b2ea6d3a3a4c20ffb47896d

Observation a181b55c-bd64-43f7-8c89-c9e8f06f9781 · outbound

This paper cites RoboTron-Drive: All-in-One Large Multimodal Model for Autonomous Driving.

A Survey on Vision-Language-Action Models for Autonomous Driving RoboTron-Drive: All-in-One Large Multimodal Model for Autonomous Driving

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.478639Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.478639Z digest=sha256:b1dc2a99826287e4cac5a6ac0196dbd6bf17929a339d182526df2b4b261f7d1e

Observation c46c8d11-f331-4e38-8b22-09967ed31c8b · outbound

This paper cites VLM-RL: A Unified Vision Language Models and Reinforcement Learning Framework for Safe Autonomous Driving.

A Survey on Vision-Language-Action Models for Autonomous Driving VLM-RL: A Unified Vision Language Models and Reinforcement Learning Framework for Safe Autonomous Driving

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.482447Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.482447Z digest=sha256:211a761d6ee376be95250ba3685b1baeb03adc744a92b278b6d5554e6a1a5c59

Observation 0bc08dcf-5eaf-4ec0-9951-e0c68652fccf · outbound

This paper cites NORA: A Small Open-Sourced Generalist Vision Language Action Model for Embodied Tasks.

A Survey on Vision-Language-Action Models for Autonomous Driving NORA: A Small Open-Sourced Generalist Vision Language Action Model for Embodied Tasks

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.485889Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.485889Z digest=sha256:784d3d38899a443e6ed1b405c453d75bf39439e2f110c11938477944da1ebfef

Observation 569b3479-5401-4fbc-8485-2bc3187b5b8a · outbound

This paper cites Yolo-v1 to yolo-v8, the rise of yolo and its complementary nature toward digital manufacturing and industrial defect detection.

A Survey on Vision-Language-Action Models for Autonomous Driving Yolo-v1 to yolo-v8, the rise of yolo and its complementary nature toward digital manufacturing and industrial defect detection

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.489388Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.489388Z digest=sha256:0bd18a66f2cea98bdd74d2ce70880eea76525fe89d8d5c5055a28d023c12b233

Observation 5224ef4c-f19f-4cbf-b687-c2984152e145 · outbound

This paper cites EMMA: End-to-End Multimodal Model for Autonomous Driving.

A Survey on Vision-Language-Action Models for Autonomous Driving EMMA: End-to-End Multimodal Model for Autonomous Driving

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.492816Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.492816Z digest=sha256:32c61ff6662c3e4ae24535ee7edfc8430859726d3232fd5198bc7c57230e575c

Observation 559edfd9-9390-47cf-841f-54cc8482836a · outbound

This paper cites DriveLMM-o1: A Step-by-Step Reasoning Dataset and Large Multimodal Model for Driving Scenario Understanding.

A Survey on Vision-Language-Action Models for Autonomous Driving DriveLMM-o1: A Step-by-Step Reasoning Dataset and Large Multimodal Model for Driving Scenario Understanding

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.496591Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.496591Z digest=sha256:961304e60a251317d9d901f18c3d2d50c594980f6b321ce1af745b21487ecf50

Observation aacfba20-a326-47dd-945b-f37dc9f8ee14 · outbound

This paper cites Narrate: Versatile lan- guage architecture for optimal control in robotics.

A Survey on Vision-Language-Action Models for Autonomous Driving Narrate: Versatile lan- guage architecture for optimal control in robotics

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.499865Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.499865Z digest=sha256:056b4cd45d2d8ab00b4533202ca5768e791c06663e22ad3a6c5bc2c4de914e18

Observation d92e1b2c-cbdb-4b14-9bc7-14bf466ee07e · outbound

This paper cites ADriver-I: A General World Model for Autonomous Driving.

A Survey on Vision-Language-Action Models for Autonomous Driving ADriver-I: A General World Model for Autonomous Driving

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.503085Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.503085Z digest=sha256:834329f9fbc2b1f2cfbec98304006bdb6c2fafe2d46a3a8b254d5c09246cb4c0

Observation 334a87f1-6670-47ce-9a9c-084f0356cfd2 · outbound

This paper cites Think twice be- fore driving: Towards scalable decoders for end-to-end au- tonomous driving.

A Survey on Vision-Language-Action Models for Autonomous Driving Think twice be- fore driving: Towards scalable decoders for end-to-end au- tonomous driving

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.506379Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.506379Z digest=sha256:d392b52ba676f380eba898be5c72083e14b6200f48384b86f5ed4ffa85b1c980

Observation 38578a71-14e3-43f1-aca0-8dd29613d89c · outbound

This paper cites Bench2Drive: Towards Multi-Ability Benchmarking of Closed-Loop End-To-End Autonomous Driving.

A Survey on Vision-Language-Action Models for Autonomous Driving Bench2Drive: Towards Multi-Ability Benchmarking of Closed-Loop End-To-End Autonomous Driving

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.509692Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.509692Z digest=sha256:1bb32db96288c25a7cd3826348bbf39486a738aea2d3a6b9b16bdd86a7a92ecb

Observation 41eabb91-c06c-40d7-9c42-900df6541d45 · outbound

This paper cites DriveTransformer: Unified Transformer for Scalable End-to-End Autonomous Driving.

A Survey on Vision-Language-Action Models for Autonomous Driving DriveTransformer: Unified Transformer for Scalable End-to-End Autonomous Driving

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.513157Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.513157Z digest=sha256:75e06c091e72a0fc6c7310d44274748343d3f417d2d3418b4c549c6344a6f6b1

Observation 224a316f-d4fa-4a7b-a4cd-896d8c42f4ef · outbound

This paper cites DiffVLA: Vision-Language Guided Diffusion Planning for Autonomous Driving.

A Survey on Vision-Language-Action Models for Autonomous Driving DiffVLA: Vision-Language Guided Diffusion Planning for Autonomous Driving

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.516550Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.516550Z digest=sha256:be23f617bd143bf0351f30d5eb1a1e5a294f906889cb0036b4548e6470fddacd

Observation e42c38b3-5667-4a3b-9987-316e834e7550 · outbound

This paper cites Senna: Bridging Large Vision-Language Models and End-to-End Autonomous Driving.

A Survey on Vision-Language-Action Models for Autonomous Driving Senna: Bridging Large Vision-Language Models and End-to-End Autonomous Driving

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.520794Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.520794Z digest=sha256:e0aba54eb520b8036214138d96e3343fe359c24e013f1b0076a482c033213e4c

Observation 5c35b144-26d3-4866-b0c6-8ca291c57bda · outbound

This paper cites Vad: Vectorized scene rep- resentation for efficient autonomous driving.

A Survey on Vision-Language-Action Models for Autonomous Driving Vad: Vectorized scene rep- resentation for efficient autonomous driving

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.524393Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.524393Z digest=sha256:29ecfa702d8341966fb7891ca4faed0d029bc0e0ffe5342329272de4e63b5f5b

Observation 8269ba50-47e2-4843-b48a-d34b15368cda · outbound

This paper cites Koma: Knowledge-driven multi- agent framework for autonomous driving with large lan- guage models.

A Survey on Vision-Language-Action Models for Autonomous Driving Koma: Knowledge-driven multi- agent framework for autonomous driving with large lan- guage models

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.527846Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.527846Z digest=sha256:57aaa9c5d23b93482d0137896d1322f7e3b06f4b36fb517a58c519898d4d1818

Observation 848168b4-dbea-4e9c-8ddf-c10b85559472 · outbound

This paper cites Communication-Aware Reinforcement Learning for Cooperative Adaptive Cruise Control.

A Survey on Vision-Language-Action Models for Autonomous Driving Communication-Aware Reinforcement Learning for Cooperative Adaptive Cruise Control

Reference 61

Resolution
verified exact
local_arxiv, observed 2026-08-06T21:31:06.177631Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T21:31:04.530858Z digest=sha256:d1d2af2266ddc5d2257bd4a352d03b3aef19d61ddb354d8faf61d1fbb9897238

Observation faec5161-b732-46ed-bc48-d3785462ad2c · outbound

This paper cites Learning to drive in a day.

A Survey on Vision-Language-Action Models for Autonomous Driving Learning to drive in a day

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.534174Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.534174Z digest=sha256:9ceabb2c9c6d04dceb4b3180b5bfd99d3cae645a05433797b7053c2d5cfb5418

Observation 653bce07-5833-4ada-90fe-b04a282e2f8f · outbound

This paper cites an unresolved cited work.

A Survey on Vision-Language-Action Models for Autonomous Driving Unresolved cited work

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.537689Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.537689Z digest=sha256:47c4782c1211dfc2ebd9f7b3aea0a8f1c2b38a1e7bf0da79d53fc5db2455be4a

Observation 75ea97e9-bd5c-4ff7-b378-05062533862b · outbound

This paper cites Fine-Tuning Vision-Language-Action Models: Optimizing Speed and Success.

A Survey on Vision-Language-Action Models for Autonomous Driving Fine-Tuning Vision-Language-Action Models: Optimizing Speed and Success

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.540761Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.540761Z digest=sha256:0f75fb3602b92ca4f0d1c77363296dbb1b5feee03a59d059a20000ac89320a55

Observation 88d37ec8-5d2f-4ecf-8d92-6d7cb7aaa2d4 · outbound

This paper cites OpenVLA: An Open-Source Vision-Language-Action Model.

A Survey on Vision-Language-Action Models for Autonomous Driving OpenVLA: An Open-Source Vision-Language-Action Model

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.544824Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.544824Z digest=sha256:b7aa77cb11bc8103b1f47a4d4698d767c16e6c5dc92a4d6d585fb422dee0f55e

Observation 0261be28-95c3-4a3f-b0dc-7f76815360f1 · outbound

This paper cites A survey on motion prediction and risk assessment for in- telligent vehicles.

A Survey on Vision-Language-Action Models for Autonomous Driving A survey on motion prediction and risk assessment for in- telligent vehicles

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.549316Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.549316Z digest=sha256:635f775f939a9057e749df5bda0a9dd7f0cca6401bb97e7e01612182f3d1271d

Observation 4acb9048-e194-4083-bc7f-d33873d302b6 · outbound

This paper cites PointVLA: Injecting the 3D World into Vision-Language-Action Models.

A Survey on Vision-Language-Action Models for Autonomous Driving PointVLA: Injecting the 3D World into Vision-Language-Action Models

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.553726Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.553726Z digest=sha256:02e5dd97ff8aa4f2a6b916a9be947a18a79627f4ef6717b2c0f3cddf6e9a4fe1

Observation d42ecad3-38ae-4afa-a415-a0a6ab560be7 · outbound

This paper cites Navigation-Guided Sparse Scene Representation for End-to-End Autonomous Driving.

A Survey on Vision-Language-Action Models for Autonomous Driving Navigation-Guided Sparse Scene Representation for End-to-End Autonomous Driving

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.557074Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.557074Z digest=sha256:b6e69cad42f4e4bf8d375bfb78289a7bfd5cabf9d775d570fe40e73a4f1da9c5

Observation 004a1db2-0aa4-40a9-94f9-8ba012b716fd · outbound

This paper cites Enhancing End-to-End Autonomous Driving with Latent World Model.

A Survey on Vision-Language-Action Models for Autonomous Driving Enhancing End-to-End Autonomous Driving with Latent World Model

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.560587Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.560587Z digest=sha256:8d71a8fc17cf24a289c14ba174af4e6370ca3bc8ff5832e49ae508035676cbd9

Observation e3f1a3fa-a0a9-4035-9bf7-2035b9e91235 · outbound

This paper cites Recogdrive: A reinforced cognitive framework for end-to-end autonomous driving, 2025.

A Survey on Vision-Language-Action Models for Autonomous Driving Recogdrive: A reinforced cognitive framework for end-to-end autonomous driving, 2025

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.564262Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.564262Z digest=sha256:0d40e7ba4c52b0ddd6604759c2ab1ec2a72599d03a384d1ea756de63d59cf84d

Observation 01b84402-1ceb-47f6-bd22-f4984485b8b3 · outbound

This paper cites Hydra-MDP: End-to-end Multimodal Planning with Multi-target Hydra-Distillation.

A Survey on Vision-Language-Action Models for Autonomous Driving Hydra-MDP: End-to-end Multimodal Planning with Multi-target Hydra-Distillation

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.568371Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.568371Z digest=sha256:4fd5d74aa0a4ad9dbe8e497077ce873e7ad897c19e9c621f9052a2bc995eba66

Observation 58283645-65e3-4dc8-be82-d85c9dc06e21 · outbound

This paper cites Generalized Trajectory Scoring for End-to-end Multimodal Planning.

A Survey on Vision-Language-Action Models for Autonomous Driving Generalized Trajectory Scoring for End-to-end Multimodal Planning

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.572098Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.572098Z digest=sha256:ec939981cc72de541665e11db89b0e900530b87d1ab9b6ce147b758575c811d2

Observation 6b37605a-5907-47e3-959b-3713cd1443eb · outbound

This paper cites Pnpnet: End-to-end per- ception and prediction with tracking in the loop.

A Survey on Vision-Language-Action Models for Autonomous Driving Pnpnet: End-to-end per- ception and prediction with tracking in the loop

Reference 73

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.576167Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.576167Z digest=sha256:700f6fa0595f4597f206e4c70291e28ab2f7c9ab57462d9dd5fd12b07cab428f

Observation e0c04cd3-a3d8-41b6-880b-70f0e1e154a1 · outbound

This paper cites Visual instruction tuning.

A Survey on Vision-Language-Action Models for Autonomous Driving Visual instruction tuning

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.579305Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.579305Z digest=sha256:adfba0d1540239e73580e89c7e5427ccbec50d00d282aa6f7602114bdfbc6be0

Observation 44ea8d2f-3f23-4c2f-ae28-109d31e4e189 · outbound

This paper cites Mtd-gpt: A multi-task decision-making gpt model for autonomous driving at unsignalized intersections.

A Survey on Vision-Language-Action Models for Autonomous Driving Mtd-gpt: A multi-task decision-making gpt model for autonomous driving at unsignalized intersections

Reference 75

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.582341Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.582341Z digest=sha256:80a55e945a1fa377d9b349c95999076ef211e25c76615083da89d9dd98e56e40

Observation 25d0286c-a370-4c1f-8d37-358293a1a5d5 · outbound

This paper cites Robomamba: Ef- ficient vision-language-action model for robotic reasoning and manipulation.

A Survey on Vision-Language-Action Models for Autonomous Driving Robomamba: Ef- ficient vision-language-action model for robotic reasoning and manipulation

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.585664Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.585664Z digest=sha256:db10b8958d3a021547eb3caf53a76096bbec0d11d6253087057b008d896b2ddd

Observation 6b80e62b-0141-4228-8bbd-3fd7451e7954 · outbound

This paper cites Fully Unified Motion Planning for End-to-End Autonomous Driving.

A Survey on Vision-Language-Action Models for Autonomous Driving Fully Unified Motion Planning for End-to-End Autonomous Driving

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.588876Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.588876Z digest=sha256:f7464309e93c15621796e18c0f9a9578880029b12576396090e384c2ce6a9430

Observation a78c074f-2c92-41a6-8e06-99d5163e47a5 · outbound

This paper cites Vlm-e2e: Enhancing end-to-end autonomous driv- ing with multimodal driver attention fusion.

A Survey on Vision-Language-Action Models for Autonomous Driving Vlm-e2e: Enhancing end-to-end autonomous driv- ing with multimodal driver attention fusion

Reference 78

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.592420Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.592420Z digest=sha256:e101074f78c4b6c3d78ae4a8729ad522c81aba942cf2f0187af3f08d9b4c77a4

Observation a9f834fe-66c9-44ec-82ad-36c0632a40ce · outbound

This paper cites Reasonplan: Unified scene prediction and decision reasoning for closed-loop au- tonomous driving.

A Survey on Vision-Language-Action Models for Autonomous Driving Reasonplan: Unified scene prediction and decision reasoning for closed-loop au- tonomous driving

Reference 79

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.595814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.595814Z digest=sha256:3a08d6308cd01101fdcba35693a47d564930e7d8ed9248ec8d0576c62d5fb255

Observation 2dd1e3bc-b9ca-4c7b-8123-1e5b6f7222f6 · outbound

This paper cites Bevfusion: Multi-task multi-sensor fusion with unified bird’s-eye view representation.

A Survey on Vision-Language-Action Models for Autonomous Driving Bevfusion: Multi-task multi-sensor fusion with unified bird’s-eye view representation

Reference 80

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.599077Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.599077Z digest=sha256:1af33028483e62fe24877e60365e9c2a071dc5c1a312509d79935668d68803c2

Observation 30478a36-9dec-494f-9251-a392c74f2b68 · outbound

This paper cites VLM-MPC: Vision Language Foundation Model (VLM)-Guided Model Predictive Controller (MPC) for Autonomous Driving.

A Survey on Vision-Language-Action Models for Autonomous Driving VLM-MPC: Vision Language Foundation Model (VLM)-Guided Model Predictive Controller (MPC) for Autonomous Driving

Reference 81

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.602293Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.602293Z digest=sha256:9a72376500c000c3f056405c6fad486ddcb4b7b503b90e6dd243cfe9fc80a72f

Observation f55be80c-fd31-45a6-a185-719224fa5fd1 · outbound

This paper cites ActiveAD: Planning-Oriented Active Learning for End-to-End Autonomous Driving.

A Survey on Vision-Language-Action Models for Autonomous Driving ActiveAD: Planning-Oriented Active Learning for End-to-End Autonomous Driving

Reference 82

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.605560Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.605560Z digest=sha256:2acc31f9e3a63c44409a46fe1e7b392e212cede798d4a42464277aedfc28c9d7

Observation 6b7a9aad-23d7-4b7d-89c1-d3bf8446c354 · outbound

This paper cites Fast and fu- rious: Real time end-to-end 3d detection, tracking and mo- tion forecasting with a single convolutional net.

A Survey on Vision-Language-Action Models for Autonomous Driving Fast and fu- rious: Real time end-to-end 3d detection, tracking and mo- tion forecasting with a single convolutional net

Reference 83

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.608939Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.608939Z digest=sha256:bd3f83b009c203eb0e115331103dda1a38976cf66245080c71cfdec6201e1136

Observation 8342d199-fd35-4553-81a6-23c516a5613a · outbound

This paper cites Dolphins: Multimodal language model for driving.

A Survey on Vision-Language-Action Models for Autonomous Driving Dolphins: Multimodal language model for driving

Reference 84

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.612085Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.612085Z digest=sha256:3ee805e9c5014114110103f87c030c8a24bf585f4d93a042fb2d2c28a17b7fdd

Observation 89742c4a-fee3-4efe-bba9-f3dac8bf713a · outbound

This paper cites A Survey on Vision-Language-Action Models for Embodied AI.

A Survey on Vision-Language-Action Models for Autonomous Driving A Survey on Vision-Language-Action Models for Embodied AI

Reference 85

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.615132Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.615132Z digest=sha256:72163fe68abec0f4858babe9a207eff9332b44d41c1a58bf22e7b2e25ceeee34

Observation 6e81063f-e9f2-41e9-b3cf-77ed23412124 · outbound

This paper cites LeapVAD: A Leap in Autonomous Driving via Cognitive Perception and Dual-Process Thinking.

A Survey on Vision-Language-Action Models for Autonomous Driving LeapVAD: A Leap in Autonomous Driving via Cognitive Perception and Dual-Process Thinking

Reference 86

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.618427Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.618427Z digest=sha256:74ce2d555ee48af56d2c083e9c2281281cd577d082e8e0ca041d8f92c6b37731

Observation cda65a11-b737-4623-b2a2-fbc9896fcb20 · outbound

This paper cites GPT-Driver: Learning to Drive with GPT.

A Survey on Vision-Language-Action Models for Autonomous Driving GPT-Driver: Learning to Drive with GPT

Reference 87

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.622875Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.622875Z digest=sha256:22dfae800925afa394ece8158da70ac752afef288d726bf3ec4caf054f293b21

Observation d02d2829-189c-4182-a3c9-f7aa54bf1d9c · outbound

This paper cites A Language Agent for Autonomous Driving.

A Survey on Vision-Language-Action Models for Autonomous Driving A Language Agent for Autonomous Driving

Reference 88

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.626480Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.626480Z digest=sha256:66abb2888fd1dcf88f177c569b9258471aa7cf0d7f9cfe103530cba47e354eb2

Observation 83a6066d-1707-4fa5-9bbb-420260d6686b · outbound

This paper cites Lingoqa: Visual question answering for autonomous driving.

A Survey on Vision-Language-Action Models for Autonomous Driving Lingoqa: Visual question answering for autonomous driving

Reference 89

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.629874Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.629874Z digest=sha256:61380eb249380fb924f9ed41c25c71204e5cc92be17108f08b0988d5e138def9

Observation 1ec33b50-ede3-45f7-a0bb-db23a2fe0d2f · outbound

This paper cites Continuously Learning, Adapting, and Improving: A Dual-Process Approach to Autonomous Driving.

A Survey on Vision-Language-Action Models for Autonomous Driving Continuously Learning, Adapting, and Improving: A Dual-Process Approach to Autonomous Driving

Reference 90

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.633849Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.633849Z digest=sha256:5fe964346735a5ce4b80c9a0b168a3303470c92d20a1ecf4119799d62034e4bc

Observation e7f686fe-f594-4a44-a100-80d971a5004e · outbound

This paper cites Chatmpc: Natural language based mpc personalization.

A Survey on Vision-Language-Action Models for Autonomous Driving Chatmpc: Natural language based mpc personalization

Reference 91

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.638377Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.638377Z digest=sha256:b052df2b5c58bc498e3df33724df04495d69a0b7bd9332e6159f63ac207fa738

Observation 4ef1b1d9-cf11-4a75-bdc3-3bef5054ae1a · outbound

This paper cites Data Scaling Laws for End-to-End Autonomous Driving.

A Survey on Vision-Language-Action Models for Autonomous Driving Data Scaling Laws for End-to-End Autonomous Driving

Reference 92

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.641543Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.641543Z digest=sha256:22e939df76c844dd49ff761cb5932f2a6cdf41f0f671027089b35db4101a3f4e

Observation aab49f27-f584-4724-a0ef-1071af9a8d3e · outbound

This paper cites Rea- son2drive: Towards interpretable and chain-based reason- ing for autonomous driving.

A Survey on Vision-Language-Action Models for Autonomous Driving Rea- son2drive: Towards interpretable and chain-based reason- ing for autonomous driving

Reference 93

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.645337Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.645337Z digest=sha256:c96e97ab9af5d46689390006f1213b4c2bddf8bae75fed5d15bcecefba936d2c

Observation 50a6587c-3ce2-410d-b1f3-67784e3c7a6a · outbound

This paper cites DINOv2: Learning Robust Visual Features without Supervision.

A Survey on Vision-Language-Action Models for Autonomous Driving DINOv2: Learning Robust Visual Features without Supervision

Reference 94

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.648737Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.648737Z digest=sha256:ac3cee742af6a68d75f8ac847dcbe6dfb3915af5db90d533d2ef5f545f825527

Observation 2bd0575f-59dd-4132-9ead-ab10afbd4ff6 · outbound

This paper cites A survey of motion planning and control techniques for self-driving urban vehicles.

A Survey on Vision-Language-Action Models for Autonomous Driving A survey of motion planning and control techniques for self-driving urban vehicles

Reference 95

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.652503Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.652503Z digest=sha256:4604ecbd1fe7873288d5be2683632cc8812533a83e19a9ef89e9e92a1e9b04a2

Observation ae903616-2e2d-4ab7-a36c-49d69a2c4e2e · outbound

This paper cites Vlp: Vision language planning for autonomous driving.

A Survey on Vision-Language-Action Models for Autonomous Driving Vlp: Vision language planning for autonomous driving

Reference 96

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.656685Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.656685Z digest=sha256:2c92753f27b28c694d6b87c97c42bc3b017b4828685a42267de9325211523045

Observation d4f55e8d-20a7-4b11-897b-961c39ea1760 · outbound

This paper cites Lego-drive: Language-enhanced goal-oriented closed-loop end-to-end autonomous driving.

A Survey on Vision-Language-Action Models for Autonomous Driving Lego-drive: Language-enhanced goal-oriented closed-loop end-to-end autonomous driving

Reference 97

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.659895Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.659895Z digest=sha256:7f745e87ab689dbf6373dba3694e917f6f47ca97dd881a2d1ca1661065cc4647

Observation 29f45af6-fa90-40c6-93e8-d9eddb9426a5 · outbound

This paper cites FAST: Efficient Action Tokenization for Vision-Language-Action Models.

A Survey on Vision-Language-Action Models for Autonomous Driving FAST: Efficient Action Tokenization for Vision-Language-Action Models

Reference 98

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.663325Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.663325Z digest=sha256:30c134c3a99743d937d23fcd15865aaec02fdb67283f23135fcd90d4e0d85aed

Observation c2c0b6f3-5add-43e1-8983-a090cf69e547 · outbound

This paper cites Agentthink: A unified framework for tool-augmented chain-of-thought reasoning in vision- language models for autonomous driving.

A Survey on Vision-Language-Action Models for Autonomous Driving Agentthink: A unified framework for tool-augmented chain-of-thought reasoning in vision- language models for autonomous driving

Reference 99

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.667282Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.667282Z digest=sha256:1ca7af6071b7a334c1d61aed4a85c21df226f304a1dfb350dae457f8834929a4

Observation bd8d2423-3ca1-44ce-bdaa-d1c46b3f86ec · outbound

This paper cites FASIONAD++ : Integrating High-Level Instruction and Information Bottleneck in FAt-Slow fusION Systems for Enhanced Safety in Autonomous Driving with Adaptive Feedback.

A Survey on Vision-Language-Action Models for Autonomous Driving FASIONAD++ : Integrating High-Level Instruction and Information Bottleneck in FAt-Slow fusION Systems for Enhanced Safety in Autonomous Driving with Adaptive Feedback

Reference 100

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.670504Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.670504Z digest=sha256:275a7fcaa5908e46ada6a28d68e8f6e0edeb2dfd4ef2316f85663e387a02f302

Observation a020dcbf-7044-407c-8e17-0afe8753c195 · outbound

This paper cites SpatialVLA: Exploring Spatial Representations for Visual-Language-Action Model.

A Survey on Vision-Language-Action Models for Autonomous Driving SpatialVLA: Exploring Spatial Representations for Visual-Language-Action Model

Reference 101

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.674245Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.674245Z digest=sha256:926fe7ba4f049f03977fc13eb20def140c6e492e0188fecce22254abae331474

Pith citing papers

Observation 6114a5cf-459a-4563-b467-772dcb481c99 · inbound

Research Challenges and Progress in the End-to-End V2X Cooperative Autonomous Driving Competition cites this paper.

Research Challenges and Progress in the End-to-End V2X Cooperative Autonomous Driving Competition A Survey on Vision-Language-Action Models for Autonomous Driving

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-06T12:39:02.315303Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:39:02.315303Z digest=sha256:e88e1afb7b2ac76548397bb2325aefe994046e52eb5562ff276069e30b70a191

Observation 0538d1b4-ec46-4733-ba7d-7fb55c70bcfc · inbound

Efficient Vision-Language-Action Models for Embodied Manipulation: A Systematic Survey cites this paper.

Efficient Vision-Language-Action Models for Embodied Manipulation: A Systematic Survey A Survey on Vision-Language-Action Models for Autonomous Driving

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-04T09:08:07.066489Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:08:07.066489Z digest=sha256:03d505add217b668a276421be77de72c8c48fe5ce2dc355ccd678b040e315c61

Observation 92d6ba2c-4203-42ed-9441-008dca2c40cc · inbound

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving cites this paper.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving A Survey on Vision-Language-Action Models for Autonomous Driving

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.416823Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.416823Z digest=sha256:db3a892c14b9fcca0435ae2db61c58141d9761ee599934359988cdb6df2daadf

Observation 3ac9cf36-3282-43d6-9301-f54beba24238 · inbound

VLSA: Vision-Language-Action Models with Plug-and-Play Safety Constraint Layer cites this paper.

VLSA: Vision-Language-Action Models with Plug-and-Play Safety Constraint Layer A Survey on Vision-Language-Action Models for Autonomous Driving

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-03T17:38:20.749342Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T17:38:20.749342Z digest=sha256:8a87adb15008beb4513e8b33d0d989e75c74e30776ce4341965e9a2a10d8b8b9

Observation ac0e4079-2379-48f1-aae3-20bfad6b6af6 · inbound

A Review of Learning-Based Motion Planning: Toward a Data-Driven Optimal Control Approach cites this paper.

A Review of Learning-Based Motion Planning: Toward a Data-Driven Optimal Control Approach A Survey on Vision-Language-Action Models for Autonomous Driving

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-03T16:52:52.413571Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:52:52.413571Z digest=sha256:d85e2cf2b0824ecba9a19c7263fd57b0ee987c40e8703383f6e42a04c1cd9555

Observation 0b1d1384-657a-4d9f-b8a6-c3c5f9bd6585 · inbound

LinMU: Multimodal Understanding Made Linear cites this paper.

LinMU: Multimodal Understanding Made Linear A Survey on Vision-Language-Action Models for Autonomous Driving

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-16T18:33:15.213150Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-16T18:32:23.951566Z digest=sha256:ad0418df42b95132255952445f6cd34ff2b0b5805de92ea2773b5c5fc4fb2fcf

Observation 04c9f689-ac34-47c7-a01d-e8440510b18d · inbound

From Representational Complementarity to Dual Systems: Synergizing VLM and Vision-Only Backbones for End-to-End Driving cites this paper.

From Representational Complementarity to Dual Systems: Synergizing VLM and Vision-Only Backbones for End-to-End Driving A Survey on Vision-Language-Action Models for Autonomous Driving

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-04T06:07:09.468768Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T06:07:09.468768Z digest=sha256:6993245ed58887fd4753782e58e0873b3c037d079ddd42e2464c110fa59d0b70

Observation 136b991e-d88f-48a0-9c4b-a5d941be43f3 · inbound

Steadily moving semi-infinite fracture in plane poroelasticity cites this paper.

Steadily moving semi-infinite fracture in plane poroelasticity A Survey on Vision-Language-Action Models for Autonomous Driving

Reference 48

Resolution
metadata mismatch
local_arxiv, observed 2026-07-05T11:41:02.544826Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-07-05T11:39:05.686584Z digest=sha256:7f9db01aa2ed2ad9baba5e7fc364ead6f127bfacff3179521f7c36f52ac3df44

Observation 85266423-3b9a-482c-b212-fadb1dfcaf91 · inbound

XEmbodied: A Foundation Model with Enhanced Geometric and Physical Cues for Large-Scale Embodied Environments cites this paper.

XEmbodied: A Foundation Model with Enhanced Geometric and Physical Cues for Large-Scale Embodied Environments A Survey on Vision-Language-Action Models for Autonomous Driving

Reference 48

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T05:51:10.420760Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-10T05:46:36.865150Z digest=sha256:43da59956ee261bb3615131c0e17f0b96a0e1fb17838019f689a92f7bc9d7a3a

Observation 24670dd2-623b-4306-9a19-12df51202285 · inbound

EgoDyn-Bench: Evaluating Ego-Motion Understanding in Vision-Centric Foundation Models for Autonomous Driving cites this paper.

EgoDyn-Bench: Evaluating Ego-Motion Understanding in Vision-Centric Foundation Models for Autonomous Driving A Survey on Vision-Language-Action Models for Autonomous Driving

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-05-10T00:19:47.159530Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-10T00:09:18.068337Z digest=sha256:905998acaf55ceef395db5394d33c60696bd4fc29dc8bf0e507faa3ac400a791

Observation 670b3c2b-9aab-4827-bfd5-e0b330b9271c · inbound

MVPruner: Dynamic Token Pruning for Accelerating Multi-view Vision-Language Models in Autonomous Driving cites this paper.

MVPruner: Dynamic Token Pruning for Accelerating Multi-view Vision-Language Models in Autonomous Driving A Survey on Vision-Language-Action Models for Autonomous Driving

Reference 15

Resolution
metadata mismatch
arxiv_id, observed 2026-06-29T18:23:51.152734Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-06-29T05:08:46.145616Z digest=sha256:65dce21e118319918c607a91688d5d00af1c5c43abedeec348e6dc2846768aa6

Observation 3158c713-bf1e-40eb-aa7b-8f6b7f88b616 · inbound

MVPruner: Dynamic Token Pruning for Accelerating Multi-view Vision-Language Models in Autonomous Driving cites this paper.

MVPruner: Dynamic Token Pruning for Accelerating Multi-view Vision-Language Models in Autonomous Driving A Survey on Vision-Language-Action Models for Autonomous Driving

Reference 15

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T21:37:24.180527Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-07-02T21:35:33.280591Z digest=sha256:5944e8b76b40d256f207ebc6a3c25b21767aa703adbb38ca816f3d289b5ee785

Observation 0d0865d9-586c-436c-b941-718cebd07d84 · inbound

PriorEye: Geospatial Visual Priors for End-to-End Autonomous Driving cites this paper.

PriorEye: Geospatial Visual Priors for End-to-End Autonomous Driving A Survey on Vision-Language-Action Models for Autonomous Driving

Reference 29

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T09:55:41.492322Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-07-01T05:58:58.075150Z digest=sha256:ca25894524c86b1d3c6a0dfd2fcdfbaa954e9cf2c6a0caee3c26f03a16ee66c4

Observation 8c159a26-866d-4ccf-a534-b4e5dccdc3dc · inbound

MoRAL: Sensor-Grounded BEV Reasoning for Compact VLMs toward Edge-Oriented Autonomous Driving cites this paper.

MoRAL: Sensor-Grounded BEV Reasoning for Compact VLMs toward Edge-Oriented Autonomous Driving A Survey on Vision-Language-Action Models for Autonomous Driving

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-04T07:11:20.776007Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T07:11:20.776007Z digest=sha256:ec4cc8a048e1c0f90a575cd070bb4476171bbb4b1bd892913e83a20a8426407d