Pith. sign in

Paper Citation Record · LEDGER

A Survey on Vision-Language-Action Models for Autonomous Driving

As of 7 August 2026, this Paper Citation Record lists 100 of 166 outbound references and 14 inbound Pith citation observations for arXiv:2506.24044.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.24044 v1

Coverage vector

measured 100 of 166 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T21:31:04.674245Z

measured 114 of 114 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 14 of 14 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T12:39:02.315303Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-05T11:41:02.543365Z

Reference resolution

100 of 166 outbound references displayed

  • verified exact1
  • verified fuzzy0
  • unresolved99
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 95972b4e-af7f-4426-b799-a6a82ba1c909 · outbound

This paper cites Flamingo: a visual language model for few-shot learn- ing.

A Survey on Vision-Language-Action Models for Autonomous Driving Flamingo: a visual language model for few-shot learn- ing

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.317875Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.317875Z digest=sha256:a97dc38ac8d500725d71329840259a66c958fae92748a27512bc452a173ee1f2

Observation 9683c8e4-9fa6-4ba5-b4de-b343ce145803 · outbound

This paper cites An lstm net- work for highway trajectory prediction.

A Survey on Vision-Language-Action Models for Autonomous Driving An lstm net- work for highway trajectory prediction

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.321970Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.321970Z digest=sha256:e002cf97f89fd9e9a6c714eeb98e6009729a3ade42b9717a1337ccdf746360cb

Observation 1d2f59fb-38cf-4392-b6ad-e788b6dbb978 · outbound

This paper cites VaViM and VaVAM: Autonomous Driving through Video Generative Modeling.

A Survey on Vision-Language-Action Models for Autonomous Driving VaViM and VaVAM: Autonomous Driving through Video Generative Modeling

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.325481Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.325481Z digest=sha256:a97c17add54a5b6afd4116bd0f27beebe45e210fefb05b6bac0d98626a9e4c37

Observation 3be9daf3-30aa-47f2-8f53-03ea0dfa3524 · outbound

This paper cites $\pi_0$: A Vision-Language-Action Flow Model for General Robot Control.

A Survey on Vision-Language-Action Models for Autonomous Driving $\pi_0$: A Vision-Language-Action Flow Model for General Robot Control

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.329364Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.329364Z digest=sha256:333f6baf359db46e2286a55b6d97c909d040997637d61cc8f05b311768a51cb1

Observation a940b45e-a0dc-4940-be05-64a3d59804c0 · outbound

This paper cites Fine-grained affective processing capabilities emerging from large lan- guage models.

A Survey on Vision-Language-Action Models for Autonomous Driving Fine-grained affective processing capabilities emerging from large lan- guage models

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.333208Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.333208Z digest=sha256:71dc962606d5bf9811ea40b6db1f2aa5bfcb0f82a4617ddb15624aca0b0ae83c

Observation 0c255595-3398-4239-bed9-78afb5ad81ef · outbound

This paper cites Language models are few-shot learners.

A Survey on Vision-Language-Action Models for Autonomous Driving Language models are few-shot learners

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.336827Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.336827Z digest=sha256:85d37cb5b7b90d6a34837c9f9ed278e9f307531c68a339b11132893fd3b26453

Observation eddffb62-ec8c-4c99-9782-f80d4cdc0146 · outbound

This paper cites nuscenes: A multimodal dataset for autonomous driving.

A Survey on Vision-Language-Action Models for Autonomous Driving nuscenes: A multimodal dataset for autonomous driving

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.341206Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.341206Z digest=sha256:4d7f20100ce26ff2eae2d75ab1c1c4dadb31cac7c1bdb33dc74a26979eae7d76

Observation 7fa2c5dd-9bfd-4e66-9c0b-c682eeeca2fa · outbound

This paper cites nuscenes: A mul- timodal dataset for autonomous driving.

A Survey on Vision-Language-Action Models for Autonomous Driving nuscenes: A mul- timodal dataset for autonomous driving

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.344745Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.344745Z digest=sha256:cce2b413f9eee82bc348e260b2f687003d0be1d0e52ad33f05520996e50c8ff8

Observation 87ae5ca0-c84f-4946-9287-e7bf4952b552 · outbound

This paper cites Learning from all vehi- cles.

A Survey on Vision-Language-Action Models for Autonomous Driving Learning from all vehi- cles

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.349406Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.349406Z digest=sha256:3a662c7084b53fb870ac9faa080a76bb0efa31934556c4a9925eca66c386edeb

Observation 15618d1e-cae1-4536-a0a4-4ad30489a0fb · outbound

This paper cites Insight: Enhancing autonomous driving safety through vision-language models on context-aware hazard detection and edge case evaluation.

A Survey on Vision-Language-Action Models for Autonomous Driving Insight: Enhancing autonomous driving safety through vision-language models on context-aware hazard detection and edge case evaluation

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.352905Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.352905Z digest=sha256:788a2525f90a150f7a64ff8912e8069e7cd27eac15dd672823bb5ff4c5ca0b33

Observation 88023e48-d173-4a4c-a9d3-d4966c1b92f8 · outbound

This paper cites TS-VLM: Text-Guided SoftSort Pooling for Vision-Language Models in Multi-View Driving Reasoning.

A Survey on Vision-Language-Action Models for Autonomous Driving TS-VLM: Text-Guided SoftSort Pooling for Vision-Language Models in Multi-View Driving Reasoning

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.357239Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.357239Z digest=sha256:4ee3c055eb89243ff5a8418decc8365da4821e9e948a16164ef7a093bd495fdc

Observation 4e41fc49-42a9-42da-9ee8-3ce930c479e6 · outbound

This paper cites What data do we need for training an av motion planner? In 2021 IEEE Inter- national Conference on Robotics and Automation (ICRA) , pages 1066–1072.

A Survey on Vision-Language-Action Models for Autonomous Driving What data do we need for training an av motion planner? In 2021 IEEE Inter- national Conference on Robotics and Automation (ICRA) , pages 1066–1072

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.361431Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.361431Z digest=sha256:d69b62d795f25a0d776177179223ef36f821561d3b60cb295ce82607f40d44aa

Observation b5cd6e6b-7abf-4712-a9a2-783a5470fee6 · outbound

This paper cites End-to-end autonomous driving: Challenges and frontiers.

A Survey on Vision-Language-Action Models for Autonomous Driving End-to-end autonomous driving: Challenges and frontiers

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.364887Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.364887Z digest=sha256:9d100a5cd32e5ea2950fda186ef66e71450adaf11cc1da20c1b0fa8ec165c5fc

Observation f48e424e-dd98-4372-b145-c8ce827bcbc7 · outbound

This paper cites VADv2: End-to-End Vectorized Autonomous Driving via Probabilistic Planning.

A Survey on Vision-Language-Action Models for Autonomous Driving VADv2: End-to-End Vectorized Autonomous Driving via Probabilistic Planning

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.369131Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.369131Z digest=sha256:b20f3484cb9cf2cb8c1e22fd5f796d20f75ac2927b5fd8533f52ec36d8ddd920

Observation 1d70c450-28a6-4dcc-8af5-14c10c1389aa · outbound

This paper cites Asynchronous large language model en- hanced planner for autonomous driving.

A Survey on Vision-Language-Action Models for Autonomous Driving Asynchronous large language model en- hanced planner for autonomous driving

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.373472Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.373472Z digest=sha256:19e05fb8f5080e8185156b51638a008c0c53d2fcb58f6137802527386b245d07

Observation 206166b7-1c33-4c8c-971f-57f8bf472cfb · outbound

This paper cites Ppad: Iterative interactions of prediction and planning for end-to-end autonomous driving.

A Survey on Vision-Language-Action Models for Autonomous Driving Ppad: Iterative interactions of prediction and planning for end-to-end autonomous driving

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.377211Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.377211Z digest=sha256:b6eb4cd148d5e8bdd72397e2a5fe99a9eb68e6bb2b1edd70bf24a165222fdad9

Observation 981fc864-c778-4305-a2b6-26039998b808 · outbound

This paper cites CoVLA: Comprehensive vision-language-action dataset for au- tonomous driving.

A Survey on Vision-Language-Action Models for Autonomous Driving CoVLA: Comprehensive vision-language-action dataset for au- tonomous driving

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.381473Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.381473Z digest=sha256:e0e18ef04ba6e580090c7db42eaf350f5b6adce474a113c1adc26e83c1457ab4

Observation ae9035c8-cebe-40f7-8752-3275413763fa · outbound

This paper cites Impromptu VLA: Open Weights and Open Data for Driving Vision-Language-Action Models.

A Survey on Vision-Language-Action Models for Autonomous Driving Impromptu VLA: Open Weights and Open Data for Driving Vision-Language-Action Models

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.384791Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.384791Z digest=sha256:48183cfe5c3d22d4144f8ad32fa527793e655ed161c693f1d4645bebc6aadd40

Observation f0a55e4d-3b65-4899-9f39-c0b41d12bc6b · outbound

This paper cites Neat: Neural attention fields for end-to-end autonomous driving.

A Survey on Vision-Language-Action Models for Autonomous Driving Neat: Neural attention fields for end-to-end autonomous driving

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.388971Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.388971Z digest=sha256:b6ab8eb6e773b4dacac4fbfcf8796a70a9973ae48e8e37dd034975bc34960965

Observation 16c070ea-5acb-463f-9fea-4626f7e83d4f · outbound

This paper cites Transfuser: Imita- tion with transformer-based sensor fusion for autonomous driving.

A Survey on Vision-Language-Action Models for Autonomous Driving Transfuser: Imita- tion with transformer-based sensor fusion for autonomous driving

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.392370Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.392370Z digest=sha256:8d5e09370e9bf78a36c9dddf0f0b5fa9ded93da339ddcb2de66105560a9bf548

Observation 230a8c27-d8da-41f2-863a-2fe5460e25b3 · outbound

This paper cites Talk2bev: Language-enhanced bird’s- eye view maps for autonomous driving.

A Survey on Vision-Language-Action Models for Autonomous Driving Talk2bev: Language-enhanced bird’s- eye view maps for autonomous driving

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.395495Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.395495Z digest=sha256:c963e9e40b444012c8262502d0a93298289280f5dad323bb91e8f423cce389d9

Observation d5233365-64c2-4c2b-9811-1693fd109563 · outbound

This paper cites A survey on multimodal large lan- guage models for autonomous driving.

A Survey on Vision-Language-Action Models for Autonomous Driving A survey on multimodal large lan- guage models for autonomous driving

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.398942Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.398942Z digest=sha256:8bcc003b3c7a3d7ee131648ef7c9a11e30d87335d4a84b664ad77d6629d71846

Observation bb544cc9-b712-435f-b5f1-ee1b4e55acd0 · outbound

This paper cites Chain-of-Thought for Autonomous Driving: A Comprehensive Survey and Future Prospects.

A Survey on Vision-Language-Action Models for Autonomous Driving Chain-of-Thought for Autonomous Driving: A Comprehensive Survey and Future Prospects

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.402197Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.402197Z digest=sha256:918b521234805e4c76f9e2c74d09bb1a0492c9617d927b0a75f9e1662fe6ad9f

Observation 97a7f9ab-fabd-4d1e-ba7a-1558608947d7 · outbound

This paper cites Hint-AD: Holistically Aligned Interpretability in End-to-End Autonomous Driving.

A Survey on Vision-Language-Action Models for Autonomous Driving Hint-AD: Holistically Aligned Interpretability in End-to-End Autonomous Driving

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.405773Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.405773Z digest=sha256:a74dd2cf0c7274439b6bc1ae5d0efa8a50fd17db7b2d018980966bf079806ae3

Observation 400d1b16-2441-4279-a5c5-cb36963b63cf · outbound

This paper cites Dualad: Disentangling the dynamic and static world for end-to-end driving.

A Survey on Vision-Language-Action Models for Autonomous Driving Dualad: Disentangling the dynamic and static world for end-to-end driving

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.409296Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.409296Z digest=sha256:19cc4349ab464bc97936d2830831eb379d31873793a35301f70610b72b42cbb0

Observation 043f54c0-3ae0-49c0-b9c5-106b8097b215 · outbound

This paper cites Carla: An open urban driv- ing simulator.

A Survey on Vision-Language-Action Models for Autonomous Driving Carla: An open urban driv- ing simulator

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.412510Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.412510Z digest=sha256:d7bc866f0c7ad1fbef4cc17fd2505c3fc7fb10cf3f7b8d912f04270afacb635a

Observation 18721357-3ffb-480b-a858-92e7014b9fc1 · outbound

This paper cites On the road to portability: Compressing end-to-end motion planner for autonomous driving.

A Survey on Vision-Language-Action Models for Autonomous Driving On the road to portability: Compressing end-to-end motion planner for autonomous driving

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.415798Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.415798Z digest=sha256:6d266d6e79cb8a83abce0a31a76afcf6ef8355508b38926cc4d2c46ab1bf5714

Observation bf68921b-35b3-4fb8-984d-61e1486323c8 · outbound

This paper cites Polarpoint-bev: Bird-eye- view perception in polar points for explainable end-to-end autonomous driving.

A Survey on Vision-Language-Action Models for Autonomous Driving Polarpoint-bev: Bird-eye- view perception in polar points for explainable end-to-end autonomous driving

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.418970Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.418970Z digest=sha256:aaa9cd345b73f683ffad879672453403526379fffc3d216f8e6a6ffb807700da

Observation c5da9684-8de4-4242-87fb-5b4afdb9c75e · outbound

This paper cites Drive like a human: Rethinking autonomous driving with large language models.

A Survey on Vision-Language-Action Models for Autonomous Driving Drive like a human: Rethinking autonomous driving with large language models

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.422092Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.422092Z digest=sha256:48d6f028db75c844ad77d3466cbf3992759679f3e62360c862176f13f2b4d6a3

Observation 25067f8d-5942-446f-b9ee-06a64cc0550d · outbound

This paper cites ORION: A Holistic End-to-End Autonomous Driving Framework by Vision-Language Instructed Action Generation.

A Survey on Vision-Language-Action Models for Autonomous Driving ORION: A Holistic End-to-End Autonomous Driving Framework by Vision-Language Instructed Action Generation

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.429364Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.429364Z digest=sha256:7c5551cfef86e3ff234dd614e5c9ed13df1f058d3c98300832e8f13faeac5a35

Observation ff5b3305-7a88-4afd-9ab0-2fc46ae66821 · outbound

This paper cites A Survey for Foundation Models in Autonomous Driving.

A Survey on Vision-Language-Action Models for Autonomous Driving A Survey for Foundation Models in Autonomous Driving

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.432603Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.432603Z digest=sha256:f1d0f511d2b49c3d5011957890849a649e36410ad56d7f272af2b57b75d4b271

Observation e064c0f6-803e-497d-814f-5ec440ab0104 · outbound

This paper cites LangCoop: Collaborative Driving with Language.

A Survey on Vision-Language-Action Models for Autonomous Driving LangCoop: Collaborative Driving with Language

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.436251Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.436251Z digest=sha256:293d75b11ab105fc8b7ba5dda5e666d4c2eddd1757b5dc2dd659bbbbf9c38a0b

Observation a06dc6e1-527a-4ff7-84eb-a8eb02c12f7c · outbound

This paper cites A review of motion planning techniques for au- tomated vehicles.

A Survey on Vision-Language-Action Models for Autonomous Driving A review of motion planning techniques for au- tomated vehicles

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.439666Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.439666Z digest=sha256:09fd137112e49ad42a274574c61096f0c32f2fd25092afdd68a7c7ead1510b3e

Observation eb80006e-0204-4af8-9719-477ef2246b51 · outbound

This paper cites iPad: Iterative Proposal-centric End-to-End Autonomous Driving.

A Survey on Vision-Language-Action Models for Autonomous Driving iPad: Iterative Proposal-centric End-to-End Autonomous Driving

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.442877Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.442877Z digest=sha256:c7b5f75f4f559938a19631b2d105ae8fba920920d2811543402620f4ede437b9

Observation f6b14606-f44c-4317-aff4-815a937b3744 · outbound

This paper cites End-to-End Autonomous Driving without Costly Modularization and 3D Manual Annotation.

A Survey on Vision-Language-Action Models for Autonomous Driving End-to-End Autonomous Driving without Costly Modularization and 3D Manual Annotation

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.446259Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.446259Z digest=sha256:254d0e32c35b2e21d2ab8899322becc62a047ba56ddaa0d8258e349b7551bae9

Observation cbe35daa-af34-47e7-9f69-a5ff07a714ed · outbound

This paper cites SURDS: Benchmarking Spatial Understanding and Reasoning in Driving Scenarios with Vision Language Models.

A Survey on Vision-Language-Action Models for Autonomous Driving SURDS: Benchmarking Spatial Understanding and Reasoning in Driving Scenarios with Vision Language Models

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.449729Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.449729Z digest=sha256:99805721c886e885cdb29fae887d41d3057207449d5a8c6d313744e763cf3350

Observation c8a4693a-d8ac-41fc-86b0-14116f3e8d94 · outbound

This paper cites Dme-driver: Integrating human decision logic and 3d scene perception in autonomous driving.

A Survey on Vision-Language-Action Models for Autonomous Driving Dme-driver: Integrating human decision logic and 3d scene perception in autonomous driving

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.453118Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.453118Z digest=sha256:cdeabe06f8b00d4d4c94c10aa96e2ec3b7bb8ff4b11fe9c085fb771074322acb

Observation ed0d1377-4c9b-413f-b84e-4ff74c4aeca0 · outbound

This paper cites Driveaction: A benchmark for exploring human-like driving decisions in vla models.

A Survey on Vision-Language-Action Models for Autonomous Driving Driveaction: A benchmark for exploring human-like driving decisions in vla models

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.456365Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.456365Z digest=sha256:3efe34de3a4c056a9e9b680ea0ec391c6f8d6558efc6e6910961c772657b0356

Observation 625c298e-8009-4e1a-815c-aeb0cfe79849 · outbound

This paper cites Urban driving with conditional imitation learning.

A Survey on Vision-Language-Action Models for Autonomous Driving Urban driving with conditional imitation learning

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.459454Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.459454Z digest=sha256:3c8cedbdac8d02281850255b7dab3c6586efeceed323f4812e73a8ae344c412c

Observation 49fdd285-272c-4511-a055-efdfe7f3aef7 · outbound

This paper cites DriveAgent: Multi-Agent Structured Reasoning with LLM and Multimodal Sensor Fusion for Autonomous Driving.

A Survey on Vision-Language-Action Models for Autonomous Driving DriveAgent: Multi-Agent Structured Reasoning with LLM and Multimodal Sensor Fusion for Autonomous Driving

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.462537Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.462537Z digest=sha256:ac0377f5d3b3cd9d6d78b69d0102a3867a38b583e3cf811498825b8e787c88a8

Observation 7222be9c-5737-4c26-b8fb-be9cc01874f8 · outbound

This paper cites Lora: Low-rank adaptation of large language models.

A Survey on Vision-Language-Action Models for Autonomous Driving Lora: Low-rank adaptation of large language models

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.465990Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.465990Z digest=sha256:174405ea895784fe015efe0dd119fc7197cd6a3d224c885049b29a0c6bcb81fa

Observation 3b76c5c3-ee14-47a7-a98d-07ad99bca1d9 · outbound

This paper cites Planning-oriented autonomous driving.

A Survey on Vision-Language-Action Models for Autonomous Driving Planning-oriented autonomous driving

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.469144Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.469144Z digest=sha256:f589c476757eee13b5512c822cd28d961350bba00b31b0d60fe8010a1a15391e

Observation 25810685-f58b-4f16-a052-e0c8c25afd6c · outbound

This paper cites Planning-oriented autonomous driving.

A Survey on Vision-Language-Action Models for Autonomous Driving Planning-oriented autonomous driving

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.472273Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.472273Z digest=sha256:bd2f22a458f2f577fafccc33d05507c60626b4684e9bb35c225bd06cd6bea5f5

Observation 3f37d0c4-e48d-4bbc-9443-1f08c6123840 · outbound

This paper cites A survey on trajectory-prediction methods for autonomous driving.

A Survey on Vision-Language-Action Models for Autonomous Driving A survey on trajectory-prediction methods for autonomous driving

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.475355Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.475355Z digest=sha256:03b4a24861b0e28febba93204d794ad550f26d4d08eb7855393c69dc40ef5ac0

Observation a181b55c-bd64-43f7-8c89-c9e8f06f9781 · outbound

This paper cites RoboTron-Drive: All-in-One Large Multimodal Model for Autonomous Driving.

A Survey on Vision-Language-Action Models for Autonomous Driving RoboTron-Drive: All-in-One Large Multimodal Model for Autonomous Driving

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.478639Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.478639Z digest=sha256:9d9de1e67da5ffd5546862d49a0c4e936df7f2d644ce36f338d65dde63310335

Observation c46c8d11-f331-4e38-8b22-09967ed31c8b · outbound

This paper cites VLM-RL: A Unified Vision Language Models and Reinforcement Learning Framework for Safe Autonomous Driving.

A Survey on Vision-Language-Action Models for Autonomous Driving VLM-RL: A Unified Vision Language Models and Reinforcement Learning Framework for Safe Autonomous Driving

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.482447Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.482447Z digest=sha256:3186f82b725684ff7dac48e4880edf8a7e9a57208081bfccc7788e85cfdeeae7

Observation 0bc08dcf-5eaf-4ec0-9951-e0c68652fccf · outbound

This paper cites NORA: A Small Open-Sourced Generalist Vision Language Action Model for Embodied Tasks.

A Survey on Vision-Language-Action Models for Autonomous Driving NORA: A Small Open-Sourced Generalist Vision Language Action Model for Embodied Tasks

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.485889Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.485889Z digest=sha256:1ffc74d1c295c7641a291eaabbaf9e1ca8924f1d6e720a81154721c136430e9a

Observation 569b3479-5401-4fbc-8485-2bc3187b5b8a · outbound

This paper cites Yolo-v1 to yolo-v8, the rise of yolo and its complementary nature toward digital manufacturing and industrial defect detection.

A Survey on Vision-Language-Action Models for Autonomous Driving Yolo-v1 to yolo-v8, the rise of yolo and its complementary nature toward digital manufacturing and industrial defect detection

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.489388Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.489388Z digest=sha256:54b0e11487236577f4dab17881d6c508d50e0067a4f6bad329cf8587b03a40e6

Observation 5224ef4c-f19f-4cbf-b687-c2984152e145 · outbound

This paper cites EMMA: End-to-End Multimodal Model for Autonomous Driving.

A Survey on Vision-Language-Action Models for Autonomous Driving EMMA: End-to-End Multimodal Model for Autonomous Driving

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.492816Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.492816Z digest=sha256:5a16ef05c1f01fc1b7c84e16c7f49dbd7e53e88afc552a3a5d20dee553a6ba5a

Observation 559edfd9-9390-47cf-841f-54cc8482836a · outbound

This paper cites DriveLMM-o1: A Step-by-Step Reasoning Dataset and Large Multimodal Model for Driving Scenario Understanding.

A Survey on Vision-Language-Action Models for Autonomous Driving DriveLMM-o1: A Step-by-Step Reasoning Dataset and Large Multimodal Model for Driving Scenario Understanding

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.496591Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.496591Z digest=sha256:a6b1f9e6963c6819ec0e8983634ff861989b01e792ec392b1cd0c1a607782ed0

Observation aacfba20-a326-47dd-945b-f37dc9f8ee14 · outbound

This paper cites Narrate: Versatile lan- guage architecture for optimal control in robotics.

A Survey on Vision-Language-Action Models for Autonomous Driving Narrate: Versatile lan- guage architecture for optimal control in robotics

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.499865Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.499865Z digest=sha256:d8baa62d067d128671d10c6064566a6e921a1a65ec1f8fa1301b55d0882462e4

Observation d92e1b2c-cbdb-4b14-9bc7-14bf466ee07e · outbound

This paper cites ADriver-I: A General World Model for Autonomous Driving.

A Survey on Vision-Language-Action Models for Autonomous Driving ADriver-I: A General World Model for Autonomous Driving

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.503085Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.503085Z digest=sha256:076d00a56a62a7125722bf2f8ba022ef632323c4dee3a9d3132f4bfbca3c5e7e

Observation 334a87f1-6670-47ce-9a9c-084f0356cfd2 · outbound

This paper cites Think twice be- fore driving: Towards scalable decoders for end-to-end au- tonomous driving.

A Survey on Vision-Language-Action Models for Autonomous Driving Think twice be- fore driving: Towards scalable decoders for end-to-end au- tonomous driving

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.506379Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.506379Z digest=sha256:998d3b465ac293eac023e68c462eda9c12796f716a30f29867635ba1e9c9d8f4

Observation 38578a71-14e3-43f1-aca0-8dd29613d89c · outbound

This paper cites Bench2Drive: Towards Multi-Ability Benchmarking of Closed-Loop End-To-End Autonomous Driving.

A Survey on Vision-Language-Action Models for Autonomous Driving Bench2Drive: Towards Multi-Ability Benchmarking of Closed-Loop End-To-End Autonomous Driving

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.509692Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.509692Z digest=sha256:2bdb5d29561ebef31485fea5f6f37cbcd9f4c453cb2deb543edc51520ba8bbf4

Observation 41eabb91-c06c-40d7-9c42-900df6541d45 · outbound

This paper cites DriveTransformer: Unified Transformer for Scalable End-to-End Autonomous Driving.

A Survey on Vision-Language-Action Models for Autonomous Driving DriveTransformer: Unified Transformer for Scalable End-to-End Autonomous Driving

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.513157Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.513157Z digest=sha256:7dccb9ed1bd24680365233d1c9da00f6b596f1eb7f125d3a4d3332d7f54c07e0

Observation 224a316f-d4fa-4a7b-a4cd-896d8c42f4ef · outbound

This paper cites DiffVLA: Vision-Language Guided Diffusion Planning for Autonomous Driving.

A Survey on Vision-Language-Action Models for Autonomous Driving DiffVLA: Vision-Language Guided Diffusion Planning for Autonomous Driving

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.516550Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.516550Z digest=sha256:8c8bda8ce4bbe060960abed9e9424e7b49cf0eb9b2b4f7fbf948dfa5f4f8bfdb

Observation e42c38b3-5667-4a3b-9987-316e834e7550 · outbound

This paper cites Senna: Bridging Large Vision-Language Models and End-to-End Autonomous Driving.

A Survey on Vision-Language-Action Models for Autonomous Driving Senna: Bridging Large Vision-Language Models and End-to-End Autonomous Driving

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.520794Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.520794Z digest=sha256:e03a641ed7633b4a3cfd4b4651dd7244a385165de0fbc023944d5e35c758d3a0

Observation 5c35b144-26d3-4866-b0c6-8ca291c57bda · outbound

This paper cites Vad: Vectorized scene rep- resentation for efficient autonomous driving.

A Survey on Vision-Language-Action Models for Autonomous Driving Vad: Vectorized scene rep- resentation for efficient autonomous driving

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.524393Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.524393Z digest=sha256:68be3c8dcc3b8d8a698f681ec6c6ad05ff8ec04c08b0f4adf5694dc640c56ad6

Observation 8269ba50-47e2-4843-b48a-d34b15368cda · outbound

This paper cites Koma: Knowledge-driven multi- agent framework for autonomous driving with large lan- guage models.

A Survey on Vision-Language-Action Models for Autonomous Driving Koma: Knowledge-driven multi- agent framework for autonomous driving with large lan- guage models

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.527846Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.527846Z digest=sha256:eabfaacd84f1d179067c4101b35ff61ee41a90895b6576892347ef51b46e0a91

Observation 848168b4-dbea-4e9c-8ddf-c10b85559472 · outbound

This paper cites Communication-Aware Reinforcement Learning for Cooperative Adaptive Cruise Control.

A Survey on Vision-Language-Action Models for Autonomous Driving Communication-Aware Reinforcement Learning for Cooperative Adaptive Cruise Control

Reference 61

Resolution
verified exact
local_arxiv, observed 2026-08-06T21:31:06.177631Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:31:04.530858Z digest=sha256:6f02f1adf445c61ef88737d597c4e15ef33438a0111ec78402b5adb79be90643

Observation faec5161-b732-46ed-bc48-d3785462ad2c · outbound

This paper cites Learning to drive in a day.

A Survey on Vision-Language-Action Models for Autonomous Driving Learning to drive in a day

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.534174Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.534174Z digest=sha256:5e865cc18dd76b48ab3e8905e4fcb105714f9a838f905e3435b216a7b7841c15

Observation 653bce07-5833-4ada-90fe-b04a282e2f8f · outbound

This paper cites an unresolved cited work.

A Survey on Vision-Language-Action Models for Autonomous Driving Unresolved cited work

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.537689Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.537689Z digest=sha256:0a02728cef3bbd9ef62f3e14401b9672790b5ef7361dfaf26e0a0cd83a20742e

Observation 75ea97e9-bd5c-4ff7-b378-05062533862b · outbound

This paper cites Fine-Tuning Vision-Language-Action Models: Optimizing Speed and Success.

A Survey on Vision-Language-Action Models for Autonomous Driving Fine-Tuning Vision-Language-Action Models: Optimizing Speed and Success

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.540761Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.540761Z digest=sha256:271ec22fe31e051a1b8849d252cd9b6a4cff62918e09e4070e297c8c09133477

Observation 88d37ec8-5d2f-4ecf-8d92-6d7cb7aaa2d4 · outbound

This paper cites OpenVLA: An Open-Source Vision-Language-Action Model.

A Survey on Vision-Language-Action Models for Autonomous Driving OpenVLA: An Open-Source Vision-Language-Action Model

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.544824Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.544824Z digest=sha256:18c865b25e6756c4caa4aced0059ca179d5df6b5f7c268920d385ea53b4fa5bf

Observation 0261be28-95c3-4a3f-b0dc-7f76815360f1 · outbound

This paper cites A survey on motion prediction and risk assessment for in- telligent vehicles.

A Survey on Vision-Language-Action Models for Autonomous Driving A survey on motion prediction and risk assessment for in- telligent vehicles

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.549316Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.549316Z digest=sha256:cf0a34094e5f009d7e043dba19bf7e2c93ade09ab1c1699f49d900906ed7c482

Observation 4acb9048-e194-4083-bc7f-d33873d302b6 · outbound

This paper cites PointVLA: Injecting the 3D World into Vision-Language-Action Models.

A Survey on Vision-Language-Action Models for Autonomous Driving PointVLA: Injecting the 3D World into Vision-Language-Action Models

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.553726Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.553726Z digest=sha256:2de4a49c46965f4347e8085f474ec18f1363c473a6dd55db02fe735b953a39b8

Observation d42ecad3-38ae-4afa-a415-a0a6ab560be7 · outbound

This paper cites Navigation-Guided Sparse Scene Representation for End-to-End Autonomous Driving.

A Survey on Vision-Language-Action Models for Autonomous Driving Navigation-Guided Sparse Scene Representation for End-to-End Autonomous Driving

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.557074Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.557074Z digest=sha256:d5fa5f65ea81652047b1f2a929a7ac090d98a3ec73b3b2b140948cfc76f93387

Observation 004a1db2-0aa4-40a9-94f9-8ba012b716fd · outbound

This paper cites Enhancing End-to-End Autonomous Driving with Latent World Model.

A Survey on Vision-Language-Action Models for Autonomous Driving Enhancing End-to-End Autonomous Driving with Latent World Model

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.560587Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.560587Z digest=sha256:335fbf2b947834c48e1bc99abc69dd1437f71ec71fc5f0fa3eb85decbc95a3d1

Observation e3f1a3fa-a0a9-4035-9bf7-2035b9e91235 · outbound

This paper cites Recogdrive: A reinforced cognitive framework for end-to-end autonomous driving, 2025.

A Survey on Vision-Language-Action Models for Autonomous Driving Recogdrive: A reinforced cognitive framework for end-to-end autonomous driving, 2025

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.564262Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.564262Z digest=sha256:2eb105d311e91491554fb0e67356e324357a690439adc3bce58c2737058c415b

Observation 01b84402-1ceb-47f6-bd22-f4984485b8b3 · outbound

This paper cites Hydra-MDP: End-to-end Multimodal Planning with Multi-target Hydra-Distillation.

A Survey on Vision-Language-Action Models for Autonomous Driving Hydra-MDP: End-to-end Multimodal Planning with Multi-target Hydra-Distillation

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.568371Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.568371Z digest=sha256:331d4f719b61b6b901b6f7e76d1a402ed43d1d4f22831c0b8a8cb8a7fe593c6f

Observation 58283645-65e3-4dc8-be82-d85c9dc06e21 · outbound

This paper cites Generalized Trajectory Scoring for End-to-end Multimodal Planning.

A Survey on Vision-Language-Action Models for Autonomous Driving Generalized Trajectory Scoring for End-to-end Multimodal Planning

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.572098Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.572098Z digest=sha256:641c6aa2ec96b4f4ac27792a2e2543d418ba16c35eb576d6a3314ac9364c022a

Observation 6b37605a-5907-47e3-959b-3713cd1443eb · outbound

This paper cites Pnpnet: End-to-end per- ception and prediction with tracking in the loop.

A Survey on Vision-Language-Action Models for Autonomous Driving Pnpnet: End-to-end per- ception and prediction with tracking in the loop

Reference 73

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.576167Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.576167Z digest=sha256:e4ec0cc0e5360d8789957d2e4e8bbc13c88620fac772325bdf462805845cd695

Observation e0c04cd3-a3d8-41b6-880b-70f0e1e154a1 · outbound

This paper cites Visual instruction tuning.

A Survey on Vision-Language-Action Models for Autonomous Driving Visual instruction tuning

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.579305Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.579305Z digest=sha256:813ac0274115930b5fb9d934fcb0ad1b39a8c19852bf4f4547e388908c6908e4

Observation 44ea8d2f-3f23-4c2f-ae28-109d31e4e189 · outbound

This paper cites Mtd-gpt: A multi-task decision-making gpt model for autonomous driving at unsignalized intersections.

A Survey on Vision-Language-Action Models for Autonomous Driving Mtd-gpt: A multi-task decision-making gpt model for autonomous driving at unsignalized intersections

Reference 75

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.582341Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.582341Z digest=sha256:9c3507c9da694a74fd703f38f957f44c3ecd0dde9a749a91e91db86f9da73bc1

Observation 25d0286c-a370-4c1f-8d37-358293a1a5d5 · outbound

This paper cites Robomamba: Ef- ficient vision-language-action model for robotic reasoning and manipulation.

A Survey on Vision-Language-Action Models for Autonomous Driving Robomamba: Ef- ficient vision-language-action model for robotic reasoning and manipulation

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.585664Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.585664Z digest=sha256:c69928c1ae712050339e4525b417d6d23bf0c0ed07b58023e9f6ee873c316834

Observation 6b80e62b-0141-4228-8bbd-3fd7451e7954 · outbound

This paper cites Fully Unified Motion Planning for End-to-End Autonomous Driving.

A Survey on Vision-Language-Action Models for Autonomous Driving Fully Unified Motion Planning for End-to-End Autonomous Driving

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.588876Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.588876Z digest=sha256:cb4a01b4514fa0597b4ff58a134814cf3e375752e6b254a94b05d30eb92ea4aa

Observation a78c074f-2c92-41a6-8e06-99d5163e47a5 · outbound

This paper cites Vlm-e2e: Enhancing end-to-end autonomous driv- ing with multimodal driver attention fusion.

A Survey on Vision-Language-Action Models for Autonomous Driving Vlm-e2e: Enhancing end-to-end autonomous driv- ing with multimodal driver attention fusion

Reference 78

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.592420Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.592420Z digest=sha256:9dc0f550f6bff89ac0e8ee4802fb84eea3bbeba2bb1005b9d0df4c0f83463e4f

Observation a9f834fe-66c9-44ec-82ad-36c0632a40ce · outbound

This paper cites Reasonplan: Unified scene prediction and decision reasoning for closed-loop au- tonomous driving.

A Survey on Vision-Language-Action Models for Autonomous Driving Reasonplan: Unified scene prediction and decision reasoning for closed-loop au- tonomous driving

Reference 79

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.595814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.595814Z digest=sha256:e6742cb89a2003ca2ecc0e902823be4efabeaab41df85c2253aa3e5013f77f9e

Observation 2dd1e3bc-b9ca-4c7b-8123-1e5b6f7222f6 · outbound

This paper cites Bevfusion: Multi-task multi-sensor fusion with unified bird’s-eye view representation.

A Survey on Vision-Language-Action Models for Autonomous Driving Bevfusion: Multi-task multi-sensor fusion with unified bird’s-eye view representation

Reference 80

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.599077Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.599077Z digest=sha256:8b903a814e65de959e66b0c246084fbcd90b342d5ffa0aaf8295667a7f9b775b

Observation 30478a36-9dec-494f-9251-a392c74f2b68 · outbound

This paper cites VLM-MPC: Vision Language Foundation Model (VLM)-Guided Model Predictive Controller (MPC) for Autonomous Driving.

A Survey on Vision-Language-Action Models for Autonomous Driving VLM-MPC: Vision Language Foundation Model (VLM)-Guided Model Predictive Controller (MPC) for Autonomous Driving

Reference 81

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.602293Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.602293Z digest=sha256:17d9031776053fe9579040ec07c23b3beb99b31a0cd6ae5a4ab02d322440001d

Observation f55be80c-fd31-45a6-a185-719224fa5fd1 · outbound

This paper cites ActiveAD: Planning-Oriented Active Learning for End-to-End Autonomous Driving.

A Survey on Vision-Language-Action Models for Autonomous Driving ActiveAD: Planning-Oriented Active Learning for End-to-End Autonomous Driving

Reference 82

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.605560Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.605560Z digest=sha256:df6671f4ae0860c4c77baed434f4c5717aec1a180583deb0d6234470d5092cb2

Observation 6b7a9aad-23d7-4b7d-89c1-d3bf8446c354 · outbound

This paper cites Fast and fu- rious: Real time end-to-end 3d detection, tracking and mo- tion forecasting with a single convolutional net.

A Survey on Vision-Language-Action Models for Autonomous Driving Fast and fu- rious: Real time end-to-end 3d detection, tracking and mo- tion forecasting with a single convolutional net

Reference 83

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.608939Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.608939Z digest=sha256:a1c9c37f2e20c7e7219adefa440aec34205b9511818acb4771efbfe39a0424cc

Observation 8342d199-fd35-4553-81a6-23c516a5613a · outbound

This paper cites Dolphins: Multimodal language model for driving.

A Survey on Vision-Language-Action Models for Autonomous Driving Dolphins: Multimodal language model for driving

Reference 84

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.612085Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.612085Z digest=sha256:1a0d04c13d98caa43adb75b63a7c3829d1355eb4e59620cdaf33e64da6c4b5b6

Observation 89742c4a-fee3-4efe-bba9-f3dac8bf713a · outbound

This paper cites A Survey on Vision-Language-Action Models for Embodied AI.

A Survey on Vision-Language-Action Models for Autonomous Driving A Survey on Vision-Language-Action Models for Embodied AI

Reference 85

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.615132Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.615132Z digest=sha256:f68cc7a1d18e69bf9a142138f501022ff97956336afee2eab70bcf8e67044f3f

Observation 6e81063f-e9f2-41e9-b3cf-77ed23412124 · outbound

This paper cites LeapVAD: A Leap in Autonomous Driving via Cognitive Perception and Dual-Process Thinking.

A Survey on Vision-Language-Action Models for Autonomous Driving LeapVAD: A Leap in Autonomous Driving via Cognitive Perception and Dual-Process Thinking

Reference 86

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.618427Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.618427Z digest=sha256:b29855d9148fdf05d72a58c15ae00319d7e7cb1cfc513da3022aa4ed4e55426c

Observation cda65a11-b737-4623-b2a2-fbc9896fcb20 · outbound

This paper cites GPT-Driver: Learning to Drive with GPT.

A Survey on Vision-Language-Action Models for Autonomous Driving GPT-Driver: Learning to Drive with GPT

Reference 87

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.622875Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.622875Z digest=sha256:6903204ecc7aead1259491ef0a356ff23ced8ec0a8fdd37d6e8202c3528ce9a8

Observation d02d2829-189c-4182-a3c9-f7aa54bf1d9c · outbound

This paper cites A Language Agent for Autonomous Driving.

A Survey on Vision-Language-Action Models for Autonomous Driving A Language Agent for Autonomous Driving

Reference 88

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.626480Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.626480Z digest=sha256:d456f3a2e463a4738cfc202e0bf96b0f8b4210b6fa8cca4bf0d129f1e529934f

Observation 83a6066d-1707-4fa5-9bbb-420260d6686b · outbound

This paper cites Lingoqa: Visual question answering for autonomous driving.

A Survey on Vision-Language-Action Models for Autonomous Driving Lingoqa: Visual question answering for autonomous driving

Reference 89

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.629874Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.629874Z digest=sha256:e18b3433313c003302e1a4992975e53ffa60d0dd095e52bbc59e0a8e0f30287e

Observation 1ec33b50-ede3-45f7-a0bb-db23a2fe0d2f · outbound

This paper cites Continuously Learning, Adapting, and Improving: A Dual-Process Approach to Autonomous Driving.

A Survey on Vision-Language-Action Models for Autonomous Driving Continuously Learning, Adapting, and Improving: A Dual-Process Approach to Autonomous Driving

Reference 90

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.633849Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.633849Z digest=sha256:34277a9ba812e269e30ddfbe4a61a08a5188c9e2d7c92a671c9d5e974fab096e

Observation e7f686fe-f594-4a44-a100-80d971a5004e · outbound

This paper cites Chatmpc: Natural language based mpc personalization.

A Survey on Vision-Language-Action Models for Autonomous Driving Chatmpc: Natural language based mpc personalization

Reference 91

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.638377Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.638377Z digest=sha256:3a23d1d1e79fa460968b5f4f1a6ebb443138d731c20633713797ffd402a14b48

Observation 4ef1b1d9-cf11-4a75-bdc3-3bef5054ae1a · outbound

This paper cites Data Scaling Laws for End-to-End Autonomous Driving.

A Survey on Vision-Language-Action Models for Autonomous Driving Data Scaling Laws for End-to-End Autonomous Driving

Reference 92

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.641543Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.641543Z digest=sha256:1baf9f4cc9c60b25d16f5afc1ded984ad780c12587b36f7e3cad12c14e9e843d

Observation aab49f27-f584-4724-a0ef-1071af9a8d3e · outbound

This paper cites Rea- son2drive: Towards interpretable and chain-based reason- ing for autonomous driving.

A Survey on Vision-Language-Action Models for Autonomous Driving Rea- son2drive: Towards interpretable and chain-based reason- ing for autonomous driving

Reference 93

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.645337Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.645337Z digest=sha256:e6b1003096454b75dc551bf981dbc43624b865ec611d99490920620c4f8937fd

Observation 50a6587c-3ce2-410d-b1f3-67784e3c7a6a · outbound

This paper cites DINOv2: Learning Robust Visual Features without Supervision.

A Survey on Vision-Language-Action Models for Autonomous Driving DINOv2: Learning Robust Visual Features without Supervision

Reference 94

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.648737Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.648737Z digest=sha256:f7c3406692ad5d8191d62235e483b34869e2cb8337c9c9c26945df1fef7076ae

Observation 2bd0575f-59dd-4132-9ead-ab10afbd4ff6 · outbound

This paper cites A survey of motion planning and control techniques for self-driving urban vehicles.

A Survey on Vision-Language-Action Models for Autonomous Driving A survey of motion planning and control techniques for self-driving urban vehicles

Reference 95

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.652503Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.652503Z digest=sha256:b6df4efde1764e3c93deb7738b0013c9b0185b3ca9943a26b8cc3b458cbc854f

Observation ae903616-2e2d-4ab7-a36c-49d69a2c4e2e · outbound

This paper cites Vlp: Vision language planning for autonomous driving.

A Survey on Vision-Language-Action Models for Autonomous Driving Vlp: Vision language planning for autonomous driving

Reference 96

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.656685Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.656685Z digest=sha256:7ab953071d9040bf2ccf44e3c0a8ebb491efb51e1a917bb6d6b844919a1de877

Observation d4f55e8d-20a7-4b11-897b-961c39ea1760 · outbound

This paper cites Lego-drive: Language-enhanced goal-oriented closed-loop end-to-end autonomous driving.

A Survey on Vision-Language-Action Models for Autonomous Driving Lego-drive: Language-enhanced goal-oriented closed-loop end-to-end autonomous driving

Reference 97

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.659895Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.659895Z digest=sha256:7cfc99f7f96513ff2685f033d2edab63c4fd5a020232847fa1f2331e07560a29

Observation 29f45af6-fa90-40c6-93e8-d9eddb9426a5 · outbound

This paper cites FAST: Efficient Action Tokenization for Vision-Language-Action Models.

A Survey on Vision-Language-Action Models for Autonomous Driving FAST: Efficient Action Tokenization for Vision-Language-Action Models

Reference 98

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.663325Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.663325Z digest=sha256:ff41427fa1baa3b88cb792bd3e1cfdb8a16934edbe1b881b88bfe0189bfa2b35

Observation c2c0b6f3-5add-43e1-8983-a090cf69e547 · outbound

This paper cites Agentthink: A unified framework for tool-augmented chain-of-thought reasoning in vision- language models for autonomous driving.

A Survey on Vision-Language-Action Models for Autonomous Driving Agentthink: A unified framework for tool-augmented chain-of-thought reasoning in vision- language models for autonomous driving

Reference 99

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.667282Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.667282Z digest=sha256:40204dfd420e708af430ac1e88d2cde92fd62c5c7396860ef0302b143e886b90

Observation bd8d2423-3ca1-44ce-bdaa-d1c46b3f86ec · outbound

This paper cites FASIONAD++ : Integrating High-Level Instruction and Information Bottleneck in FAt-Slow fusION Systems for Enhanced Safety in Autonomous Driving with Adaptive Feedback.

A Survey on Vision-Language-Action Models for Autonomous Driving FASIONAD++ : Integrating High-Level Instruction and Information Bottleneck in FAt-Slow fusION Systems for Enhanced Safety in Autonomous Driving with Adaptive Feedback

Reference 100

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.670504Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.670504Z digest=sha256:e256d72606062f2d9be4929fc4be2ca7c079f1e2841d83c299a7c6f0af624315

Observation a020dcbf-7044-407c-8e17-0afe8753c195 · outbound

This paper cites SpatialVLA: Exploring Spatial Representations for Visual-Language-Action Model.

A Survey on Vision-Language-Action Models for Autonomous Driving SpatialVLA: Exploring Spatial Representations for Visual-Language-Action Model

Reference 101

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.674245Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.674245Z digest=sha256:056f95aa5b15154652f433dd90891afd1dcb7d0e3fa437371f36d74cc0173df9

Pith citing papers

Observation 6114a5cf-459a-4563-b467-772dcb481c99 · inbound

Research Challenges and Progress in the End-to-End V2X Cooperative Autonomous Driving Competition cites this paper.

Research Challenges and Progress in the End-to-End V2X Cooperative Autonomous Driving Competition A Survey on Vision-Language-Action Models for Autonomous Driving

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-06T12:39:02.315303Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:39:02.315303Z digest=sha256:56200539a32832a04b75a80ea5a2991e1244e962b853095e9c6015af8e68e246

Observation 0538d1b4-ec46-4733-ba7d-7fb55c70bcfc · inbound

Efficient Vision-Language-Action Models for Embodied Manipulation: A Systematic Survey cites this paper.

Efficient Vision-Language-Action Models for Embodied Manipulation: A Systematic Survey A Survey on Vision-Language-Action Models for Autonomous Driving

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-04T09:08:07.066489Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:08:07.066489Z digest=sha256:2f8a639b2bbf9f79cabccda797a0ac70d8e236aed348da176db81d51cff9cc67

Observation 92d6ba2c-4203-42ed-9441-008dca2c40cc · inbound

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving cites this paper.

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving A Survey on Vision-Language-Action Models for Autonomous Driving

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-03T18:37:03.416823Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:37:03.416823Z digest=sha256:69f6cec7208e194ebd020c0c318e84e3489b8c5217e6d70f9a2554a5a6805764

Observation 3ac9cf36-3282-43d6-9301-f54beba24238 · inbound

VLSA: Vision-Language-Action Models with Plug-and-Play Safety Constraint Layer cites this paper.

VLSA: Vision-Language-Action Models with Plug-and-Play Safety Constraint Layer A Survey on Vision-Language-Action Models for Autonomous Driving

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-03T17:38:20.749342Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T17:38:20.749342Z digest=sha256:1fd2357d2adff969344c30531bb6346eb03ed5fff66279a92da3186244587c89

Observation ac0e4079-2379-48f1-aae3-20bfad6b6af6 · inbound

A Review of Learning-Based Motion Planning: Toward a Data-Driven Optimal Control Approach cites this paper.

A Review of Learning-Based Motion Planning: Toward a Data-Driven Optimal Control Approach A Survey on Vision-Language-Action Models for Autonomous Driving

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-03T16:52:52.413571Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:52:52.413571Z digest=sha256:b5b609aa06373ad6650fc707b9c0687a666cf3b51d180a4e389fbc6082c25fbf

Observation 0b1d1384-657a-4d9f-b8a6-c3c5f9bd6585 · inbound

LinMU: Multimodal Understanding Made Linear cites this paper.

LinMU: Multimodal Understanding Made Linear A Survey on Vision-Language-Action Models for Autonomous Driving

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-16T18:33:15.213150Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-16T18:32:23.951566Z digest=sha256:8c4742759d3a7928437e3bb13c9dc8c433f15a4833cf97d40c616f0430a06126

Observation 04c9f689-ac34-47c7-a01d-e8440510b18d · inbound

From Representational Complementarity to Dual Systems: Synergizing VLM and Vision-Only Backbones for End-to-End Driving cites this paper.

From Representational Complementarity to Dual Systems: Synergizing VLM and Vision-Only Backbones for End-to-End Driving A Survey on Vision-Language-Action Models for Autonomous Driving

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-04T06:07:09.468768Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T06:07:09.468768Z digest=sha256:66026c0d4bef09fc5a3b3c881ed44be6aa6b16222e3b202aba948e1a17ee235c

Observation 136b991e-d88f-48a0-9c4b-a5d941be43f3 · inbound

Steadily moving semi-infinite fracture in plane poroelasticity cites this paper.

Steadily moving semi-infinite fracture in plane poroelasticity A Survey on Vision-Language-Action Models for Autonomous Driving

Reference 48

Resolution
metadata mismatch
local_arxiv, observed 2026-07-05T11:41:02.544826Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-07-05T11:39:05.686584Z digest=sha256:09357712503e3f2a5fdf2373692e1c386facc99eefd3ddbc04bb6463e34e7943

Observation 85266423-3b9a-482c-b212-fadb1dfcaf91 · inbound

XEmbodied: A Foundation Model with Enhanced Geometric and Physical Cues for Large-Scale Embodied Environments cites this paper.

XEmbodied: A Foundation Model with Enhanced Geometric and Physical Cues for Large-Scale Embodied Environments A Survey on Vision-Language-Action Models for Autonomous Driving

Reference 48

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T05:51:10.420760Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-10T05:46:36.865150Z digest=sha256:7542255ffd4a37d213eaa41b0884f116a176015b4756a15f70bbe8e82c03ec39

Observation 24670dd2-623b-4306-9a19-12df51202285 · inbound

EgoDyn-Bench: Evaluating Ego-Motion Understanding in Vision-Centric Foundation Models for Autonomous Driving cites this paper.

EgoDyn-Bench: Evaluating Ego-Motion Understanding in Vision-Centric Foundation Models for Autonomous Driving A Survey on Vision-Language-Action Models for Autonomous Driving

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-05-10T00:19:47.159530Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-10T00:09:18.068337Z digest=sha256:2646ef1b4d397c285d1d15827847574b30b38cd07f45bab121272cdbee3d2f90

Observation 670b3c2b-9aab-4827-bfd5-e0b330b9271c · inbound

MVPruner: Dynamic Token Pruning for Accelerating Multi-view Vision-Language Models in Autonomous Driving cites this paper.

MVPruner: Dynamic Token Pruning for Accelerating Multi-view Vision-Language Models in Autonomous Driving A Survey on Vision-Language-Action Models for Autonomous Driving

Reference 15

Resolution
metadata mismatch
arxiv_id, observed 2026-06-29T18:23:51.152734Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-29T05:08:46.145616Z digest=sha256:1aa0d76b9e76358e8428010dfa907d108ae0502d88c1ac9ad49a4245a93438ee

Observation 3158c713-bf1e-40eb-aa7b-8f6b7f88b616 · inbound

MVPruner: Dynamic Token Pruning for Accelerating Multi-view Vision-Language Models in Autonomous Driving cites this paper.

MVPruner: Dynamic Token Pruning for Accelerating Multi-view Vision-Language Models in Autonomous Driving A Survey on Vision-Language-Action Models for Autonomous Driving

Reference 15

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T21:37:24.180527Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-07-02T21:35:33.280591Z digest=sha256:68905d8476fed03327221abc23d99d2824e39aa4b0b67adcfb65e82df5f2c5d6

Observation 0d0865d9-586c-436c-b941-718cebd07d84 · inbound

PriorEye: Geospatial Visual Priors for End-to-End Autonomous Driving cites this paper.

PriorEye: Geospatial Visual Priors for End-to-End Autonomous Driving A Survey on Vision-Language-Action Models for Autonomous Driving

Reference 29

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T09:55:41.492322Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-07-01T05:58:58.075150Z digest=sha256:9b96d0b7fdf40426baefe92f6d529cd93b2030b071a77396743fad761cbce137

Observation 8c159a26-866d-4ccf-a534-b4e5dccdc3dc · inbound

MoRAL: Sensor-Grounded BEV Reasoning for Compact VLMs toward Edge-Oriented Autonomous Driving cites this paper.

MoRAL: Sensor-Grounded BEV Reasoning for Compact VLMs toward Edge-Oriented Autonomous Driving A Survey on Vision-Language-Action Models for Autonomous Driving

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-04T07:11:20.776007Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T07:11:20.776007Z digest=sha256:3def242533dff224bc12c103fd2f1f462b04699a7621e968de79523a6ab8e8e0