Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T06:02:13.187292Z
Paper Citation Record · LEDGER
As of 16 August 2026, this Paper Citation Record lists 76 of 76 outbound references and 6 inbound Pith citation observations for arXiv:2506.06218.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T06:02:13.187292Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-06T00:08:21.388777Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-07-09T05:46:01.450385Z
76 of 76 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 578c97c1-444d-4a6e-b1e1-13d8c2717b6f · outbound
STSBench: A Spatio-temporal Scenario Benchmark for Multi-modal Large Language Models in Autonomous Driving LLaMA 3.2: Open Foundation and Instruction Models, 2024
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 63cdeb96-d1e5-4291-a68f-ce8c9d47efe4 · outbound
STSBench: A Spatio-temporal Scenario Benchmark for Multi-modal Large Language Models in Autonomous Driving Flamingo: a Visual Language Model for Few-Shot Learning
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 905ea11b-5021-431c-8a9d-558743278fe3 · outbound
STSBench: A Spatio-temporal Scenario Benchmark for Multi-modal Large Language Models in Autonomous Driving CoVLA: Comprehensive Vision-Language-Action Dataset for Autonomous Driving
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 13c45644-8a12-4bbd-a3a6-e11adc0b4cb0 · outbound
STSBench: A Spatio-temporal Scenario Benchmark for Multi-modal Large Language Models in Autonomous Driving Qwen2.5-VL Technical Report
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7392d7c1-043c-49b9-bff4-33f1a3996cf2 · outbound
STSBench: A Spatio-temporal Scenario Benchmark for Multi-modal Large Language Models in Autonomous Driving Lang, Sourabh V ora, Venice Erin Liong, Qiang Xu, Anush Krishnan, Yu Pan, Giancarlo Baldan, and Oscar Beijbom
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation eaa9f319-698d-4da0-b5ab-8f591cb92aa1 · outbound
STSBench: A Spatio-temporal Scenario Benchmark for Multi-modal Large Language Models in Autonomous Driving MAPLM: A Real-World Large-Scale Vision-Language Dataset for Map and Traffic Scene Understanding
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 29e0b0a2-5b83-403e-8b58-c4e0f603a994 · outbound
STSBench: A Spatio-temporal Scenario Benchmark for Multi-modal Large Language Models in Autonomous Driving How far are we to gpt-4v? closing the gap to commercial multimodal models with open-source suites
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation be51db96-d4d2-40ce-b9b5-0fbd258ccd95 · outbound
STSBench: A Spatio-temporal Scenario Benchmark for Multi-modal Large Language Models in Autonomous Driving Internvl: Scaling up vision foundation models and aligning for generic visual-linguistic tasks
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 75893966-066d-4887-8676-651159a8b5fd · outbound
STSBench: A Spatio-temporal Scenario Benchmark for Multi-modal Large Language Models in Autonomous Driving The Cityscapes Dataset for Semantic Urban Scene Understanding
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation efd415f6-72e2-431c-b563-cf6736bd4c49 · outbound
STSBench: A Spatio-temporal Scenario Benchmark for Multi-modal Large Language Models in Autonomous Driving InstructBLIP: Towards General-purpose Vision-Language Models with Instruction Tuning
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 2dc79cff-9d03-4d53-809d-9ff63e8df15b · outbound
STSBench: A Spatio-temporal Scenario Benchmark for Multi-modal Large Language Models in Autonomous Driving DeepSeek-V3 Technical Report
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f4de31c4-3919-481f-a0d0-74aaf4f3a52f · outbound
STSBench: A Spatio-temporal Scenario Benchmark for Multi-modal Large Language Models in Autonomous Driving Smith, Hannaneh Hajishirzi, Ross Girshick, Ali Farhadi, and Aniruddha Kembhavi
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 6def661e-8df5-4e99-bb0b-ff7a47980a23 · outbound
STSBench: A Spatio-temporal Scenario Benchmark for Multi-modal Large Language Models in Autonomous Driving Talk2Car: Taking control of your self-driving car
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 88631727-4d17-4ebb-aa1a-f5478a3cf72f · outbound
STSBench: A Spatio-temporal Scenario Benchmark for Multi-modal Large Language Models in Autonomous Driving Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e4a84a4e-1f60-4d2c-b8b8-495c0e52d839 · outbound
STSBench: A Spatio-temporal Scenario Benchmark for Multi-modal Large Language Models in Autonomous Driving Holistic Autonomous Driving Understanding by Bird’s-Eye-View Injected Multi-Modal Large Models
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 601b8b08-f511-4c9f-98aa-b2353f4078fd · outbound
STSBench: A Spatio-temporal Scenario Benchmark for Multi-modal Large Language Models in Autonomous Driving CARLA: An Open Urban Driving Simulator
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation d40f0edc-972c-4867-aa84-a7a60001d6e0 · outbound
STSBench: A Spatio-temporal Scenario Benchmark for Multi-modal Large Language Models in Autonomous Driving ORION: A Holistic End-to-End Autonomous Driving Framework by Vision-Language Instructed Action Generation
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ab061dd3-fe7b-4ad6-adcb-b025364a38fe · outbound
STSBench: A Spatio-temporal Scenario Benchmark for Multi-modal Large Language Models in Autonomous Driving Are we ready for autonomous driving? The KITTI vision benchmark suite
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation eebbe239-95f0-44ab-85c8-545ce79f0b2b · outbound
STSBench: A Spatio-temporal Scenario Benchmark for Multi-modal Large Language Models in Autonomous Driving SURDS: Benchmarking Spatial Understanding and Reasoning in Driving Scenarios with Vision Language Models
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9e87d558-b20f-4cef-ac97-7223b62a16b7 · outbound
STSBench: A Spatio-temporal Scenario Benchmark for Multi-modal Large Language Models in Autonomous Driving Planning-oriented Autonomous Driving
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 83e38495-3cfb-4db7-937a-892a4f4ae3f1 · outbound
STSBench: A Spatio-temporal Scenario Benchmark for Multi-modal Large Language Models in Autonomous Driving RoboTron-Drive: All-in-One Large Multimodal Model for Autonomous Driving
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d24a4d13-87c9-47ef-b6dd-2082928fa097 · outbound
STSBench: A Spatio-temporal Scenario Benchmark for Multi-modal Large Language Models in Autonomous Driving Making Large Language Models Better Planners with Reasoning-Decision Alignment
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation cefcec4e-63ab-43cc-9f01-be84f09718ae · outbound
STSBench: A Spatio-temporal Scenario Benchmark for Multi-modal Large Language Models in Autonomous Driving EMMA: End-to-End Multimodal Model for Autonomous Driving
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4dd1bbdd-79ef-4d20-a514-4a992137034b · outbound
STSBench: A Spatio-temporal Scenario Benchmark for Multi-modal Large Language Models in Autonomous Driving NuScenes-MQA: Integrated Evaluation of Captions and QA for Autonomous Driving Datasets Using Markup Annotations
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 5fd9b368-d0d3-4e47-881c-78b07fcce794 · outbound
STSBench: A Spatio-temporal Scenario Benchmark for Multi-modal Large Language Models in Autonomous Driving DriveLMM-o1: A Step-by-Step Reasoning Dataset and Large Multimodal Model for Driving Scenario Understanding
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ddd1bf57-9e7b-4b46-8e0f-8c99f47b4e06 · outbound
STSBench: A Spatio-temporal Scenario Benchmark for Multi-modal Large Language Models in Autonomous Driving Bench2Drive: Towards Multi-Ability Benchmarking of Closed-Loop End-To-End Autonomous Driving
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation ed812cee-cb1b-45e8-aa2c-1fff4ba8c535 · outbound
STSBench: A Spatio-temporal Scenario Benchmark for Multi-modal Large Language Models in Autonomous Driving V AD: Vectorized Scene Representation for Efficient Autonomous Driving
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 41261a53-7866-4512-91a1-9a6668329127 · outbound
STSBench: A Spatio-temporal Scenario Benchmark for Multi-modal Large Language Models in Autonomous Driving Senna: Bridging Large Vision-Language Models and End-to-End Autonomous Driving
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 94ebaa67-f270-4961-aa0e-bd90ebb95002 · outbound
STSBench: A Spatio-temporal Scenario Benchmark for Multi-modal Large Language Models in Autonomous Driving TOD3Cap: Towards 3D Dense Captioning in Outdoor Scenes
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 19f1f03c-586a-497f-a89a-7ac07cf885c6 · outbound
STSBench: A Spatio-temporal Scenario Benchmark for Multi-modal Large Language Models in Autonomous Driving Textual Explana- tions for Self-Driving Vehicles
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 3551ae27-52c8-49a8-a625-373da8b077b1 · outbound
STSBench: A Spatio-temporal Scenario Benchmark for Multi-modal Large Language Models in Autonomous Driving Unresolved cited work
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation dbaa139d-cb07-4b0c-b657-91268b39a258 · outbound
STSBench: A Spatio-temporal Scenario Benchmark for Multi-modal Large Language Models in Autonomous Driving BLIP-2: bootstrapping language-image pre-training with frozen image encoders and large language models
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8b01edf7-5b07-49ad-a107-6d521729b595 · outbound
STSBench: A Spatio-temporal Scenario Benchmark for Multi-modal Large Language Models in Autonomous Driving Generative Planning with 3D-vision Language Pre-training for End-to-End Autonomous Driving
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 132f6d3e-81e1-402f-a17b-9bad4f2ea6cb · outbound
STSBench: A Spatio-temporal Scenario Benchmark for Multi-modal Large Language Models in Autonomous Driving Automated Evaluation of Large Vision-Language Models on Self-driving Corner Cases
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 2700579a-bb60-4ee0-9928-9f50f1406279 · outbound
STSBench: A Spatio-temporal Scenario Benchmark for Multi-modal Large Language Models in Autonomous Driving BEVFormer: Learning Bird’s-Eye-View Representation from Multi-Camera Images via Spatiotemporal Transformers
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 240f776d-cfbb-458d-bf3c-6dbcf0b47e46 · outbound
STSBench: A Spatio-temporal Scenario Benchmark for Multi-modal Large Language Models in Autonomous Driving Unresolved cited work
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 002482ce-25a7-45be-8c9b-1c2510cab280 · outbound
STSBench: A Spatio-temporal Scenario Benchmark for Multi-modal Large Language Models in Autonomous Driving Visual Instruction Tuning
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 318f772c-2a14-4b7e-b0b1-3ab52449499d · outbound
STSBench: A Spatio-temporal Scenario Benchmark for Multi-modal Large Language Models in Autonomous Driving Improved Baselines with Visual Instruction Tuning
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bbd26351-f20e-4f2c-a10b-ca53f9602d7b · outbound
STSBench: A Spatio-temporal Scenario Benchmark for Multi-modal Large Language Models in Autonomous Driving Llava-next: Improved reasoning, ocr, and world knowledge, 2024
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 71a49614-32f8-45c9-8c13-79e52a879db5 · outbound
STSBench: A Spatio-temporal Scenario Benchmark for Multi-modal Large Language Models in Autonomous Driving Dolphins: Multimodal Language Model for Driving
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 013047aa-4b10-48d5-8b9b-9d903c2d2029 · outbound
STSBench: A Spatio-temporal Scenario Benchmark for Multi-modal Large Language Models in Autonomous Driving DRAMA: Joint Risk Localization and Captioning in Driving
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 41a19cf7-1158-4448-a836-ef7911dd0dc7 · outbound
STSBench: A Spatio-temporal Scenario Benchmark for Multi-modal Large Language Models in Autonomous Driving One Million Scenes for Autonomous Driving: ONCE Dataset
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 780325de-d703-4d4a-a403-3389f88e7dae · outbound
STSBench: A Spatio-temporal Scenario Benchmark for Multi-modal Large Language Models in Autonomous Driving LingoQA: Video Question Answering for Autonomous Driving
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation cb004d6a-7285-419a-972f-081708cce445 · outbound
STSBench: A Spatio-temporal Scenario Benchmark for Multi-modal Large Language Models in Autonomous Driving Reason2Drive: Towards Interpretable and Chain-based Reasoning for Autonomous Driving
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation cf4f7d79-26b2-4640-b2c3-9b424c0eaba3 · outbound
STSBench: A Spatio-temporal Scenario Benchmark for Multi-modal Large Language Models in Autonomous Driving GPT-4 Technical Report
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 134754d4-a4e7-447c-809d-0dce2daf5388 · outbound
STSBench: A Spatio-temporal Scenario Benchmark for Multi-modal Large Language Models in Autonomous Driving VLP: Vision Language Planning for Autonomous Driving
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation bcc8709b-3780-4a6a-af70-212dca4e3f26 · outbound
STSBench: A Spatio-temporal Scenario Benchmark for Multi-modal Large Language Models in Autonomous Driving NuScenes-QA: A Multi-Modal Visual Question Answering Benchmark for Autonomous Driving Scenario
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 97156a9e-4eca-40cb-9148-7172c33361f5 · outbound
STSBench: A Spatio-temporal Scenario Benchmark for Multi-modal Large Language Models in Autonomous Driving Learning Transferable Visual Models From Natural Language Supervision
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9763a031-1998-44a4-86f8-1f149144cc7b · outbound
STSBench: A Spatio-temporal Scenario Benchmark for Multi-modal Large Language Models in Autonomous Driving Rerun: A Visualization SDK for Multimodal Data, 2024
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 134d81e8-f02b-46f6-9c80-612713c386b6 · outbound
STSBench: A Spatio-temporal Scenario Benchmark for Multi-modal Large Language Models in Autonomous Driving Rank2Tell: A Multimodal Driving Dataset for Joint Importance Ranking and Reasoning
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 1593a902-c933-491f-b71a-d6ebb71ea76f · outbound
STSBench: A Spatio-temporal Scenario Benchmark for Multi-modal Large Language Models in Autonomous Driving Waslander, Yu Liu, and Hong- sheng Li
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 87b7c163-2644-4d34-b0d8-dffa49d438b4 · outbound
STSBench: A Spatio-temporal Scenario Benchmark for Multi-modal Large Language Models in Autonomous Driving DriveLM: Driving with Graph Visual Question Answering
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 41fddff6-2566-4c61-a976-f856dc4612b2 · outbound
STSBench: A Spatio-temporal Scenario Benchmark for Multi-modal Large Language Models in Autonomous Driving InsightDrive: Insight Scene Representation for End-to-End Autonomous Driving
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7b77396c-1291-433d-82d4-6ebf7af29670 · outbound
STSBench: A Spatio-temporal Scenario Benchmark for Multi-modal Large Language Models in Autonomous Driving Scalability in Perception for Autonomous Driving: Waymo Open Dataset
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation bb8925f4-fb50-4f58-8bc6-6e3acc12d6ab · outbound
STSBench: A Spatio-temporal Scenario Benchmark for Multi-modal Large Language Models in Autonomous Driving NuScenes-SpatialQA: A Spatial Understanding and Reasoning Benchmark for Vision-Language Models in Autonomous Driving
Reference 56
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f1ce3a9e-3718-4ac9-a315-490601e658e7 · outbound
STSBench: A Spatio-temporal Scenario Benchmark for Multi-modal Large Language Models in Autonomous Driving DriveVLM: The Convergence of Autonomous Driving and Large Vision-Language Models
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 860292ef-8592-4576-9437-ee71171cbe26 · outbound
STSBench: A Spatio-temporal Scenario Benchmark for Multi-modal Large Language Models in Autonomous Driving Object Referring in Videos With Language and Human Gaze
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 33cdbd46-4a8c-4df3-959f-0e485c860f1c · outbound
STSBench: A Spatio-temporal Scenario Benchmark for Multi-modal Large Language Models in Autonomous Driving Exploring Object- Centric Temporal Modeling for Efficient Multi-View 3D Object Detection
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation f89684e8-d231-42ff-8cd8-d8119f7c033c · outbound
STSBench: A Spatio-temporal Scenario Benchmark for Multi-modal Large Language Models in Autonomous Driving OmniDrive: A Holistic Vision-Language Dataset for Autonomous Driving with Counterfactual Reasoning
Reference 60
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cb4cbe43-40cb-4e80-8d76-ddc71f041efa · outbound
STSBench: A Spatio-temporal Scenario Benchmark for Multi-modal Large Language Models in Autonomous Driving Chain-of-thought prompting elicits reasoning in large language models
Reference 61
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d88d3f1b-67ba-4b72-8aba-72c12dd19460 · outbound
STSBench: A Spatio-temporal Scenario Benchmark for Multi-modal Large Language Models in Autonomous Driving Argoverse 2: Next Generation Datasets for Self-driving Perception and Forecasting
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation dfaf150f-7bbd-48cc-a930-5b5a60c4a71a · outbound
STSBench: A Spatio-temporal Scenario Benchmark for Multi-modal Large Language Models in Autonomous Driving BEVDriver: Leveraging BEV Maps in LLMs for Robust Closed-Loop Driving
Reference 63
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ad5f8d31-1693-43e6-83ed-ccb971147ec1 · outbound
STSBench: A Spatio-temporal Scenario Benchmark for Multi-modal Large Language Models in Autonomous Driving Language Prompt for Autonomous Driving
Reference 64
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation d4d6aed3-8deb-41f1-8318-014e26189f73 · outbound
STSBench: A Spatio-temporal Scenario Benchmark for Multi-modal Large Language Models in Autonomous Driving Are VLMs Ready for Autonomous Driving? An Empirical Study from the Reliability, Data, and Metric Perspectives
Reference 65
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 07fa72f2-17f2-4265-9bb6-c46b7207bb8c · outbound
STSBench: A Spatio-temporal Scenario Benchmark for Multi-modal Large Language Models in Autonomous Driving Explainable Object-Induced Action Decision for Autonomous Vehicles
Reference 66
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation e032a932-7c28-4897-8aa4-fbd41d5d5bad · outbound
STSBench: A Spatio-temporal Scenario Benchmark for Multi-modal Large Language Models in Autonomous Driving Drivegpt4: Interpretable end-to-end autonomous driving via large language model
Reference 67
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 223ca51b-e1e8-43c3-b3d9-18af9342e671 · outbound
STSBench: A Spatio-temporal Scenario Benchmark for Multi-modal Large Language Models in Autonomous Driving BDD100K: A Diverse Driving Dataset for Heterogeneous Multitask Learning
Reference 68
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 292ac349-af25-448f-a8da-beacad958f26 · outbound
STSBench: A Spatio-temporal Scenario Benchmark for Multi-modal Large Language Models in Autonomous Driving Are Vision LLMs Road-Ready? A Comprehensive Benchmark for Safety-Critical Driving Video Understanding
Reference 69
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c7220aac-9563-4639-9e3e-d34b828a3844 · outbound
STSBench: A Spatio-temporal Scenario Benchmark for Multi-modal Large Language Models in Autonomous Driving Sigmoid Loss for Language Image Pre-Training
Reference 70
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation dc59d335-10d2-41a7-887b-feecbf0eedf1 · outbound
STSBench: A Spatio-temporal Scenario Benchmark for Multi-modal Large Language Models in Autonomous Driving Sce2DriveX: A Generalized MLLM Framework for Scene-to-Drive Learning, 2025
Reference 71
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation dd65be83-e74e-4cc0-a07d-55f577fadc0f · outbound
STSBench: A Spatio-temporal Scenario Benchmark for Multi-modal Large Language Models in Autonomous Driving Large Language Models Are Not Robust Multiple Choice Selectors
Reference 72
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 0c025d56-caaf-488c-8a4b-a9c382724fa6 · outbound
STSBench: A Spatio-temporal Scenario Benchmark for Multi-modal Large Language Models in Autonomous Driving Doe-1: Closed-Loop Autonomous Driving with Large World Model
Reference 73
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 213a1e93-b9ef-4b95-b707-d07d4e80dfbd · outbound
STSBench: A Spatio-temporal Scenario Benchmark for Multi-modal Large Language Models in Autonomous Driving HERMES: A Unified Self-Driving World Model for Simultaneous 3D Scene Understanding and Generation
Reference 74
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eafc358d-10b8-4a46-b629-3beb1bc903f1 · outbound
STSBench: A Spatio-temporal Scenario Benchmark for Multi-modal Large Language Models in Autonomous Driving Unresolved cited work
Reference 75
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 05f11f15-3548-4c0d-9279-c56be095989d · outbound
STSBench: A Spatio-temporal Scenario Benchmark for Multi-modal Large Language Models in Autonomous Driving Embodied Understanding of Driving Scenarios
Reference 76
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation c267b539-9536-4bf2-8eaa-69e75c2ce4c1 · outbound
STSBench: A Spatio-temporal Scenario Benchmark for Multi-modal Large Language Models in Autonomous Driving Ego turning left,
Reference 77
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 57803ea2-d4b9-4db3-bc67-e67f7b813e83 · inbound
TPS-Drive: Task-Guided Representation Purification for VLM-based Autonomous Driving STSBench: A Spatio-temporal Scenario Benchmark for Multi-modal Large Language Models in Autonomous Driving
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 2a0ecd49-8cb3-4f76-8ee6-0f4f7b711920 · inbound
MVPruner: Dynamic Token Pruning for Accelerating Multi-view Vision-Language Models in Autonomous Driving STSBench: A Spatio-temporal Scenario Benchmark for Multi-modal Large Language Models in Autonomous Driving
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 2d3d9db3-be22-4727-97a6-fc5ab8d8f675 · inbound
MVPruner: Dynamic Token Pruning for Accelerating Multi-view Vision-Language Models in Autonomous Driving STSBench: A Spatio-temporal Scenario Benchmark for Multi-modal Large Language Models in Autonomous Driving
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 761ad56b-b155-434d-a190-58ebc25690a6 · inbound
CARLA-GS: Decoupling Representation, Reasoning, and Physics Simulation for Autonomous Driving Corner-Case Synthesis STSBench: A Spatio-temporal Scenario Benchmark for Multi-modal Large Language Models in Autonomous Driving
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 13b4fbaf-9454-4e66-91ec-dc9f651d025f · inbound
ObsDriveBench: Benchmarking Multimodal Understanding under Adverse Weather with Observability Awareness STSBench: A Spatio-temporal Scenario Benchmark for Multi-modal Large Language Models in Autonomous Driving
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9b79c5e9-0f91-48c2-a21f-3c066ff5d836 · inbound
STAR-VLM: Spatiotemporal Grounding Vision-Language Models for Motion and Velocity Estimation via Automotive Radar Supervision STSBench: A Spatio-temporal Scenario Benchmark for Multi-modal Large Language Models in Autonomous Driving
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.