Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-15T16:26:15.704069Z
Paper Citation Record · LEDGER
As of 22 August 2026, this Paper Citation Record lists 24 of 24 outbound references and 9 inbound Pith citation observations for arXiv:2509.05578.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-15T16:26:15.704069Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-22T06:32:14.747728+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-02T05:06:33.852576Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-03T16:48:39.596474Z
24 of 24 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation a4b40f5c-cdb8-4a1f-95d9-83410d760f0d · outbound
OccVLA: Vision-Language-Action Model with Implicit 3D Occupancy Supervision PaliGemma: A versatile 3B VLM for transfer
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 819333d0-c886-42c1-9d99-6b3e344ae725 · outbound
OccVLA: Vision-Language-Action Model with Implicit 3D Occupancy Supervision VAD: Vectorized Scene Representation for Efficient Autonomous Driving
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 73d694dd-ebb0-4090-a1af-96ff63c461aa · outbound
OccVLA: Vision-Language-Action Model with Implicit 3D Occupancy Supervision OccTransformer: Improving BEVFormer for 3D camera-only occupancy prediction
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9d3f5af6-f056-44de-949c-e41152da9c8f · outbound
OccVLA: Vision-Language-Action Model with Implicit 3D Occupancy Supervision Decoupled Weight Decay Regularization
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fa17e969-aa3e-46e7-b6e7-506c6ea1da85 · outbound
OccVLA: Vision-Language-Action Model with Implicit 3D Occupancy Supervision URLhttps: //aclanthology.org/2023.emnlp-demo.13
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 25d3b5ef-29ff-4d55-80cb-aa3b36177bc3 · outbound
OccVLA: Vision-Language-Action Model with Implicit 3D Occupancy Supervision PaliGemma 2: A Family of Versatile VLMs for Transfer
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 082a111f-0e5b-48a0-b8b0-563b5fea803a · outbound
OccVLA: Vision-Language-Action Model with Implicit 3D Occupancy Supervision DriveVLM: The Convergence of Autonomous Driving and Large Vision-Language Models
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6e7633ad-96e6-4c6f-afa0-43398143f7d1 · outbound
OccVLA: Vision-Language-Action Model with Implicit 3D Occupancy Supervision LLaMA: Open and Efficient Foundation Language Models
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9529e215-2148-429d-b64d-a27a92768539 · outbound
OccVLA: Vision-Language-Action Model with Implicit 3D Occupancy Supervision 12 Preprint
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9c5496ad-3f5b-4efd-bd2b-2e6cf53373d4 · outbound
OccVLA: Vision-Language-Action Model with Implicit 3D Occupancy Supervision OccLLaMA: An Occupancy-Language-Action Generative World Model for Autonomous Driving
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a4577109-337f-4ca3-a1f6-62ee9d179133 · outbound
OccVLA: Vision-Language-Action Model with Implicit 3D Occupancy Supervision Neural Map Prior for Autonomous Driving
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 88242e4b-b939-4d61-8a3c-37e140382b10 · outbound
OccVLA: Vision-Language-Action Model with Implicit 3D Occupancy Supervision DriveGPT4: Interpretable End-to-end Autonomous Driving via Large Language Model
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 60ce2586-7d5f-403a-87c1-2c073a800c28 · outbound
OccVLA: Vision-Language-Action Model with Implicit 3D Occupancy Supervision LiDAR-LLM: Exploring the Potential of Large Language Models for 3D LiDAR Understanding
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6b063f9d-131a-4842-b71c-41bda42fa56b · outbound
OccVLA: Vision-Language-Action Model with Implicit 3D Occupancy Supervision Rethinking the Open-Loop Evaluation of End-to-End Autonomous Driving in nuScenes
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9a38f73e-77f9-459d-b1a7-197c79fd9cbf · outbound
OccVLA: Vision-Language-Action Model with Implicit 3D Occupancy Supervision OccFormer: Dual-path Transformer for Vision-based 3D Semantic Occupancy Prediction
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eb094116-c308-4135-9b69-edf76732ea58 · outbound
OccVLA: Vision-Language-Action Model with Implicit 3D Occupancy Supervision DriveAgent-R1: Advancing VLM-based Autonomous Driving with Active Perception and Hybrid Thinking
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f5f2124b-ce8d-45d3-99ee-fc272b619a11 · outbound
OccVLA: Vision-Language-Action Model with Implicit 3D Occupancy Supervision Unresolved cited work
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 18b6d784-41e2-4682-ad8a-83096becfd5e · outbound
OccVLA: Vision-Language-Action Model with Implicit 3D Occupancy Supervision NuScenes-QA: A Multi-modal Visual Question Answering Benchmark for Autonomous Driving Scenario
Reference 2017
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e9d7a048-9be0-4e12-a0d1-5dcbd3522690 · outbound
OccVLA: Vision-Language-Action Model with Implicit 3D Occupancy Supervision Bridging Stereo Geometry and BEV Representation with Reliable Mutual Interaction for Semantic Scene Completion
Reference 2019
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 805e8997-b38a-4cd5-b1b7-7b2e99b2ea31 · outbound
OccVLA: Vision-Language-Action Model with Implicit 3D Occupancy Supervision Adapters: A unified li- brary for parameter-efficient and modular transfer learning
Reference 2020
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 089e9a7c-596e-4ced-a9e9-c6c33fdb4d75 · outbound
OccVLA: Vision-Language-Action Model with Implicit 3D Occupancy Supervision ST-P3: End-to-end Vision-based Autonomous Driving via Spatial-Temporal Feature Learning
Reference 2022
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eea1e51a-4e1e-435e-b1f1-86b1b4a37f72 · outbound
OccVLA: Vision-Language-Action Model with Implicit 3D Occupancy Supervision Tri-Perspective View for Vision-Based 3D Semantic Occupancy Prediction
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a37b57bd-8291-4f65-96e3-5f1846e09a0d · outbound
OccVLA: Vision-Language-Action Model with Implicit 3D Occupancy Supervision GPT-4 Technical Report
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fc05f3a1-bb1c-445f-8eda-3a47119e9884 · outbound
OccVLA: Vision-Language-Action Model with Implicit 3D Occupancy Supervision LMFusion: Adapting Pretrained Language Models for Multimodal Generation
Reference 2025
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5f1683f6-57be-4be2-9d80-8e61963892de · inbound
ShelfGaussian: Shelf-Supervised Open-Vocabulary Gaussian-based 3D Scene Understanding OccVLA: Vision-Language-Action Model with Implicit 3D Occupancy Supervision
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 9f063d58-4fb3-4dd6-8835-ef025ac71ba9 · inbound
Monocular Open Vocabulary Occupancy Prediction for Indoor Scenes OccVLA: Vision-Language-Action Model with Implicit 3D Occupancy Supervision
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 5ffcdbe2-573d-4277-9544-3875d8f2d94a · inbound
The Latent Space: Foundation, Evolution, Mechanism, Ability, and Outlook OccVLA: Vision-Language-Action Model with Implicit 3D Occupancy Supervision
Reference 124
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 129cd057-4381-4156-be54-a34a21c1153c · inbound
Height-Guided Projection Reparameterization for Camera-LiDAR Occupancy OccVLA: Vision-Language-Action Model with Implicit 3D Occupancy Supervision
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 698acbad-fd1c-4fd1-8a42-201e9edaf122 · inbound
Height-Guided Projection Reparameterization for Camera-LiDAR Occupancy OccVLA: Vision-Language-Action Model with Implicit 3D Occupancy Supervision
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 207aa84a-c5bf-4bc7-96ca-eaf65bf772c0 · inbound
TPS-Drive: Task-Guided Representation Purification for VLM-based Autonomous Driving OccVLA: Vision-Language-Action Model with Implicit 3D Occupancy Supervision
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 673c8019-968e-4da9-89d5-713c18b52b3e · inbound
M$^3$Eval: Multi-Modal Memory Evaluation through Cognitively-Grounded Video Tasks OccVLA: Vision-Language-Action Model with Implicit 3D Occupancy Supervision
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 1a5a9899-08ed-4210-ba36-436a01495968 · inbound
Teaching Vision-Language-Action Models What to See and Where to Look OccVLA: Vision-Language-Action Model with Implicit 3D Occupancy Supervision
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 8d12a04d-ee7a-4419-9c93-78de1d3c9166 · inbound
GPOcc++: Unified Sparse Gaussian Occupancy Prediction with Visual Geometry Priors OccVLA: Vision-Language-Action Model with Implicit 3D Occupancy Supervision
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.