Pith. sign in

Paper Citation Record · LEDGER

CarLLaVA: Vision language models for camera-only closed-loop driving

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 18 inbound Pith citation observations for arXiv:2406.10165.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2406.10165 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 18 of 18 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 18 of 18 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:24:07.325451Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T16:48:39.563623Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 6756f4cd-a63c-44c5-91cd-c115d8e8f6f7 · inbound

ORION: A Holistic End-to-End Autonomous Driving Framework by Vision-Language Instructed Action Generation cites this paper.

ORION: A Holistic End-to-End Autonomous Driving Framework by Vision-Language Instructed Action Generation CarLLaVA: Vision language models for camera-only closed-loop driving

Reference 46

Resolution
verified exact
arxiv_id, observed 2026-05-17T08:10:39.392246Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-17T08:10:39.240168Z digest=sha256:8772e19a811c224a1a54ed240cda4fcee366ad6ee86190530c8d1364f0b46c1d

Observation 3b29bd27-3d8c-4f05-8e06-16e447ee5871 · inbound

Generative AI for Autonomous Driving: A Review cites this paper.

Generative AI for Autonomous Driving: A Review CarLLaVA: Vision language models for camera-only closed-loop driving

Reference 253

Resolution
unresolved
no resolver link, observed 2026-08-07T15:24:07.325451Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:24:07.325451Z digest=sha256:00a751ab3f794541566e036a2bab23899d974c29992ed87e26c5cdf45485ea8b

Observation 72e4307d-33d0-4567-ae10-668bb23074c8 · inbound

Generalized Trajectory Scoring for End-to-end Multimodal Planning cites this paper.

Generalized Trajectory Scoring for End-to-end Multimodal Planning CarLLaVA: Vision language models for camera-only closed-loop driving

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T05:59:18.570814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:59:18.570814Z digest=sha256:fbc49317879f40586fdb80dfed2ee73ce5f87de5c457803ccdff3f59beb277c3

Observation db19ec5a-0383-4d55-927e-91ddafaea5de · inbound

ETA: Efficiency through Thinking Ahead, A Dual Approach to Self-Driving with Large Models cites this paper.

ETA: Efficiency through Thinking Ahead, A Dual Approach to Self-Driving with Large Models CarLLaVA: Vision language models for camera-only closed-loop driving

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T05:32:32.675801Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:32:32.675801Z digest=sha256:41a4ee1dc784c5575029ac72ce045c865a87d9c6309bdc096d268bb98b232a07

Observation dfcea971-7b05-491e-ad11-895c84303b51 · inbound

AutoVLA: A Vision-Language-Action Model for End-to-End Autonomous Driving with Adaptive Reasoning and Reinforcement Fine-Tuning cites this paper.

AutoVLA: A Vision-Language-Action Model for End-to-End Autonomous Driving with Adaptive Reasoning and Reinforcement Fine-Tuning CarLLaVA: Vision language models for camera-only closed-loop driving

Reference 41

Resolution
verified exact
arxiv_id, observed 2026-05-14T21:46:44.177559Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-14T21:46:43.955825Z digest=sha256:b517bf9a5495536a49e488ff9e00e74ec1850662eeefa420daf575c8df70fcef

Observation 10e4463e-a45e-4816-8eb5-f213050b1545 · inbound

LaViPlan : Language-Guided Visual Path Planning with RLVR cites this paper.

LaViPlan : Language-Guided Visual Path Planning with RLVR CarLLaVA: Vision language models for camera-only closed-loop driving

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T16:40:04.319609Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:40:04.319609Z digest=sha256:7a43f058f0a0e985ec8ff55572b773eb4309a7c1882956c594242044c3a44742

Observation 9ec1907e-0dfe-4b4b-88c5-64f011d8ef7b · inbound

DriveQA: Passing the Driving Knowledge Test cites this paper.

DriveQA: Passing the Driving Knowledge Test CarLLaVA: Vision language models for camera-only closed-loop driving

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-05T13:58:10.219174Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T13:58:10.219174Z digest=sha256:2b7dc3e6abb2dc14104f51fff038905cb19bcc2fff9f8234d947efd974409b3a

Observation 5152c52a-e923-4061-bc27-7bf10ca16f53 · inbound

CogDriver: Integrating Cognitive Inertia for Temporally Coherent Planning in Autonomous Driving cites this paper.

CogDriver: Integrating Cognitive Inertia for Temporally Coherent Planning in Autonomous Driving CarLLaVA: Vision language models for camera-only closed-loop driving

Reference 32

Resolution
verified exact
arxiv_id, observed 2026-05-18T20:11:50.808741Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-18T20:10:43.416488Z digest=sha256:7d2dbe0171fd9f2a853335533d25266f6ff237c719f8607682d5e9a1655670c9

Observation 390ac7d7-d185-435a-bca2-de21f12fd849 · inbound

DriveVLA-W0: World Models Amplify Data Scaling Law in Autonomous Driving cites this paper.

DriveVLA-W0: World Models Amplify Data Scaling Law in Autonomous Driving CarLLaVA: Vision language models for camera-only closed-loop driving

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-05-17T06:48:01.087290Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-17T06:48:00.943591Z digest=sha256:08d8ef568e1f5d3d1d80723b7e7ca8f120babaab1e646a17b6395dde39b28851

Observation 37984482-9bc7-4476-898c-5b83ac98484d · inbound

A Review of Learning-Based Motion Planning: Toward a Data-Driven Optimal Control Approach cites this paper.

A Review of Learning-Based Motion Planning: Toward a Data-Driven Optimal Control Approach CarLLaVA: Vision language models for camera-only closed-loop driving

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-03T16:52:55.423521Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:52:55.423521Z digest=sha256:71f0a4310d7a819dd19de856534a3821f0920f7db6789106eab33d155bc52888

Observation dbbaca0c-d952-4cf0-9b47-521a441d1a54 · inbound

AlignDrive: Aligned Lateral-Longitudinal Planning for End-to-End Autonomous Driving cites this paper.

AlignDrive: Aligned Lateral-Longitudinal Planning for End-to-End Autonomous Driving CarLLaVA: Vision language models for camera-only closed-loop driving

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-05-16T18:41:11.162098Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-16T18:39:49.416119Z digest=sha256:4b3db6ba3658b71d686ae8964229ed646031ef2b79e702079ce798bd97f9bfa3

Observation 71fddd0e-e487-4fac-85eb-8e36dc2e83ba · inbound

AlignDrive: Aligned Lateral-Longitudinal Planning for End-to-End Autonomous Driving cites this paper.

AlignDrive: Aligned Lateral-Longitudinal Planning for End-to-End Autonomous Driving CarLLaVA: Vision language models for camera-only closed-loop driving

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-03T12:48:35.811131Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T12:48:35.811131Z digest=sha256:54ac9fa15c0366ffcd29e81ff116c61a9f1ed6a243338d584f2170f4eabe8782

Observation 989695e6-72a6-4d0a-8ec7-4a54304ae37a · inbound

TaCarla: A comprehensive benchmarking dataset for end-to-end autonomous driving cites this paper.

TaCarla: A comprehensive benchmarking dataset for end-to-end autonomous driving CarLLaVA: Vision language models for camera-only closed-loop driving

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-02T20:21:31.075670Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T20:21:31.075670Z digest=sha256:95f8411b5ebcc392d824ce6a90e6f800bfa058df49260fd6a9d864e891558520

Observation 89353404-a066-42c3-8092-3df05a1a6272 · inbound

DVGT-2: Vision-Geometry-Action Model for Autonomous Driving at Scale cites this paper.

DVGT-2: Vision-Geometry-Action Model for Autonomous Driving at Scale CarLLaVA: Vision language models for camera-only closed-loop driving

Reference 59

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T23:03:25.120064Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-13T22:59:57.744015Z digest=sha256:ecbc9bb13fdff0409eec5129b1f71ec3611ac8dc48064c0486b857fb819be966

Observation 0c567676-f9eb-4692-ae88-df2d94edf05a · inbound

DeepSight: Long-Horizon World Modeling via Latent States Prediction for End-to-End Autonomous Driving cites this paper.

DeepSight: Long-Horizon World Modeling via Latent States Prediction for End-to-End Autonomous Driving CarLLaVA: Vision language models for camera-only closed-loop driving

Reference 53

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T06:31:25.869965Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-12T04:13:37.421188Z digest=sha256:65433d35fe862245f722be718cf192f5a353ba5596125451b0ffc9be27f6cf03

Observation 091db9d8-e3fb-4eb6-b07a-abe8e2cbf683 · inbound

LVDrive: Latent Visual Representation Enhanced Vision-Language-Action Autonomous Driving Model cites this paper.

LVDrive: Latent Visual Representation Enhanced Vision-Language-Action Autonomous Driving Model CarLLaVA: Vision language models for camera-only closed-loop driving

Reference 41

Resolution
verified exact
arxiv_id, observed 2026-05-22T07:41:15.037981Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-22T07:37:11.292270Z digest=sha256:a820001bef62ab4dd116dd6f741792acccaa420e546ff153467078dcc8cbd11d

Observation f7076c21-e939-4102-aeda-032225d7d17d · inbound

VLGA: Vision-Language-Geometry-Action Models for Autonomous Driving cites this paper.

VLGA: Vision-Language-Geometry-Action Models for Autonomous Driving CarLLaVA: Vision language models for camera-only closed-loop driving

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-07-03T11:08:03.416062Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-27T09:40:51.407286Z digest=sha256:d768999e0fadb9cc01258d332f6dda9e819126582829558babf0d54107cb71ab

Observation 4843126b-cb6e-4163-8a59-290a4e8d47e7 · inbound

Teaching Vision-Language-Action Models What to See and Where to Look cites this paper.

Teaching Vision-Language-Action Models What to See and Where to Look CarLLaVA: Vision language models for camera-only closed-loop driving

Reference 44

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T16:48:39.564964Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-07-03T16:42:13.520913Z digest=sha256:de0a1f30957e5d0529d65147a751e16d8a59afebdf3a8ca4427e6cbcf2ff5c09