Pith. sign in

Paper Citation Record · LEDGER

VLM-RL: A Unified Vision Language Models and Reinforcement Learning Framework for Safe Autonomous Driving

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 14 inbound Pith citation observations for arXiv:2412.15544.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2412.15544 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 14 of 14 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 14 of 14 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:13:41.080485Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-05T11:41:02.586060Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 028119a4-8591-4506-b4d8-fd400be7334f · inbound

HCRMP: A LLM-Hinted Contextual Reinforcement Learning Framework for Autonomous Driving cites this paper.

HCRMP: A LLM-Hinted Contextual Reinforcement Learning Framework for Autonomous Driving VLM-RL: A Unified Vision Language Models and Reinforcement Learning Framework for Safe Autonomous Driving

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T15:13:41.080485Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:13:41.080485Z digest=sha256:b1a33c627664fa173f2dcda1e81dd0803d8038b194112089617db53799b2caed

Observation 2f4ced2b-76a2-48c8-a2af-6c52e7d28f00 · inbound

Simulating the Unseen: Crash Prediction Must Learn from What Did Not Happen cites this paper.

Simulating the Unseen: Crash Prediction Must Learn from What Did Not Happen VLM-RL: A Unified Vision Language Models and Reinforcement Learning Framework for Safe Autonomous Driving

Reference 149

Resolution
unresolved
no resolver link, observed 2026-08-07T13:28:34.208933Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:28:34.208933Z digest=sha256:298cff08ae1dfb7d52cc4885f1bf578b209b486c35de61626747e7c16d17d840

Observation 1c2138c5-3554-40fe-b0f0-fa912caa0c04 · inbound

ROAD: Responsibility-Oriented Reward Design for Reinforcement Learning in Autonomous Driving cites this paper.

ROAD: Responsibility-Oriented Reward Design for Reinforcement Learning in Autonomous Driving VLM-RL: A Unified Vision Language Models and Reinforcement Learning Framework for Safe Autonomous Driving

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T12:32:06.483874Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:32:06.483874Z digest=sha256:f0780cd0010ae76410767656888875e33e04bd4c3741c8cec497676d541db1d1

Observation 58f4aa1f-e3bd-4561-ae73-8b69edcf6da4 · inbound

SEAL: Vision-Language Model-Based Safe End-to-End Cooperative Autonomous Driving with Adaptive Long-Tail Modeling cites this paper.

SEAL: Vision-Language Model-Based Safe End-to-End Cooperative Autonomous Driving with Adaptive Long-Tail Modeling VLM-RL: A Unified Vision Language Models and Reinforcement Learning Framework for Safe Autonomous Driving

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-06T22:40:47.335989Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:40:47.335989Z digest=sha256:a4a93636efa6013b6a7f1ee315a0f5ed71dca50cc171937a149457625573432d

Observation c46c8d11-f331-4e38-8b22-09967ed31c8b · inbound

A Survey on Vision-Language-Action Models for Autonomous Driving cites this paper.

A Survey on Vision-Language-Action Models for Autonomous Driving VLM-RL: A Unified Vision Language Models and Reinforcement Learning Framework for Safe Autonomous Driving

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.482447Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.482447Z digest=sha256:3186f82b725684ff7dac48e4880edf8a7e9a57208081bfccc7788e85cfdeeae7

Observation eb5915ce-7eb7-47b9-893a-10c4cb984cb7 · inbound

One Object, Multiple Lies: A Benchmark for Cross-task Adversarial Attack on Unified Vision-Language Models cites this paper.

One Object, Multiple Lies: A Benchmark for Cross-task Adversarial Attack on Unified Vision-Language Models VLM-RL: A Unified Vision Language Models and Reinforcement Learning Framework for Safe Autonomous Driving

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T18:40:53.704274Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:40:53.704274Z digest=sha256:d10ccd143324aa41050050bfba62a8987a72b9ee4eb4bd8a15fe96acb9eacb78

Observation f0c4480c-62e3-4933-a1c5-481fa5d4ee31 · inbound

Edge-Based Multimodal Sensor Data Fusion with Vision Language Models (VLMs) for Real-time Autonomous Vehicle Accident Avoidance cites this paper.

Edge-Based Multimodal Sensor Data Fusion with Vision Language Models (VLMs) for Real-time Autonomous Vehicle Accident Avoidance VLM-RL: A Unified Vision Language Models and Reinforcement Learning Framework for Safe Autonomous Driving

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-06T05:57:51.136342Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T05:57:51.136342Z digest=sha256:940660a24a853363e357e568775c4e49a9ef4820572bf78d95a9ad912a0d7d1a

Observation b57e86bd-7545-46b9-8b81-d17fe70ea8f7 · inbound

AutoNeural: Co-Designing Vision-Language Models for NPU Inference cites this paper.

AutoNeural: Co-Designing Vision-Language Models for NPU Inference VLM-RL: A Unified Vision Language Models and Reinforcement Learning Framework for Safe Autonomous Driving

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-03T18:56:58.078201Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:56:58.078201Z digest=sha256:6ec568d3ac8e32fe06ca90e4ae47cf3eee0574b2de35777c35e7ee728a67446d

Observation 3aa3a8c6-6904-4002-94b6-e498c2a4ccf4 · inbound

AnySlot: Goal-Conditioned Vision-Language-Action Policies for Zero-Shot Slot-Level Placement cites this paper.

AnySlot: Goal-Conditioned Vision-Language-Action Policies for Zero-Shot Slot-Level Placement VLM-RL: A Unified Vision Language Models and Reinforcement Learning Framework for Safe Autonomous Driving

Reference 13

Resolution
unresolved
no resolver link, observed 2026-07-12T22:40:19.969447Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T22:40:19.969447Z digest=sha256:af43c7417a8699b328f6288c270cfc7018a1f473602c74f66162cd0bd84dfc12

Observation b3b1b85c-1c3d-45b4-bf9b-b9239330a597 · inbound

SignReasoner: Compositional Reasoning for Complex Traffic Sign Understanding via Functional Structure Units cites this paper.

SignReasoner: Compositional Reasoning for Complex Traffic Sign Understanding via Functional Structure Units VLM-RL: A Unified Vision Language Models and Reinforcement Learning Framework for Safe Autonomous Driving

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-11T09:31:04.630341Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T15:58:48.213961Z digest=sha256:f7464f561865a813a856114242e095e9fe775338b9f517469ecd183d72bad38c

Observation dec2d05e-bc38-4115-874d-5fe13a9604cf · inbound

Steadily moving semi-infinite fracture in plane poroelasticity cites this paper.

Steadily moving semi-infinite fracture in plane poroelasticity VLM-RL: A Unified Vision Language Models and Reinforcement Learning Framework for Safe Autonomous Driving

Reference 40

Resolution
metadata mismatch
local_arxiv, observed 2026-07-05T11:41:02.587922Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-07-05T11:39:05.686584Z digest=sha256:d6d6aa85dfb52413787023a33912c55806bf4080f3ff464107a0c07aa5b645ce

Observation 762982d2-7824-4920-86da-e02d1c584585 · inbound

XEmbodied: A Foundation Model with Enhanced Geometric and Physical Cues for Large-Scale Embodied Environments cites this paper.

XEmbodied: A Foundation Model with Enhanced Geometric and Physical Cues for Large-Scale Embodied Environments VLM-RL: A Unified Vision Language Models and Reinforcement Learning Framework for Safe Autonomous Driving

Reference 40

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T05:51:10.323450Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T05:46:36.865150Z digest=sha256:727ec2febe1a2ecae019adb65e3807e507b47224bc2ad2db3140fe23b8a7d58b

Observation 2fe49ae3-9b5c-49d6-a295-1261ba8bbf74 · inbound

Language-Driven Cost Optimization for Autonomous Driving cites this paper.

Language-Driven Cost Optimization for Autonomous Driving VLM-RL: A Unified Vision Language Models and Reinforcement Learning Framework for Safe Autonomous Driving

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-07-03T05:17:40.113249Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-27T13:21:51.059859Z digest=sha256:f950deb1ed0690cbf5972817cc6c40ceb4b9c89b44839b244fb232265fdaa3b9

Observation 07ea0aa7-5160-4499-a7c5-007f0a15717c · inbound

Seeing Before Colliding: Anticipatory Safe RL with Frozen Vision-Language Models cites this paper.

Seeing Before Colliding: Anticipatory Safe RL with Frozen Vision-Language Models VLM-RL: A Unified Vision Language Models and Reinforcement Learning Framework for Safe Autonomous Driving

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-07-03T04:17:36.743391Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-27T14:05:48.011688Z digest=sha256:11bbc905ef13a469f92f1062db32e86c02658282406d292fa75618f5b99e5cab