Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 16 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 34 inbound Pith citation observations for arXiv:2405.10292.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-15T20:59:41.116130Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z
0 of 0 outbound references displayed
External citation measurements
8
arxiv_reference, observed 2026-08-05T02:28:24.338817Z
No outbound reference observations are available for this paper version.
Observation 354caf9c-1843-4282-97a9-d1b7c9a86b19 · inbound
Cambrian-1: A Fully Open, Vision-Centric Exploration of Multimodal LLMs Fine-Tuning Large Vision-Language Models as Decision-Making Agents via Reinforcement Learning
Reference 146
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 486a9b4b-31ee-43b4-a7e8-7f04da745aa0 · inbound
Generalist Virtual Agents: A Survey on Autonomous Agents Across Digital Platforms Fine-Tuning Large Vision-Language Models as Decision-Making Agents via Reinforcement Learning
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 685f869f-e7bc-4de9-9059-9fbdf4b8a96e · inbound
VisGraphVar: A Benchmark Generator for Assessing Variability in Graph Analysis Using Large Vision-Language Models Fine-Tuning Large Vision-Language Models as Decision-Making Agents via Reinforcement Learning
Reference 63
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4137cb0c-8afa-49a0-85d7-81fbf13774a8 · inbound
Large Language Model-Brained GUI Agents: A Survey Fine-Tuning Large Vision-Language Models as Decision-Making Agents via Reinforcement Learning
Reference 269
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation bc3804be-aeb5-4b4c-ac57-48e69c29ee40 · inbound
GRAPE: Generalizing Robot Policy via Preference Alignment Fine-Tuning Large Vision-Language Models as Decision-Making Agents via Reinforcement Learning
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1f0dc90a-fa6e-47f6-9815-d97d33952d16 · inbound
Dynamic Cross-Modal Alignment for Robust Semantic Location Prediction Fine-Tuning Large Vision-Language Models as Decision-Making Agents via Reinforcement Learning
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5a7b1f18-ff53-4682-aed8-2f50cb567739 · inbound
Dynamic Cross-Modal Alignment for Robust Semantic Location Prediction Fine-Tuning Large Vision-Language Models as Decision-Making Agents via Reinforcement Learning
Reference 2022
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 81961021-753c-4c11-b087-7455143f4495 · inbound
Bridging Vision and Language: Modeling Causality and Temporality in Video Narratives Fine-Tuning Large Vision-Language Models as Decision-Making Agents via Reinforcement Learning
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation efdb86d5-6a57-4f20-8773-6fa05cfdb689 · inbound
Optimizing Vision-Language Interactions Through Decoder-Only Models Fine-Tuning Large Vision-Language Models as Decision-Making Agents via Reinforcement Learning
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 13e4d032-0556-4830-963c-ac3152413ad6 · inbound
Temporal Contrastive Learning for Video Temporal Reasoning in Large Vision-Language Models Fine-Tuning Large Vision-Language Models as Decision-Making Agents via Reinforcement Learning
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 14ef6d89-a516-4e00-8e59-0c5756570613 · inbound
Leveraging Retrieval-Augmented Tags for Large Vision-Language Understanding in Complex Scenes Fine-Tuning Large Vision-Language Models as Decision-Making Agents via Reinforcement Learning
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 72f67eb6-1f26-478f-8704-989518038344 · inbound
Online Preference-based Reinforcement Learning with Self-augmented Feedback from Large Language Model Fine-Tuning Large Vision-Language Models as Decision-Making Agents via Reinforcement Learning
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f0b0febd-18cd-41bb-9ee6-6268daf9c69d · inbound
Training Software Engineering Agents and Verifiers with SWE-Gym Fine-Tuning Large Vision-Language Models as Decision-Making Agents via Reinforcement Learning
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation cc763e41-2d8b-49bc-a969-3b72d7af4a05 · inbound
An Overview and Discussion on Using Large Language Models for Implementation Generation of Solutions to Open-Ended Problems Fine-Tuning Large Vision-Language Models as Decision-Making Agents via Reinforcement Learning
Reference 105
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f803b801-e120-4d73-accd-1777b2049def · inbound
Large language models for artificial general intelligence (AGI): A survey of foundational principles and approaches Fine-Tuning Large Vision-Language Models as Decision-Making Agents via Reinforcement Learning
Reference 296
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 44d5c870-070d-4de4-99e3-785240b86460 · inbound
Dynamic Knowledge Integration for Enhanced Vision-Language Reasoning Fine-Tuning Large Vision-Language Models as Decision-Making Agents via Reinforcement Learning
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6af1a7e6-4817-452c-8227-94dc0a889695 · inbound
Improving Vision-Language-Action Model with Online Reinforcement Learning Fine-Tuning Large Vision-Language Models as Decision-Making Agents via Reinforcement Learning
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2c5a9231-ac47-4a73-8c04-215769c5d976 · inbound
RLS3: RL-Based Synthetic Sample Selection to Enhance Spatial Reasoning in Vision-Language Models for Indoor Autonomous Perception Fine-Tuning Large Vision-Language Models as Decision-Making Agents via Reinforcement Learning
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e7f93f0d-3498-41e3-b7d5-2d4e0b4658ce · inbound
CodeSteer: Symbolic-Augmented Language Models via Code/Text Guidance Fine-Tuning Large Vision-Language Models as Decision-Making Agents via Reinforcement Learning
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d0f68463-bb30-4748-9bd5-897400fff841 · inbound
Digi-Q: Learning Q-Value Functions for Training Device-Control Agents Fine-Tuning Large Vision-Language Models as Decision-Making Agents via Reinforcement Learning
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e1b7d85e-be24-4f27-acb9-2c7dc0e7d20f · inbound
Sample Efficient Reinforcement Learning via Large Vision Language Model Distillation Fine-Tuning Large Vision-Language Models as Decision-Making Agents via Reinforcement Learning
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bba994c9-ecbb-4d74-89c2-672229ac6d0a · inbound
Grounded Reinforcement Learning for Visual Reasoning Fine-Tuning Large Vision-Language Models as Decision-Making Agents via Reinforcement Learning
Reference 87
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 53650ddf-b229-4629-83b1-d4d6810ccd85 · inbound
Thinking vs. Doing: Agents that Reason by Scaling Test-Time Interaction Fine-Tuning Large Vision-Language Models as Decision-Making Agents via Reinforcement Learning
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a4ff4c6b-9d2e-4e44-9182-ff5299d37b8d · inbound
HALO: Human Preference Aligned Offline Reward Learning for Robot Navigation Fine-Tuning Large Vision-Language Models as Decision-Making Agents via Reinforcement Learning
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b18ca34b-7c07-4afa-a6ef-4bc5352c6030 · inbound
UAV-VL-R1: Generalizing Vision-Language Models via Supervised Fine-Tuning and Multi-Stage GRPO for UAV Visual Reasoning Fine-Tuning Large Vision-Language Models as Decision-Making Agents via Reinforcement Learning
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation b3a8be4e-9ebd-45f0-a26e-e792535251c5 · inbound
TCPO: Thought-Centric Preference Optimization for Effective Embodied Decision-making Fine-Tuning Large Vision-Language Models as Decision-Making Agents via Reinforcement Learning
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 35288331-b6b7-471c-92c2-4973e245f65b · inbound
GUI-Libra: Training Native GUI Agents to Reason and Act with Action-aware Supervision and Partially Verifiable RL Fine-Tuning Large Vision-Language Models as Decision-Making Agents via Reinforcement Learning
Reference 74
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ae5a3308-4ade-4f84-ba32-d0d685819d3e · inbound
PRISM: Perception Reasoning Interleaved for Sequential Decision Making Fine-Tuning Large Vision-Language Models as Decision-Making Agents via Reinforcement Learning
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 7cc824c7-d025-4952-8f58-3ee68f85ef0b · inbound
Reasoning Portability: Guiding Continual Learning for MLLMs in the RLVR Era Fine-Tuning Large Vision-Language Models as Decision-Making Agents via Reinforcement Learning
Reference 64
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation c20353e4-cf20-4934-97b5-fc85df66e18a · inbound
Qwen-RobotWorld Technical Report: Unifying Embodied World Modeling through Language-Conditioned Video Generation Fine-Tuning Large Vision-Language Models as Decision-Making Agents via Reinforcement Learning
Reference 234
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 72adbdd6-61dc-484b-97e8-4343b05cb815 · inbound
Rank-Then-Act: Reward-Free Control from Frame-Order Progress Fine-Tuning Large Vision-Language Models as Decision-Making Agents via Reinforcement Learning
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 577c25f7-95a5-4198-bcee-4fbe5d33cfec · inbound
Qwen-Audio-VAE Technical Report Fine-Tuning Large Vision-Language Models as Decision-Making Agents via Reinforcement Learning
Reference 241
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f1f80f45-9484-4442-bc68-1e528424f735 · inbound
PRISM: Priority-aware Rubric Internalization via Structured Multimodal Data Synthesis Fine-Tuning Large Vision-Language Models as Decision-Making Agents via Reinforcement Learning
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9c3be8fd-6365-445b-83df-0b0e4e6de7e3 · inbound
PRISM: Priority-aware Rubric Internalization via Structured Multimodal Data Synthesis Fine-Tuning Large Vision-Language Models as Decision-Making Agents via Reinforcement Learning
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.