Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-15T20:21:40.641794Z
Paper Citation Record · LEDGER
As of 16 August 2026, this Paper Citation Record lists 48 of 48 outbound references and 12 inbound Pith citation observations for arXiv:2505.13426.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-15T20:21:40.641794Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-15T23:21:12.064728Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z
48 of 48 outbound references displayed
External citation measurements
0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z
Observation 632309ab-ceac-4221-83a5-87b08d498007 · outbound
G1: Bootstrapping Perception and Reasoning Abilities of Vision-Language Model via Reinforcement Learning Unresolved cited work
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 24134682-60cd-461a-b811-80fa3ec1cec0 · outbound
G1: Bootstrapping Perception and Reasoning Abilities of Vision-Language Model via Reinforcement Learning Qwen2.5-VL Technical Report
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 279d5243-cd23-4ba9-affe-6dd70ebc6c66 · outbound
G1: Bootstrapping Perception and Reasoning Abilities of Vision-Language Model via Reinforcement Learning Deep blue.Artificial intelligence, 134(1-2):57–83, 2002
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3fec197c-b599-4d52-8449-4d313ca7646f · outbound
G1: Bootstrapping Perception and Reasoning Abilities of Vision-Language Model via Reinforcement Learning R1-v: Reinforcing super generalization ability in vision-language models with less than $3
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation f83a8618-2e4b-4f9d-80ce-38bf088742d0 · outbound
G1: Bootstrapping Perception and Reasoning Abilities of Vision-Language Model via Reinforcement Learning Next token prediction towards multimodal intelligence: A comprehensive survey, 2024
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 369135b3-c9ed-42dd-8a5e-39620b1cf007 · outbound
G1: Bootstrapping Perception and Reasoning Abilities of Vision-Language Model via Reinforcement Learning Pca-bench: Evaluating multimodal large language models in perception-cognition-action chain, 2024
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 1a6d9cb4-f391-46e4-aaeb-b0cca2fb1a16 · outbound
G1: Bootstrapping Perception and Reasoning Abilities of Vision-Language Model via Reinforcement Learning Can VLMs Play Action Role-Playing Games? Take Black Myth Wukong as a Study Case
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation efd27942-730f-4235-bcc4-794e94b13ce8 · outbound
G1: Bootstrapping Perception and Reasoning Abilities of Vision-Language Model via Reinforcement Learning Unresolved cited work
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 67ab1055-ff75-4e04-977b-dcc0f3ffde41 · outbound
G1: Bootstrapping Perception and Reasoning Abilities of Vision-Language Model via Reinforcement Learning GTBench: Uncovering the Strategic Reasoning Limitations of LLMs via Game-Theoretic Evaluations
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1c4d1528-a51b-4f91-89a1-44542b429875 · outbound
G1: Bootstrapping Perception and Reasoning Abilities of Vision-Language Model via Reinforcement Learning FlowReasoner: Reinforcing Query-Level Meta-Agents
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c351b643-633e-4b3b-8341-6134f6b7f1a8 · outbound
G1: Bootstrapping Perception and Reasoning Abilities of Vision-Language Model via Reinforcement Learning VLM Q-Learning: Aligning Vision-Language Models for Interactive Decision-Making
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9669f4b8-7798-426c-a7bb-20595bb0425e · outbound
G1: Bootstrapping Perception and Reasoning Abilities of Vision-Language Model via Reinforcement Learning Mmevalpro: Calibrating multimodal benchmarks towards trustworthy and efficient evaluation, 2025
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 8f7321ce-28c0-413c-8135-2ff7b7a81489 · outbound
G1: Bootstrapping Perception and Reasoning Abilities of Vision-Language Model via Reinforcement Learning Vision-R1: Incentivizing Reasoning Capability in Multimodal Large Language Models
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4dc716f5-6070-476c-85a1-73e1b3daf1c9 · outbound
G1: Bootstrapping Perception and Reasoning Abilities of Vision-Language Model via Reinforcement Learning Learning multiple layers of features from tiny images
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5854a8e8-56d2-44d4-8836-6ee06d98abf4 · outbound
G1: Bootstrapping Perception and Reasoning Abilities of Vision-Language Model via Reinforcement Learning JARVIS-VLA: Post-Training Large-Scale Vision Language Models to Play Visual Games with Keyboards and Mouse
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2acb931c-f258-4e73-a7b0-304dbf0627d0 · outbound
G1: Bootstrapping Perception and Reasoning Abilities of Vision-Language Model via Reinforcement Learning MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 554de052-0550-4e9e-8ce0-6b395012918b · outbound
G1: Bootstrapping Perception and Reasoning Abilities of Vision-Language Model via Reinforcement Learning Playing atari with deep reinforcement learning, 2013
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0520e4b8-2b95-49be-bc5d-68c3f11b8e40 · outbound
G1: Bootstrapping Perception and Reasoning Abilities of Vision-Language Model via Reinforcement Learning Unresolved cited work
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1a33699d-6603-444a-bd60-73df5d3a53f2 · outbound
G1: Bootstrapping Perception and Reasoning Abilities of Vision-Language Model via Reinforcement Learning Gpt-4v(ision) system card
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fa0cbb04-9ab0-4d0b-8e6c-9f2ed2d4b218 · outbound
G1: Bootstrapping Perception and Reasoning Abilities of Vision-Language Model via Reinforcement Learning hello-gpt-4o, 2024
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 4bb05e01-a1b2-4078-8281-07f3f2c907c3 · outbound
G1: Bootstrapping Perception and Reasoning Abilities of Vision-Language Model via Reinforcement Learning BALROG: Benchmarking Agentic LLM and VLM Reasoning On Games
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 937eff5f-c397-41e7-9755-fe62a921f8a5 · outbound
G1: Bootstrapping Perception and Reasoning Abilities of Vision-Language Model via Reinforcement Learning Generative agents: Interactive simulacra of human behavior
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 8360e5ef-6d8b-4245-9f09-f53d9aac24f2 · outbound
G1: Bootstrapping Perception and Reasoning Abilities of Vision-Language Model via Reinforcement Learning Llms are greedy agents: Effects of rl fine-tuning on decision-making abilities, 2025
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation feeea2f0-153e-4c9d-9859-c3dff08eeac0 · outbound
G1: Bootstrapping Perception and Reasoning Abilities of Vision-Language Model via Reinforcement Learning Unresolved cited work
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation aed021ba-1089-49ad-b927-6d7003827db8 · outbound
G1: Bootstrapping Perception and Reasoning Abilities of Vision-Language Model via Reinforcement Learning VLM-R1: A Stable and Generalizable R1-style Large Vision-Language Model
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 37e6be08-a1fd-43a2-b726-7067a8694a7d · outbound
G1: Bootstrapping Perception and Reasoning Abilities of Vision-Language Model via Reinforcement Learning HybridFlow: A Flexible and Efficient RLHF Framework
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e0eaf1a4-9c0b-49a7-88a5-79cf97e50fc1 · outbound
G1: Bootstrapping Perception and Reasoning Abilities of Vision-Language Model via Reinforcement Learning Maddison, Arthur Guez, L
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5a6cad25-9eab-41cf-82e2-5771c660ffbb · outbound
G1: Bootstrapping Perception and Reasoning Abilities of Vision-Language Model via Reinforcement Learning Mastering the game of go with deep neural networks and tree search.nature, 529(7587):484–489, 2016
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e768e2fc-e5e3-4485-bdd2-97e99ee52e0d · outbound
G1: Bootstrapping Perception and Reasoning Abilities of Vision-Language Model via Reinforcement Learning Cradle: Empowering Foundation Agents Towards General Computer Control
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 75fc430f-ef66-4d7d-bd53-b8427f8ad369 · outbound
G1: Bootstrapping Perception and Reasoning Abilities of Vision-Language Model via Reinforcement Learning Unresolved cited work
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 4082641c-36e1-4276-ada1-bc9f938798ce · outbound
G1: Bootstrapping Perception and Reasoning Abilities of Vision-Language Model via Reinforcement Learning Kimi k1.5: Scaling Reinforcement Learning with LLMs
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 33a03cbd-d4af-4138-b4a4-741a8feb7f18 · outbound
G1: Bootstrapping Perception and Reasoning Abilities of Vision-Language Model via Reinforcement Learning Unresolved cited work
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5c4bc011-e5b3-4f25-bb36-774bc6d6d5e7 · outbound
G1: Bootstrapping Perception and Reasoning Abilities of Vision-Language Model via Reinforcement Learning Unresolved cited work
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 31b151e0-0773-48b5-b0e1-ef04023fb839 · outbound
G1: Bootstrapping Perception and Reasoning Abilities of Vision-Language Model via Reinforcement Learning Can Large Language Models Play Text Games Well? Current State-of-the-Art and Open Questions
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5389e299-c36a-421e-b639-b6871230a2fa · outbound
G1: Bootstrapping Perception and Reasoning Abilities of Vision-Language Model via Reinforcement Learning Are large vision language models good game players?, 2025
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7a376b86-59df-4aeb-b987-b5ce65fecef5 · outbound
G1: Bootstrapping Perception and Reasoning Abilities of Vision-Language Model via Reinforcement Learning Are Large Vision Language Models Good Game Players?
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e8c5e5f7-f051-476a-84ad-7f46718e2760 · outbound
G1: Bootstrapping Perception and Reasoning Abilities of Vision-Language Model via Reinforcement Learning Describe, Explain, Plan and Select: Interactive Planning with Large Language Models Enables Open-World Multi-Task Agents
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 88cd7c0a-e4be-42db-8f9b-9298ab9bb032 · outbound
G1: Bootstrapping Perception and Reasoning Abilities of Vision-Language Model via Reinforcement Learning Waytowich, Devin White, MD Sunbeam, and Vinicius G
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 3c5628f0-b510-41b6-afb7-d5133c637cf5 · outbound
G1: Bootstrapping Perception and Reasoning Abilities of Vision-Language Model via Reinforcement Learning SmartPlay: A Benchmark for LLMs as Intelligent Agents
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 49c54ada-c4b2-4a80-bd78-2547dfcf2463 · outbound
G1: Bootstrapping Perception and Reasoning Abilities of Vision-Language Model via Reinforcement Learning R1-Onevision: Advancing Generalized Multimodal Reasoning through Cross-Modal Formalization
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4f129bf9-27fd-4024-8774-5aec1655a0ac · outbound
G1: Bootstrapping Perception and Reasoning Abilities of Vision-Language Model via Reinforcement Learning Fine-tuning large vision-language models as decision-making agents via reinforcement learning
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 30b7371c-8f7b-4be9-ba6d-6c5884c4719a · outbound
G1: Bootstrapping Perception and Reasoning Abilities of Vision-Language Model via Reinforcement Learning Easyr1: An efficient, scalable, multi-modality rl training framework
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 16eff4c8-22ad-4de0-b4b0-9a1c1f59bb4d · outbound
G1: Bootstrapping Perception and Reasoning Abilities of Vision-Language Model via Reinforcement Learning Llamafactory: Unified efficient fine-tuning of 100+ language models
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 192e4191-6ac2-43c1-8814-1fbb634907e1 · outbound
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 742fffcb-b53a-4d78-9b03-812468680c52 · outbound
G1: Bootstrapping Perception and Reasoning Abilities of Vision-Language Model via Reinforcement Learning Output Format Description First describe the board in <perception></perception>
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 5946c486-f3bb-4bcd-8867-09cd084c5fa0 · outbound
G1: Bootstrapping Perception and Reasoning Abilities of Vision-Language Model via Reinforcement Learning (row1,col1) (row2,col2)
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation af7af115-aad6-4ed2-9647-c2359f32d881 · outbound
G1: Bootstrapping Perception and Reasoning Abilities of Vision-Language Model via Reinforcement Learning Unresolved cited work
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation ce5ad0eb-10e1-4d55-855b-a909c6ce8a03 · outbound
G1: Bootstrapping Perception and Reasoning Abilities of Vision-Language Model via Reinforcement Learning Output Format Description First describe the board in <perception></perception>
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation b6a2d442-edd3-43fe-b415-f99bd643a4e8 · inbound
Perception, Reason, Think, and Plan: A Survey on Large Multimodal Reasoning Models G1: Bootstrapping Perception and Reasoning Abilities of Vision-Language Model via Reinforcement Learning
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 562becc3-0bf1-4cd9-b8f0-829ea6af43a6 · inbound
UAV-VL-R1: Generalizing Vision-Language Models via Supervised Fine-Tuning and Multi-Stage GRPO for UAV Visual Reasoning G1: Bootstrapping Perception and Reasoning Abilities of Vision-Language Model via Reinforcement Learning
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation be41000b-fa61-4774-89c7-5dbac2f99424 · inbound
Explain Before You Answer: A Survey on Compositional Visual Reasoning G1: Bootstrapping Perception and Reasoning Abilities of Vision-Language Model via Reinforcement Learning
Reference 159
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3c920247-9f73-4a74-b31d-de700b27b369 · inbound
The Landscape of Agentic Reinforcement Learning for LLMs: A Survey G1: Bootstrapping Perception and Reasoning Abilities of Vision-Language Model via Reinforcement Learning
Reference 239
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation bd874107-7871-46f7-a6e0-00b32ff204d2 · inbound
A Survey of Reinforcement Learning for Large Reasoning Models G1: Bootstrapping Perception and Reasoning Abilities of Vision-Language Model via Reinforcement Learning
Reference 61
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 956c631a-0ef5-4bf5-8f4f-2c7720d85ba0 · inbound
Gym-V: A Unified Vision Environment System for Agentic Vision Research G1: Bootstrapping Perception and Reasoning Abilities of Vision-Language Model via Reinforcement Learning
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 58c72bb5-6fd8-4fc1-87b9-46224ea591f5 · inbound
Odysseus: Scaling VLMs to 100+ Turn Decision-Making in Games via Reinforcement Learning G1: Bootstrapping Perception and Reasoning Abilities of Vision-Language Model via Reinforcement Learning
Reference 74
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation c91ea36f-f7fd-4ea2-bbdc-c0320bde6f33 · inbound
Towards Generalist Game Players: An Investigation of Foundation Models in the Game Multiverse G1: Bootstrapping Perception and Reasoning Abilities of Vision-Language Model via Reinforcement Learning
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 9da3c042-ac19-4f9d-b2fa-fe0fe1f26042 · inbound
Towards Generalist Game Players: An Investigation of Foundation Models in the Game Multiverse G1: Bootstrapping Perception and Reasoning Abilities of Vision-Language Model via Reinforcement Learning
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 69838789-0ba6-4fc1-aaae-0b6f5097178e · inbound
RSICCLLM: A Multimodal Large Language Model for Remote Sensing Image Change Captioning G1: Bootstrapping Perception and Reasoning Abilities of Vision-Language Model via Reinforcement Learning
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 3f6d3dd1-ceba-448d-8587-55ad11df59ff · inbound
Before Thinking, Learn to Decide: Proactive Routing for Efficient Visual Reasoning G1: Bootstrapping Perception and Reasoning Abilities of Vision-Language Model via Reinforcement Learning
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 2e6a6b4c-2bcb-419c-8434-262108d00612 · inbound
CAST: Game Solvers as Turn-Level Teachers for LLM Agents G1: Bootstrapping Perception and Reasoning Abilities of Vision-Language Model via Reinforcement Learning
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.