Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T12:22:14.226407Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 41 of 41 outbound references and 7 inbound Pith citation observations for arXiv:2505.24718.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T12:22:14.226407Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-06T23:29:31.751173Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-02T17:27:15.553972Z
41 of 41 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation c9badecf-09ff-4624-b35a-5e04c047cb19 · outbound
Reinforcing Video Reasoning with Focused Thinking DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d4ce1a9a-d3b4-49aa-9a09-05aed9e84c1f · outbound
Reinforcing Video Reasoning with Focused Thinking R1-v: Reinforcing super generalization ability in vision-language models with less than $3,
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a8d80d51-c2f5-47cb-a6a5-1ac8e721ad9f · outbound
Reinforcing Video Reasoning with Focused Thinking Vision-R1: Incentivizing Reasoning Capability in Multimodal Large Language Models
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation def96e6b-1f53-4bcc-84cc-98fe0f272012 · outbound
Reinforcing Video Reasoning with Focused Thinking MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9cef48be-5148-48ee-8727-049a8ab0dbc7 · outbound
Reinforcing Video Reasoning with Focused Thinking R1-Zero's "Aha Moment" in Visual Reasoning on a 2B Non-SFT Model
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 77cdd7ca-ef91-47f0-b384-caf56fe79085 · outbound
Reinforcing Video Reasoning with Focused Thinking LMM-R1: Empowering 3B LMMs with Strong Reasoning Abilities Through Two-Stage Rule-Based RL
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 11a643dc-ec93-43fe-ac47-f4c49d1f0933 · outbound
Reinforcing Video Reasoning with Focused Thinking Video-R1: Reinforcing Video Reasoning in MLLMs
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7c4361f0-b713-4b46-a04a-8ee3ac1a028f · outbound
Reinforcing Video Reasoning with Focused Thinking VideoChat-R1: Enhancing Spatio-Temporal Perception via Reinforcement Fine-Tuning
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9a269062-f33e-4c8c-8661-ef6baf68d974 · outbound
Reinforcing Video Reasoning with Focused Thinking DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9f385f3f-9ae7-46d2-8098-d70880665cf6 · outbound
Reinforcing Video Reasoning with Focused Thinking MME-CoT: Benchmarking Chain-of-Thought in Large Multimodal Models for Reasoning Quality, Robustness, and Efficiency
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3c644961-f39d-4b35-85d2-385fb0255c0c · outbound
Reinforcing Video Reasoning with Focused Thinking R1-VL: Learning to Reason with Multimodal Large Language Models via Step-wise Group Relative Policy Optimization
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 39800715-77d3-4a6b-8dcb-9616d86a5364 · outbound
Reinforcing Video Reasoning with Focused Thinking CLEVRER: CoLlision Events for Video REpresentation and Reasoning
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 63dc5c0e-fd1e-4515-8820-7df1a7617a02 · outbound
Reinforcing Video Reasoning with Focused Thinking Can i trust your answer? visually grounded video ques- tion answering,
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b10404a4-85df-4172-9be5-4bc22869d712 · outbound
Reinforcing Video Reasoning with Focused Thinking MMVU: Measuring Expert-Level Multi-Discipline Video Understanding
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 76b2b73f-acf5-4bf4-b0e6-2208a7941102 · outbound
Reinforcing Video Reasoning with Focused Thinking OpenAI o1 System Card
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8070ccbe-b763-47b6-b467-3b9d0d7c8850 · outbound
Reinforcing Video Reasoning with Focused Thinking Secrets of RLHF in Large Language Models Part I: PPO
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0613e258-0573-4692-ae18-f61740a8bef4 · outbound
Reinforcing Video Reasoning with Focused Thinking Vision-R1: Evolving Human-Free Alignment in Large Vision-Language Models via Vision-Guided Reinforcement Learning
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7b4aba80-d73a-4225-914e-cd2b8ce53e20 · outbound
Reinforcing Video Reasoning with Focused Thinking SEED-GRPO: Semantic Entropy Enhanced GRPO for Uncertainty-Aware Policy Optimization
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8a5119cd-5d9d-4291-8035-90a5405ecd91 · outbound
Reinforcing Video Reasoning with Focused Thinking VideoLLaMA 2: Advancing Spatial-Temporal Modeling and Audio Understanding in Video-LLMs
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 868cd186-217e-47e2-9345-540ad555343e · outbound
Reinforcing Video Reasoning with Focused Thinking Audio-Visual LLM for Video Understanding
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1be92ba5-8340-4a3b-9337-dce1fb7f4931 · outbound
Reinforcing Video Reasoning with Focused Thinking Visa: Reasoning video object segmentation via large language models,
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 87bdd1af-243d-4b5e-b13e-7a70c8daa3c9 · outbound
Reinforcing Video Reasoning with Focused Thinking Will Pre-Training Ever End? A First Step Toward Next-Generation Foundation MLLMs via Self-Improving Systematic Cognition
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7bd86063-dc8f-43ee-b910-3d8236c3ae10 · outbound
Reinforcing Video Reasoning with Focused Thinking LLaVA-UHD v2: an MLLM Integrating High-Resolution Semantic Pyramid via Hierarchical Window Transformer
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a68fac34-d6bf-45f2-9f76-6ba773877cc4 · outbound
Reinforcing Video Reasoning with Focused Thinking Forking Paths in Neural Text Generation
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 570faee5-66db-4c4e-812c-0e3a9d8f6b1c · outbound
Reinforcing Video Reasoning with Focused Thinking Critical Tokens Matter: Token-Level Contrastive Estimation Enhances LLM's Reasoning Capability
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 605d6a0f-b46e-45a6-b1fd-25588ab8c0ca · outbound
Reinforcing Video Reasoning with Focused Thinking DAPO: An Open-Source LLM Reinforcement Learning System at Scale
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0dc28f2d-2237-4f73-bb25-60c69b46ac44 · outbound
Reinforcing Video Reasoning with Focused Thinking Mvbench: A comprehensive multi-modal video understanding benchmark,
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation cf3020a1-1f48-4b8c-81aa-788731fa7a08 · outbound
Reinforcing Video Reasoning with Focused Thinking TempCompass: Do Video LLMs Really Understand Videos?
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bed7a548-6ece-4d28-8108-5f014352f3ba · outbound
Reinforcing Video Reasoning with Focused Thinking Video-MME: The First-Ever Comprehensive Evaluation Benchmark of Multi-modal LLMs in Video Analysis
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ff5b7fee-4772-452d-8a82-d0de440785c9 · outbound
Reinforcing Video Reasoning with Focused Thinking Llama-vid: An image is worth 2 tokens in large language models,
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 708677d5-b6a1-4e7b-88c6-5af249e4da12 · outbound
Reinforcing Video Reasoning with Focused Thinking Long Context Transfer from Language to Vision
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 81d94e55-924f-4bef-8a01-09d71e5a273c · outbound
Reinforcing Video Reasoning with Focused Thinking Unhackable Temporal Rewarding for Scalable Video MLLMs
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6443ad05-adcd-41d2-b2e6-1dc1fa460b2e · outbound
Reinforcing Video Reasoning with Focused Thinking Kangaroo: A Powerful Video-Language Model Supporting Long-context Video Input
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ea08b928-4410-4cbe-9613-cb0320456b33 · outbound
Reinforcing Video Reasoning with Focused Thinking Qwen2.5-VL Technical Report
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b23e32ea-f4c2-45bd-a21a-5b55d020807b · outbound
Reinforcing Video Reasoning with Focused Thinking STAR: A Benchmark for Situated Reasoning in Real-World Videos
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 04fb0489-b2e9-4f57-b528-5d163eb5e669 · outbound
Reinforcing Video Reasoning with Focused Thinking The Unreasonable Effectiveness of Entropy Minimization in LLM Reasoning
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation faa37307-72ef-439b-9160-a7171164c542 · outbound
Reinforcing Video Reasoning with Focused Thinking Right Question is Already Half the Answer: Fully Unsupervised LLM Reasoning Incentivization
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 059a37cd-0a8a-4db3-8832-8abd7abb548e · outbound
Reinforcing Video Reasoning with Focused Thinking Learning to Reason without External Rewards
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9a939a17-bd5f-4ae7-8863-1005e0f9acf6 · outbound
Reinforcing Video Reasoning with Focused Thinking Scalable best-of-n selection for large language models via self-certainty,
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cba31ede-d6d4-4a2b-884f-d19b251f9087 · outbound
Reinforcing Video Reasoning with Focused Thinking Proximal Policy Optimization Algorithms
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9d76de55-bffc-4eba-982d-7a74f3e6c844 · outbound
Reinforcing Video Reasoning with Focused Thinking Trl: Transformer reinforcement learning,
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 11ef9459-0e33-44f8-b644-437f7bf2568f · inbound
Towards Reasoning Era: A Survey of Long Chain-of-Thought for Reasoning Large Language Models Reinforcing Video Reasoning with Focused Thinking
Reference 152
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d580f246-b69a-4ce7-9e75-ff8a42927870 · inbound
MUSEG: Reinforcing Video Temporal Understanding via Timestamp-Aware Multi-Segment Grounding Reinforcing Video Reasoning with Focused Thinking
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f5846c80-5fad-47c0-a4f5-ee604d4987cf · inbound
MUPA: Towards Multi-Path Agentic Reasoning for Grounded Video Question Answering Reinforcing Video Reasoning with Focused Thinking
Reference 65
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 64b23fc0-7c1f-43fa-b7cd-ae758657a52e · inbound
A Survey of Reinforcement Learning for Large Reasoning Models Reinforcing Video Reasoning with Focused Thinking
Reference 101
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation fa44740c-da3e-4761-9ada-d3c4e94c2d7c · inbound
Watch Before You Answer: Learning from Visually Grounded Post-Training Reinforcing Video Reasoning with Focused Thinking
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 47928768-5cd9-4810-bb8b-6b6ceda6a724 · inbound
Beyond Perceptual Shortcuts: Causal-Inspired Debiasing Optimization for Generalizable Video Reasoning in Lightweight MLLMs Reinforcing Video Reasoning with Focused Thinking
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 836b67cd-9967-44af-b27b-fc1230f82719 · inbound
Watch, Remember, Reason: Human-View Video Understanding with MLLMs Reinforcing Video Reasoning with Focused Thinking
Reference 184
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.