Pith. sign in

Paper Citation Record · LEDGER

Reinforcing Video Reasoning with Focused Thinking

As of 7 August 2026, this Paper Citation Record lists 41 of 41 outbound references and 7 inbound Pith citation observations for arXiv:2505.24718.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.24718 v3

Coverage vector

measured 41 of 41 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T12:22:14.226407Z

measured 48 of 48 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 7 of 7 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T23:29:31.751173Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-02T17:27:15.553972Z

Reference resolution

41 of 41 outbound references displayed

  • verified exact0
  • verified fuzzy6
  • unresolved35
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation c9badecf-09ff-4624-b35a-5e04c047cb19 · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

Reinforcing Video Reasoning with Focused Thinking DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T12:22:10.960839Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:22:10.960839Z digest=sha256:1e728c0ef15dde0436e66c6381c321426d269b16700f2d819265c381d7a391f0

Observation d4ce1a9a-d3b4-49aa-9a09-05aed9e84c1f · outbound

This paper cites R1-v: Reinforcing super generalization ability in vision-language models with less than $3,.

Reinforcing Video Reasoning with Focused Thinking R1-v: Reinforcing super generalization ability in vision-language models with less than $3,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:22:15.828030Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:22:11.084671Z digest=sha256:40da6f8c5c3e340feab61998ab294ada467411bca8a8f00a4b2f311b2670682b

Observation a8d80d51-c2f5-47cb-a6a5-1ac8e721ad9f · outbound

This paper cites Vision-R1: Incentivizing Reasoning Capability in Multimodal Large Language Models.

Reinforcing Video Reasoning with Focused Thinking Vision-R1: Incentivizing Reasoning Capability in Multimodal Large Language Models

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T12:22:11.183874Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:22:11.183874Z digest=sha256:475c3cd63940c77225344a467b0d834137c967f682957f9b08b75b5011bea9b9

Observation def96e6b-1f53-4bcc-84cc-98fe0f272012 · outbound

This paper cites MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning.

Reinforcing Video Reasoning with Focused Thinking MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T12:22:11.241111Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:22:11.241111Z digest=sha256:1c174fbf436720eb0209486632d7bf12d0e408f3f73970516c7be1e2d30ba9cd

Observation 9cef48be-5148-48ee-8727-049a8ab0dbc7 · outbound

This paper cites R1-Zero's "Aha Moment" in Visual Reasoning on a 2B Non-SFT Model.

Reinforcing Video Reasoning with Focused Thinking R1-Zero's "Aha Moment" in Visual Reasoning on a 2B Non-SFT Model

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T12:22:11.311509Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:22:11.311509Z digest=sha256:adf3366b0a2244e2f9404f775679ce7be92d5011238beb24e0d3d67fd58ca820

Observation 77cdd7ca-ef91-47f0-b384-caf56fe79085 · outbound

This paper cites LMM-R1: Empowering 3B LMMs with Strong Reasoning Abilities Through Two-Stage Rule-Based RL.

Reinforcing Video Reasoning with Focused Thinking LMM-R1: Empowering 3B LMMs with Strong Reasoning Abilities Through Two-Stage Rule-Based RL

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T12:22:11.359058Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:22:11.359058Z digest=sha256:091f1a191c061205149a65b365fe443c769c16ed49c606e137f6a27497f03c8e

Observation 11a643dc-ec93-43fe-ac47-f4c49d1f0933 · outbound

This paper cites Video-R1: Reinforcing Video Reasoning in MLLMs.

Reinforcing Video Reasoning with Focused Thinking Video-R1: Reinforcing Video Reasoning in MLLMs

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T12:22:11.399735Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:22:11.399735Z digest=sha256:d417ca4ec51e166952a227790587e172abafc86769ff5904152f997c5322f758

Observation 7c4361f0-b713-4b46-a04a-8ee3ac1a028f · outbound

This paper cites VideoChat-R1: Enhancing Spatio-Temporal Perception via Reinforcement Fine-Tuning.

Reinforcing Video Reasoning with Focused Thinking VideoChat-R1: Enhancing Spatio-Temporal Perception via Reinforcement Fine-Tuning

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T12:22:11.468450Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:22:11.468450Z digest=sha256:f89cf12762c23b152c9554a235496a214bb2ee374ef69582c154ef1df41da457

Observation 9a269062-f33e-4c8c-8661-ef6baf68d974 · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

Reinforcing Video Reasoning with Focused Thinking DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T12:22:11.489252Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:22:11.489252Z digest=sha256:ca0fb55dd8c68c14e75da0b660810c20aee80d04e699495ba8ccb493d076f4f4

Observation 9f385f3f-9ae7-46d2-8098-d70880665cf6 · outbound

This paper cites MME-CoT: Benchmarking Chain-of-Thought in Large Multimodal Models for Reasoning Quality, Robustness, and Efficiency.

Reinforcing Video Reasoning with Focused Thinking MME-CoT: Benchmarking Chain-of-Thought in Large Multimodal Models for Reasoning Quality, Robustness, and Efficiency

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T12:22:11.542607Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:22:11.542607Z digest=sha256:5318de26177ca6186e4c932ca60437a8edca738cb7d894497789970e807fb53f

Observation 3c644961-f39d-4b35-85d2-385fb0255c0c · outbound

This paper cites R1-VL: Learning to Reason with Multimodal Large Language Models via Step-wise Group Relative Policy Optimization.

Reinforcing Video Reasoning with Focused Thinking R1-VL: Learning to Reason with Multimodal Large Language Models via Step-wise Group Relative Policy Optimization

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T12:22:11.606319Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:22:11.606319Z digest=sha256:152a449ab47f56a7d090241511acd5c34603474a1fbda4229a18ef6624f83908

Observation 39800715-77d3-4a6b-8dcb-9616d86a5364 · outbound

This paper cites CLEVRER: CoLlision Events for Video REpresentation and Reasoning.

Reinforcing Video Reasoning with Focused Thinking CLEVRER: CoLlision Events for Video REpresentation and Reasoning

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T12:22:11.671883Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:22:11.671883Z digest=sha256:9a11fd1e0ad28909f3f3d918fadf682cc7f68ad4678bdad869304b1790cbdffb

Observation 63dc5c0e-fd1e-4515-8820-7df1a7617a02 · outbound

This paper cites Can i trust your answer? visually grounded video ques- tion answering,.

Reinforcing Video Reasoning with Focused Thinking Can i trust your answer? visually grounded video ques- tion answering,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:22:15.577845Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:22:11.761989Z digest=sha256:3c585a6c6fabd87a4621bf917b94394abfdf88ffff9826afdc2382f0d6fd0ce7

Observation b10404a4-85df-4172-9be5-4bc22869d712 · outbound

This paper cites MMVU: Measuring Expert-Level Multi-Discipline Video Understanding.

Reinforcing Video Reasoning with Focused Thinking MMVU: Measuring Expert-Level Multi-Discipline Video Understanding

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T12:22:11.825894Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:22:11.825894Z digest=sha256:c2843ded439aafaffa1962a70cdb538b82f5c7ac87d2433893648677fb9248cf

Observation 76b2b73f-acf5-4bf4-b0e6-2208a7941102 · outbound

This paper cites OpenAI o1 System Card.

Reinforcing Video Reasoning with Focused Thinking OpenAI o1 System Card

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T12:22:11.951541Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:22:11.951541Z digest=sha256:e7a6ac969af04eb98f8e01dab4f323e8fa47c98d0ef5605f5d172aa9dce26720

Observation 8070ccbe-b763-47b6-b467-3b9d0d7c8850 · outbound

This paper cites Secrets of RLHF in Large Language Models Part I: PPO.

Reinforcing Video Reasoning with Focused Thinking Secrets of RLHF in Large Language Models Part I: PPO

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T12:22:12.045876Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:22:12.045876Z digest=sha256:de61dee1ec2f9121f12bb883c31b0bfce0019705d41709beaca288a3ad8fa4a2

Observation 0613e258-0573-4692-ae18-f61740a8bef4 · outbound

This paper cites Vision-R1: Evolving Human-Free Alignment in Large Vision-Language Models via Vision-Guided Reinforcement Learning.

Reinforcing Video Reasoning with Focused Thinking Vision-R1: Evolving Human-Free Alignment in Large Vision-Language Models via Vision-Guided Reinforcement Learning

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T12:22:12.120909Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:22:12.120909Z digest=sha256:2c52c072c93e293be7dd2f1a971a1b89ff62c245531707bbb1b753f6175a1825

Observation 7b4aba80-d73a-4225-914e-cd2b8ce53e20 · outbound

This paper cites SEED-GRPO: Semantic Entropy Enhanced GRPO for Uncertainty-Aware Policy Optimization.

Reinforcing Video Reasoning with Focused Thinking SEED-GRPO: Semantic Entropy Enhanced GRPO for Uncertainty-Aware Policy Optimization

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T12:22:12.190500Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:22:12.190500Z digest=sha256:1bbea8c8a1141c41ee0711d8799cb4a38d5817e96fca18981ed734b0ab029632

Observation 8a5119cd-5d9d-4291-8035-90a5405ecd91 · outbound

This paper cites VideoLLaMA 2: Advancing Spatial-Temporal Modeling and Audio Understanding in Video-LLMs.

Reinforcing Video Reasoning with Focused Thinking VideoLLaMA 2: Advancing Spatial-Temporal Modeling and Audio Understanding in Video-LLMs

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T12:22:12.277375Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:22:12.277375Z digest=sha256:b22a360dd54d351dd59db3da59182242e62396602ad56ab3f903a7c8e275267f

Observation 868cd186-217e-47e2-9345-540ad555343e · outbound

This paper cites Audio-Visual LLM for Video Understanding.

Reinforcing Video Reasoning with Focused Thinking Audio-Visual LLM for Video Understanding

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T12:22:12.360528Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:22:12.360528Z digest=sha256:5fe77e3da7a2d6f9d42af79080cd195ba45b6cf8d1a15afe1c04c440ba7cf454

Observation 1be92ba5-8340-4a3b-9337-dce1fb7f4931 · outbound

This paper cites Visa: Reasoning video object segmentation via large language models,.

Reinforcing Video Reasoning with Focused Thinking Visa: Reasoning video object segmentation via large language models,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:22:15.384290Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:22:12.453345Z digest=sha256:b35a718f814879b293595126d98349e37137f79bd4124e71b442c7456a662e2b

Observation 87bdd1af-243d-4b5e-b13e-7a70c8daa3c9 · outbound

This paper cites Will Pre-Training Ever End? A First Step Toward Next-Generation Foundation MLLMs via Self-Improving Systematic Cognition.

Reinforcing Video Reasoning with Focused Thinking Will Pre-Training Ever End? A First Step Toward Next-Generation Foundation MLLMs via Self-Improving Systematic Cognition

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T12:22:12.546086Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:22:12.546086Z digest=sha256:13cc0469ec651a1e0060462281ce2f3c0161eb2eb824e506c539f72a8d70c77b

Observation 7bd86063-dc8f-43ee-b910-3d8236c3ae10 · outbound

This paper cites LLaVA-UHD v2: an MLLM Integrating High-Resolution Semantic Pyramid via Hierarchical Window Transformer.

Reinforcing Video Reasoning with Focused Thinking LLaVA-UHD v2: an MLLM Integrating High-Resolution Semantic Pyramid via Hierarchical Window Transformer

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T12:22:12.643287Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:22:12.643287Z digest=sha256:0cb66a0b5e71452e0953512fcf87e2e508bc6f15fb548daac67639f5ecfa6a4e

Observation a68fac34-d6bf-45f2-9f76-6ba773877cc4 · outbound

This paper cites Forking Paths in Neural Text Generation.

Reinforcing Video Reasoning with Focused Thinking Forking Paths in Neural Text Generation

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T12:22:12.716029Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:22:12.716029Z digest=sha256:2f52cab42943a912a83aef4b8633419b278420624ec942e9f3663c758b05edf3

Observation 570faee5-66db-4c4e-812c-0e3a9d8f6b1c · outbound

This paper cites Critical Tokens Matter: Token-Level Contrastive Estimation Enhances LLM's Reasoning Capability.

Reinforcing Video Reasoning with Focused Thinking Critical Tokens Matter: Token-Level Contrastive Estimation Enhances LLM's Reasoning Capability

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T12:22:12.794644Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:22:12.794644Z digest=sha256:773d091d3c45908683db79420a5d595e1d2991da729577630aa596fd28c3dc05

Observation 605d6a0f-b46e-45a6-b1fd-25588ab8c0ca · outbound

This paper cites DAPO: An Open-Source LLM Reinforcement Learning System at Scale.

Reinforcing Video Reasoning with Focused Thinking DAPO: An Open-Source LLM Reinforcement Learning System at Scale

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T12:22:12.867048Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:22:12.867048Z digest=sha256:6ff37b2eac1f010b3ac4cba6d3261df5128994f8f853f02c5e831a762feef5ec

Observation 0dc28f2d-2237-4f73-bb25-60c69b46ac44 · outbound

This paper cites Mvbench: A comprehensive multi-modal video understanding benchmark,.

Reinforcing Video Reasoning with Focused Thinking Mvbench: A comprehensive multi-modal video understanding benchmark,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:22:15.240891Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:22:12.956300Z digest=sha256:d94abf274f861a64b67596b734370cd3fc1fa1db9561333f284c55500d37fa17

Observation cf3020a1-1f48-4b8c-81aa-788731fa7a08 · outbound

This paper cites TempCompass: Do Video LLMs Really Understand Videos?.

Reinforcing Video Reasoning with Focused Thinking TempCompass: Do Video LLMs Really Understand Videos?

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T12:22:13.038077Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:22:13.038077Z digest=sha256:0de22bbf52cd70605ea398157a8245f9b04fe4e8c18b772989bb2e1897164981

Observation bed7a548-6ece-4d28-8108-5f014352f3ba · outbound

This paper cites Video-MME: The First-Ever Comprehensive Evaluation Benchmark of Multi-modal LLMs in Video Analysis.

Reinforcing Video Reasoning with Focused Thinking Video-MME: The First-Ever Comprehensive Evaluation Benchmark of Multi-modal LLMs in Video Analysis

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T12:22:13.162144Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:22:13.162144Z digest=sha256:616330b071e7451f6f871476dd6d1ad6274015e07569099205b3ab34b0ad1b22

Observation ff5b7fee-4772-452d-8a82-d0de440785c9 · outbound

This paper cites Llama-vid: An image is worth 2 tokens in large language models,.

Reinforcing Video Reasoning with Focused Thinking Llama-vid: An image is worth 2 tokens in large language models,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:22:15.104174Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:22:13.236797Z digest=sha256:91c7673cfaf4fb81565c2939366da36071cd51798d8ca052b6aad4b12c48f2d0

Observation 708677d5-b6a1-4e7b-88c6-5af249e4da12 · outbound

This paper cites Long Context Transfer from Language to Vision.

Reinforcing Video Reasoning with Focused Thinking Long Context Transfer from Language to Vision

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T12:22:13.330350Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:22:13.330350Z digest=sha256:42605fe5a888d1ac2f2fca9836eb436a0e0387a1af267f9e54cc7d462646e74c

Observation 81d94e55-924f-4bef-8a01-09d71e5a273c · outbound

This paper cites Unhackable Temporal Rewarding for Scalable Video MLLMs.

Reinforcing Video Reasoning with Focused Thinking Unhackable Temporal Rewarding for Scalable Video MLLMs

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T12:22:13.428738Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:22:13.428738Z digest=sha256:becd52007ba7c5053fa7c1d9fbb64272c56ba56243b262b6a628b367d6991983

Observation 6443ad05-adcd-41d2-b2e6-1dc1fa460b2e · outbound

This paper cites Kangaroo: A Powerful Video-Language Model Supporting Long-context Video Input.

Reinforcing Video Reasoning with Focused Thinking Kangaroo: A Powerful Video-Language Model Supporting Long-context Video Input

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T12:22:13.508706Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:22:13.508706Z digest=sha256:0fd8169d13170b9b11659c987f5d1f6768daabf382677cce35de21a99dd45af8

Observation ea08b928-4410-4cbe-9613-cb0320456b33 · outbound

This paper cites Qwen2.5-VL Technical Report.

Reinforcing Video Reasoning with Focused Thinking Qwen2.5-VL Technical Report

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T12:22:13.630605Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:22:13.630605Z digest=sha256:2e3b5287df9dcb62bd89fce26f3aeb3cd0be9a241f8f8957a075cdc9deb92015

Observation b23e32ea-f4c2-45bd-a21a-5b55d020807b · outbound

This paper cites STAR: A Benchmark for Situated Reasoning in Real-World Videos.

Reinforcing Video Reasoning with Focused Thinking STAR: A Benchmark for Situated Reasoning in Real-World Videos

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T12:22:13.702673Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:22:13.702673Z digest=sha256:c3f971b98e4a61c61e366db76102a83d61f328801017b697f8a710a9e8fc96c3

Observation 04fb0489-b2e9-4f57-b528-5d163eb5e669 · outbound

This paper cites The Unreasonable Effectiveness of Entropy Minimization in LLM Reasoning.

Reinforcing Video Reasoning with Focused Thinking The Unreasonable Effectiveness of Entropy Minimization in LLM Reasoning

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T12:22:13.785924Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:22:13.785924Z digest=sha256:b703b1f282f50db8e3cfc73d24e57da1928ba082b45444d2089fad7c5e471fb4

Observation faa37307-72ef-439b-9160-a7171164c542 · outbound

This paper cites Right Question is Already Half the Answer: Fully Unsupervised LLM Reasoning Incentivization.

Reinforcing Video Reasoning with Focused Thinking Right Question is Already Half the Answer: Fully Unsupervised LLM Reasoning Incentivization

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T12:22:13.915142Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:22:13.915142Z digest=sha256:6425d0543b9dc5d1cff3723eb3b42c6725e27f2857b718646c35429630e42d41

Observation 059a37cd-0a8a-4db3-8832-8abd7abb548e · outbound

This paper cites Learning to Reason without External Rewards.

Reinforcing Video Reasoning with Focused Thinking Learning to Reason without External Rewards

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T12:22:14.022769Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:22:14.022769Z digest=sha256:212c3a07dfcb0e68dd985c4c8f2a2332d21e6c7e131af74aeaeea7cd16c481df

Observation 9a939a17-bd5f-4ae7-8863-1005e0f9acf6 · outbound

This paper cites Scalable best-of-n selection for large language models via self-certainty,.

Reinforcing Video Reasoning with Focused Thinking Scalable best-of-n selection for large language models via self-certainty,

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T12:22:14.098816Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:22:14.098816Z digest=sha256:153e8a1935533e2e2c818b5951faa083a5d1d4c04186e6404c8af2e05a603241

Observation cba31ede-d6d4-4a2b-884f-d19b251f9087 · outbound

This paper cites Proximal Policy Optimization Algorithms.

Reinforcing Video Reasoning with Focused Thinking Proximal Policy Optimization Algorithms

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T12:22:14.151223Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:22:14.151223Z digest=sha256:f21437c3ffc35cd8d6ef109978fd667c42d59ec13e76afec798978bca97b0604

Observation 9d76de55-bffc-4eba-982d-7a74f3e6c844 · outbound

This paper cites Trl: Transformer reinforcement learning,.

Reinforcing Video Reasoning with Focused Thinking Trl: Transformer reinforcement learning,

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:22:14.955173Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:22:14.226407Z digest=sha256:117a0418a367939da9543a5f6ccb15e8eef2ff5c712810870acb54434e3d4a04

Pith citing papers

Observation 11ef9459-0e33-44f8-b644-437f7bf2568f · inbound

Towards Reasoning Era: A Survey of Long Chain-of-Thought for Reasoning Large Language Models cites this paper.

Towards Reasoning Era: A Survey of Long Chain-of-Thought for Reasoning Large Language Models Reinforcing Video Reasoning with Focused Thinking

Reference 152

Resolution
verified exact
arxiv_id, observed 2026-05-12T08:40:42.109496Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-12T08:40:40.910461Z digest=sha256:9bc22cae2981bc74bc5a4e11a8d346ecd9783efc2e833cd9fd159e0836a22986

Observation d580f246-b69a-4ce7-9e75-ff8a42927870 · inbound

MUSEG: Reinforcing Video Temporal Understanding via Timestamp-Aware Multi-Segment Grounding cites this paper.

MUSEG: Reinforcing Video Temporal Understanding via Timestamp-Aware Multi-Segment Grounding Reinforcing Video Reasoning with Focused Thinking

Reference 31

Resolution
verified exact
arxiv_id, observed 2026-05-19T13:17:18.498226Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-19T13:13:40.485342Z digest=sha256:1e39cf5943f09fb92e7daa0aec2d06acf67c8af360ff236da33913a3d988e37f

Observation f5846c80-5fad-47c0-a4f5-ee604d4987cf · inbound

MUPA: Towards Multi-Path Agentic Reasoning for Grounded Video Question Answering cites this paper.

MUPA: Towards Multi-Path Agentic Reasoning for Grounded Video Question Answering Reinforcing Video Reasoning with Focused Thinking

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-06T23:29:31.751173Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:29:31.751173Z digest=sha256:91ecf9113fe32499a76feb35aacb03db584f4bd4fbdf5d37af00a380d3031aae

Observation 64b23fc0-7c1f-43fa-b7cd-ae758657a52e · inbound

A Survey of Reinforcement Learning for Large Reasoning Models cites this paper.

A Survey of Reinforcement Learning for Large Reasoning Models Reinforcing Video Reasoning with Focused Thinking

Reference 101

Resolution
verified exact
arxiv_id, observed 2026-05-18T00:02:25.240758Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-18T00:02:24.352947Z digest=sha256:cdcd3046267718f86d043afece37a0eada1fdf775286d81548068401f6ba5064

Observation fa44740c-da3e-4761-9ada-d3c4e94c2d7c · inbound

Watch Before You Answer: Learning from Visually Grounded Post-Training cites this paper.

Watch Before You Answer: Learning from Visually Grounded Post-Training Reinforcing Video Reasoning with Focused Thinking

Reference 17

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T22:20:46.742418Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-10T20:01:13.305374Z digest=sha256:0c9b66e4664bf85e58b61ed701bcda2772f8fe50b008f42cbb532b26380371b6

Observation 47928768-5cd9-4810-bb8b-6b6ceda6a724 · inbound

Beyond Perceptual Shortcuts: Causal-Inspired Debiasing Optimization for Generalizable Video Reasoning in Lightweight MLLMs cites this paper.

Beyond Perceptual Shortcuts: Causal-Inspired Debiasing Optimization for Generalizable Video Reasoning in Lightweight MLLMs Reinforcing Video Reasoning with Focused Thinking

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-11T16:56:06.417753Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-09T14:30:56.297653Z digest=sha256:322f2734bdbfb2e0481ea799f3bdc11952444c43efcb83c514b6de2275580043

Observation 836b67cd-9967-44af-b27b-fc1230f82719 · inbound

Watch, Remember, Reason: Human-View Video Understanding with MLLMs cites this paper.

Watch, Remember, Reason: Human-View Video Understanding with MLLMs Reinforcing Video Reasoning with Focused Thinking

Reference 184

Resolution
verified exact
arxiv_id, observed 2026-07-02T17:27:15.555931Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-27T22:00:28.350003Z digest=sha256:e1680e2a73c7eb5c101388d719b392cc0cb99b20cc733cd165a074d41c3ae561