Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 29 inbound Pith citation observations for arXiv:2501.12380.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T12:45:50.222550Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-03T20:48:55.978205Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation d530259e-2af9-40f2-9f55-ce0aeab71c2a · inbound
Position: Multimodal Large Language Models Can Significantly Advance Scientific Reasoning MMVU: Measuring Expert-Level Multi-Discipline Video Understanding
Reference 258
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 90f534eb-cd28-45cb-90de-ceb85eb367c2 · inbound
Qwen2.5-VL Technical Report MMVU: Measuring Expert-Level Multi-Discipline Video Understanding
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 692271aa-60f7-4281-a106-cb3b4aaa99a2 · inbound
Video-R1: Reinforcing Video Reasoning in MLLMs MMVU: Measuring Expert-Level Multi-Discipline Video Understanding
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 1d76e4df-1500-4305-bb78-4964e978c0f8 · inbound
Seed1.5-VL Technical Report MMVU: Measuring Expert-Level Multi-Discipline Video Understanding
Reference 176
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 5b120fd5-be85-403a-8b63-6e8a245f676d · inbound
Video-Holmes: Can MLLM Think Like Holmes for Complex Video Reasoning? MMVU: Measuring Expert-Level Multi-Discipline Video Understanding
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 59f62b73-7618-40f5-89b7-ea8e69ca370f · inbound
VF-Eval: Evaluating Multimodal LLMs for Generating Feedback on AIGC Videos MMVU: Measuring Expert-Level Multi-Discipline Video Understanding
Reference 61
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b10404a4-85df-4172-9be5-4bc22869d712 · inbound
Reinforcing Video Reasoning with Focused Thinking MMVU: Measuring Expert-Level Multi-Discipline Video Understanding
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3ec4580c-f95e-4276-833e-804df956f725 · inbound
ReAgent-V: A Reward-Driven Multi-Agent Framework for Video Understanding MMVU: Measuring Expert-Level Multi-Discipline Video Understanding
Reference 58
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 568a8494-8f7a-48f4-b86c-3834ef178039 · inbound
VideoCap-R1: Enhancing MLLMs for Video Captioning via Structured Thinking MMVU: Measuring Expert-Level Multi-Discipline Video Understanding
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 81a57794-523a-4f6d-aa33-8a70e34563d8 · inbound
Reinforcement Learning Tuning for VideoLLMs: Reward Design and Data Efficiency MMVU: Measuring Expert-Level Multi-Discipline Video Understanding
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f5963aa4-80c5-4e97-9d84-92b464ed8acd · inbound
EOC-Bench: Can MLLMs Identify, Recall, and Forecast Objects in an Egocentric World? MMVU: Measuring Expert-Level Multi-Discipline Video Understanding
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 18779145-4b15-48c6-9e9a-80262decbb4c · inbound
Wait, We Don't Need to "Wait"! Removing Thinking Tokens Improves Reasoning Efficiency MMVU: Measuring Expert-Level Multi-Discipline Video Understanding
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 737f9633-f141-463b-b19e-f15e39a533b1 · inbound
VRBench: A Benchmark for Multi-Step Reasoning in Long Narrative Videos MMVU: Measuring Expert-Level Multi-Discipline Video Understanding
Reference 92
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cfb50d06-db4a-4122-b1a9-f8f7b1410dc9 · inbound
GRPO-CARE: Consistency-Aware Reinforcement Learning for Multimodal Reasoning MMVU: Measuring Expert-Level Multi-Discipline Video Understanding
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 41b50db4-1a75-499d-8957-a94752de91bb · inbound
HumanOmniV2: From Understanding to Omni-Modal Reasoning with Context MMVU: Measuring Expert-Level Multi-Discipline Video Understanding
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 70847a66-37ea-4fe9-bdbd-0dbeee5c7c97 · inbound
ExpStar: Towards Automatic Commentary Generation for Multi-discipline Scientific Experiments MMVU: Measuring Expert-Level Multi-Discipline Video Understanding
Reference 63
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 391f0e94-7cc2-4f5a-802d-bf20799a183f · inbound
"Harmless to You, Hurtful to Me!": Investigating the Detection of Toxic Languages Grounded in the Perspective of Youth MMVU: Measuring Expert-Level Multi-Discipline Video Understanding
Reference 98
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bf9ba7c1-8894-43ca-9c45-32e62ae6b963 · inbound
VLM4D: Towards Spatiotemporal Awareness in Vision Language Models MMVU: Measuring Expert-Level Multi-Discipline Video Understanding
Reference 99
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7b5b0a94-6e41-4ee8-b6ee-f5e1eea5ba8b · inbound
HumanPCR: Probing MLLM Capabilities in Diverse Human-Centric Scenes MMVU: Measuring Expert-Level Multi-Discipline Video Understanding
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 348ce4e7-9091-4aea-b1fb-7ee6108a8402 · inbound
Video Reasoning without Training MMVU: Measuring Expert-Level Multi-Discipline Video Understanding
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 36ffb9e0-b6ca-4cde-814c-736e7b3be412 · inbound
Kimi K2.5: Visual Agentic Intelligence MMVU: Measuring Expert-Level Multi-Discipline Video Understanding
Reference 82
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 5f8d3008-4f27-4050-8dfb-ec67cbb3e8e5 · inbound
EasyVideoR1: Easier RL for Video Understanding MMVU: Measuring Expert-Level Multi-Discipline Video Understanding
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 786b7083-5a8f-484c-b750-94f3e1903855 · inbound
Video-ToC: Video Tree-of-Cue Reasoning MMVU: Measuring Expert-Level Multi-Discipline Video Understanding
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation d640528f-36f7-4894-b4f4-c61acfd8f2a0 · inbound
Co-Evolving Policy Distillation MMVU: Measuring Expert-Level Multi-Discipline Video Understanding
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation d319652e-f40d-41c5-a331-7082a0a4f2d7 · inbound
EgoCoT-Bench: Benchmarking Grounded and Verifiable Operation-Centric Chain of Thought Reasoning for MLLMs MMVU: Measuring Expert-Level Multi-Discipline Video Understanding
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 21dbcd8c-ac52-4bdf-aa4a-6a54472cf9e9 · inbound
VTI-CoT: Visual-Textual Interleaved Chain of Thought for Video Reasoning MMVU: Measuring Expert-Level Multi-Discipline Video Understanding
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 7ef3b79a-710a-4320-8528-13a890075f05 · inbound
Zone of Proximal Policy Optimization: Teacher in Prompts, Not Gradients MMVU: Measuring Expert-Level Multi-Discipline Video Understanding
Reference 151
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 7f85be1a-23a9-4c37-ab92-71f14c27d7e5 · inbound
DELTAVID: Enhancing Fine-Grained Spatiotemporal Perception with Cross-Video Differences MMVU: Measuring Expert-Level Multi-Discipline Video Understanding
Reference 64
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d4c2cb37-99d7-49b1-8cd2-1efa60cbe954 · inbound
Latent Visual Cache for Video Reasoning MMVU: Measuring Expert-Level Multi-Discipline Video Understanding
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.