Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 38 inbound Pith citation observations for arXiv:2405.09711.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T14:33:32.602468Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z
0 of 0 outbound references displayed
External citation measurements
22
arxiv_reference, observed 2026-08-05T02:28:24.338817Z
No outbound reference observations are available for this paper version.
Observation 2c6edefc-924a-4814-96bb-d0ab6ccd99b4 · inbound
LLaVA-NeXT-Interleave: Tackling Multi-image, Video, and 3D in Large Multimodal Models STAR: A Benchmark for Situated Reasoning in Real-World Videos
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation de30d0f5-b527-42d3-bfed-85727a24d0ad · inbound
LLaVA-Video: Video Instruction Tuning With Synthetic Data STAR: A Benchmark for Situated Reasoning in Real-World Videos
Reference 135
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 64665e60-14af-4d31-81d3-bcfef72206ec · inbound
Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling STAR: A Benchmark for Situated Reasoning in Real-World Videos
Reference 260
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 062db950-583c-4fbb-871d-02d519762ad8 · inbound
Rethinking Causal Mask Attention for Vision-Language Inference STAR: A Benchmark for Situated Reasoning in Real-World Videos
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7ee1a5e5-7b2c-498f-90ae-a196f9785ab4 · inbound
ScaleLong: A Multi-Timescale Benchmark for Long Video Understanding STAR: A Benchmark for Situated Reasoning in Real-World Videos
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b23e32ea-f4c2-45bd-a21a-5b55d020807b · inbound
Reinforcing Video Reasoning with Focused Thinking STAR: A Benchmark for Situated Reasoning in Real-World Videos
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 759cfd64-7b93-4d6e-9e2e-1b9fa7556250 · inbound
InterRVOS: Interaction-aware Referring Video Object Segmentation STAR: A Benchmark for Situated Reasoning in Real-World Videos
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 28fd6f61-85cf-4000-9df3-5bdea4755f80 · inbound
MAGNET: A Multi-agent Framework for Finding Audio-Visual Needles by Reasoning over Multi-Video Haystacks STAR: A Benchmark for Situated Reasoning in Real-World Videos
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b0beb357-0bc5-4803-9b93-2e37a27c3837 · inbound
TOGA: Temporally Grounded Open-Ended Video QA with Weak Supervision STAR: A Benchmark for Situated Reasoning in Real-World Videos
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9a8e578c-f3ed-4db6-b45f-5eba1d035eca · inbound
VRBench: A Benchmark for Multi-Step Reasoning in Long Narrative Videos STAR: A Benchmark for Situated Reasoning in Real-World Videos
Reference 76
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9fc06ce8-f3a8-4ffd-a395-9a0bd9cf8611 · inbound
IPFormer-VideoLLM: Enhancing Multi-modal Video Understanding for Multi-shot Scenes STAR: A Benchmark for Situated Reasoning in Real-World Videos
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fd965cf1-4793-44c8-9b71-6bfdcfe742e1 · inbound
InterAct-Video: Reasoning-Rich Video QA for Urban Traffic STAR: A Benchmark for Situated Reasoning in Real-World Videos
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bccc7492-cf70-4dea-a402-134cf0aea9e6 · inbound
POVQA: Preference-Optimized Video Question Answering with Rationales for Data Efficiency STAR: A Benchmark for Situated Reasoning in Real-World Videos
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 905837be-2b9a-4bfc-bfc4-713a1d332ae9 · inbound
VKnowU: Evaluating Visual Knowledge Understanding in Multimodal LLMs STAR: A Benchmark for Situated Reasoning in Real-World Videos
Reference 94
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 735132a7-13bc-4a9b-b53a-e64c0780247c · inbound
Agentic Physical AI toward a Domain-Specific Foundation Model for Nuclear Reactor Control STAR: A Benchmark for Situated Reasoning in Real-World Videos
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 120f12dd-3a86-4f5d-9f94-c11e9fbba5a2 · inbound
Mimic Human Cognition, Master Multi-Image Reasoning: A Meta-Action Framework for Enhanced Visual Understanding STAR: A Benchmark for Situated Reasoning in Real-World Videos
Reference 73
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 66a596ee-822f-4326-9cc1-e40f7e93a353 · inbound
Seeing the Scene Matters: Revealing Forgetting in Video Understanding Models with a Scene-Aware Long-Video Benchmark STAR: A Benchmark for Situated Reasoning in Real-World Videos
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation cc5e434d-a353-4b6b-8158-55286c0d98ed · inbound
VideoStir: Understanding Long Videos via Spatio-Temporally Structured and Intent-Aware RAG STAR: A Benchmark for Situated Reasoning in Real-World Videos
Reference 74
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 123b9b79-0ab7-486c-a028-10ec04398d9d · inbound
Chain-of-Glimpse: Search-Guided Progressive Object-Grounded Reasoning for Video Understanding STAR: A Benchmark for Situated Reasoning in Real-World Videos
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 1ed24a58-b4f4-4690-898a-625f32d58e0e · inbound
Chain-of-Glimpse: Search-Guided Progressive Object-Grounded Reasoning for Video Understanding STAR: A Benchmark for Situated Reasoning in Real-World Videos
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 5aa56276-6bf3-47d2-9082-8289ec639140 · inbound
NICE FACT: Diagnosing and Calibrating VLMs in Quantitative Reasoning for Kinematic Physics STAR: A Benchmark for Situated Reasoning in Real-World Videos
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation fb947a5c-c059-4932-a558-f2947109e8d2 · inbound
EgoMemReason: A Memory-Driven Reasoning Benchmark for Long-Horizon Egocentric Video Understanding STAR: A Benchmark for Situated Reasoning in Real-World Videos
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation f67abafe-e20d-4377-87db-50dc8928964e · inbound
TOC-Bench: A Temporal Object Consistency Benchmark for Video Large Language Models STAR: A Benchmark for Situated Reasoning in Real-World Videos
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 3da4fe23-a46b-445d-9362-f21ccd5487f8 · inbound
TOC-Bench: A Temporal Object Consistency Benchmark for Video Large Language Models STAR: A Benchmark for Situated Reasoning in Real-World Videos
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 15ac86a5-8aaf-41b1-ad16-b6159506d969 · inbound
OProver: A Unified Framework for Agentic Formal Theorem Proving STAR: A Benchmark for Situated Reasoning in Real-World Videos
Reference 83
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 27867a6f-0e53-492e-93be-22038c983a16 · inbound
EvoVid: Temporal-Centric Self-Evolution for Video Large Language Models STAR: A Benchmark for Situated Reasoning in Real-World Videos
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 0e77163e-16be-4107-86bd-f14845461c73 · inbound
What-If World: A Causal Benchmark for General World Models in Embodied Scenarios STAR: A Benchmark for Situated Reasoning in Real-World Videos
Reference 68
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation f6646ba2-5194-4238-8731-f1f8dca3350d · inbound
StoryVideoQA: Scaling Deep Video Understanding with a Large-Scale, Multi-Genre and Auto-Generated Dataset STAR: A Benchmark for Situated Reasoning in Real-World Videos
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation a0677d9f-3162-40e3-93f3-a10432fe406a · inbound
ChronoPhyBench: Do MLLMs Truly Understand the World or Merely Exploit Language Priors? STAR: A Benchmark for Situated Reasoning in Real-World Videos
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation c14895da-190b-4922-aee1-fe6cb16ea015 · inbound
FeVOS: Foresight Expression Video Object Segmentation STAR: A Benchmark for Situated Reasoning in Real-World Videos
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation ae2fc7de-4929-4e7f-91f1-86a672fb19dc · inbound
Reflect-R1: Evidence-Driven Reflection for Self-Correction in Long Video Understanding STAR: A Benchmark for Situated Reasoning in Real-World Videos
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 735146d8-3a12-4ed6-b5a0-2d2961f19a2f · inbound
Reflect-R1: Evidence-Driven Reflection for Self-Correction in Long Video Understanding STAR: A Benchmark for Situated Reasoning in Real-World Videos
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 80288e54-4781-49cd-ad48-7730585a28a2 · inbound
Reflect-R1: Evidence-Driven Reflection for Self-Correction in Long Video Understanding STAR: A Benchmark for Situated Reasoning in Real-World Videos
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 13711394-58a0-4b2c-a88a-453d55d6f546 · inbound
HumanMoveVQA: Can Video MLLMs reason about human movement in videos? STAR: A Benchmark for Situated Reasoning in Real-World Videos
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 0978f925-f5ae-4cf8-97a3-fee8566e1573 · inbound
HumanMoveVQA: Can Video MLLMs reason about human movement in videos? STAR: A Benchmark for Situated Reasoning in Real-World Videos
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation a4642844-0455-4f73-8e94-04cae7b4956a · inbound
H-GRPO: Permutation-Invariant Reinforcement Learning for Grounded Visual Reasoning STAR: A Benchmark for Situated Reasoning in Real-World Videos
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 6bf62ec9-5422-40a6-9e7f-ff5da6e0f4dd · inbound
MuseBench: Benchmarking Intent-Level Audiovisual Arts Understanding in MLLMs STAR: A Benchmark for Situated Reasoning in Real-World Videos
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 88dcadad-d535-4981-8826-bc546bb0920b · inbound
IMBench: A Benchmark for Intuitive Robotic Manipulation STAR: A Benchmark for Situated Reasoning in Real-World Videos
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.