Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 24 inbound Pith citation observations for arXiv:2412.12075.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T12:32:50.149539Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-04T00:49:17.542022Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation 6fedff7f-aaec-4ec1-8767-dea3b12c325d · inbound
MUSEG: Reinforcing Video Temporal Understanding via Timestamp-Aware Multi-Segment Grounding CG-Bench: Clue-grounded Question Answering Benchmark for Long Video Understanding
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation e150a606-02fb-4a4e-8b6f-c23c42111155 · inbound
DisTime: Distribution-based Time Representation for Video Large Language Models CG-Bench: Clue-grounded Question Answering Benchmark for Long Video Understanding
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 02223deb-a7f6-4379-8ff1-a37f4262b12d · inbound
AV-Reasoner: Improving and Benchmarking Clue-Grounded Audio-Visual Counting for MLLMs CG-Bench: Clue-grounded Question Answering Benchmark for Long Video Understanding
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e150c5ae-358f-4b63-a745-006588ae60ac · inbound
Bridging Perspectives: A Survey on Cross-view Collaborative Intelligence with Egocentric-Exocentric Vision CG-Bench: Clue-grounded Question Answering Benchmark for Long Video Understanding
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d426b64f-a31c-4848-bfb2-48061f7ec977 · inbound
Movie Facts and Fibs (MF$^2$): A Benchmark for Long Movie Understanding CG-Bench: Clue-grounded Question Answering Benchmark for Long Video Understanding
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation aa3c841a-826e-4063-b441-59d7aee0678b · inbound
VRBench: A Benchmark for Multi-Step Reasoning in Long Narrative Videos CG-Bench: Clue-grounded Question Answering Benchmark for Long Video Understanding
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f0c1fecd-0d5e-46e2-96fd-1bf79aa69b5e · inbound
OmniEval: A Benchmark for Evaluating Omni-modal Models with Visual, Auditory, and Textual Inputs CG-Bench: Clue-grounded Question Answering Benchmark for Long Video Understanding
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 19e47eb9-e53e-4ef6-b09f-014ff4bff5ee · inbound
A Survey on Video Temporal Grounding with Multimodal Large Language Model CG-Bench: Clue-grounded Question Answering Benchmark for Long Video Understanding
Reference 56
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 60dad2f5-2f7a-48c4-8107-13874d6f42b4 · inbound
EMCompress: Video-LLMs with Endomorphic Multimodal Compression CG-Bench: Clue-grounded Question Answering Benchmark for Long Video Understanding
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 5b5217f4-7191-472c-8ce7-66cc9f303bd9 · inbound
REVISOR: Beyond Textual Reflection, Towards Multimodal Introspective Reasoning in Long-Form Video Understanding CG-Bench: Clue-grounded Question Answering Benchmark for Long Video Understanding
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation df201b4e-fc59-42cd-b341-1f002533f6fa · inbound
VideoThinker: Building Agentic VideoLLMs with LLM-Guided Tool Reasoning CG-Bench: Clue-grounded Question Answering Benchmark for Long Video Understanding
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 8cb99da1-f860-46d7-964b-44a6e0bcaba9 · inbound
LongVideo-R1: Smart Navigation for Low-cost Long Video Understanding CG-Bench: Clue-grounded Question Answering Benchmark for Long Video Understanding
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 1e641f19-8de6-4e96-990a-0d80cdbc04a6 · inbound
Efficient Inference for Large Vision-Language Models: Bottlenecks, Techniques, and Prospects CG-Bench: Clue-grounded Question Answering Benchmark for Long Video Understanding
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 267929d3-e852-4c10-a1e2-0fc519398c0b · inbound
POINTS-Long: Adaptive Dual-Mode Visual Reasoning in MLLMs CG-Bench: Clue-grounded Question Answering Benchmark for Long Video Understanding
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation e610b167-f3da-415c-8fe0-d640e963b60a · inbound
Sink-Token-Aware Pruning for Fine-Grained Video Understanding in Efficient Video LLMs CG-Bench: Clue-grounded Question Answering Benchmark for Long Video Understanding
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 420a45c4-ccb8-4b28-ac77-fb6637033bd0 · inbound
Towards Temporal Compositional Reasoning in Long-Form Sports Videos CG-Bench: Clue-grounded Question Answering Benchmark for Long Video Understanding
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bb6e4c7b-c7d6-46af-9ab0-a9de1463aa6a · inbound
EgoMemReason: A Memory-Driven Reasoning Benchmark for Long-Horizon Egocentric Video Understanding CG-Bench: Clue-grounded Question Answering Benchmark for Long Video Understanding
Reference 85
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation c8e94c6d-64a2-4df4-bab7-3e695f1db87e · inbound
VideoOdyssey: A Benchmark for Ultra-Long-Context and Omni-Modal Video Understanding CG-Bench: Clue-grounded Question Answering Benchmark for Long Video Understanding
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 6aa79895-1600-4daf-9aec-21d29475123a · inbound
CaST-Bench: Benchmarking Causal Chain-Grounded Spatio-Temporal Reasoning for Video Question Answering CG-Bench: Clue-grounded Question Answering Benchmark for Long Video Understanding
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 9c8f22a7-42e6-4db0-b5e9-7ccde862993c · inbound
CaST-Bench: Benchmarking Causal Chain-Grounded Spatio-Temporal Reasoning for Video Question Answering CG-Bench: Clue-grounded Question Answering Benchmark for Long Video Understanding
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation f043ac01-df25-4a2e-86ed-07be12eda13a · inbound
Test-Time Scaling in Multimodal Foundation Models: A Comprehensive Survey of Generation and Reasoning CG-Bench: Clue-grounded Question Answering Benchmark for Long Video Understanding
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 6d90e396-1fe9-4aa3-8d49-02c641377eef · inbound
Rethinking RAG in Long Videos: What to Retrieve and How to Use It? CG-Bench: Clue-grounded Question Answering Benchmark for Long Video Understanding
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 432ebe94-bbaf-4073-84f0-3a6c78567a71 · inbound
Incentivizing Vision Language Models to Search for Long Video Question Answering CG-Bench: Clue-grounded Question Answering Benchmark for Long Video Understanding
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a86fe1e0-42ac-4ff6-a65f-c6d62ca74575 · inbound
TimeThink: Reasoning with Time for Video LLMs CG-Bench: Clue-grounded Question Answering Benchmark for Long Video Understanding
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.