Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 22 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 27 inbound Pith citation observations for arXiv:2308.09126.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-22T06:32:14.747728+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-16T10:50:21.460707Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-04T16:39:57.306434Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation 64afd474-cd84-470f-a8c7-086f8848c8d4 · inbound
MVBench: A Comprehensive Multi-modal Video Understanding Benchmark EgoSchema: A Diagnostic Benchmark for Very Long-form Video Language Understanding
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 46e6c3fb-224b-45ad-9dca-67673151045b · inbound
VCA: Video Curious Agent for Long Video Understanding EgoSchema: A Diagnostic Benchmark for Very Long-form Video Language Understanding
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8e5744cd-a924-45b2-8ab4-c4eb5b1f5cca · inbound
Online Video Understanding: OVBench and VideoChat-Online EgoSchema: A Diagnostic Benchmark for Very Long-form Video Language Understanding
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 19a7a086-642e-4a1c-8375-70f6a74947f8 · inbound
$\infty$-Video: A Training-Free Approach to Long Video Understanding via Continuous-Time Memory Consolidation EgoSchema: A Diagnostic Benchmark for Very Long-form Video Language Understanding
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 68a5db79-232f-41c0-be4a-2b388ff281c4 · inbound
MASR: Self-Reflective Reasoning through Multimodal Hierarchical Attention Focusing for Agent-based Video Understanding EgoSchema: A Diagnostic Benchmark for Very Long-form Video Language Understanding
Reference 64
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2942772e-a2f3-4211-b5d8-6c2a222f5199 · inbound
VideoLLM Benchmarks and Evaluation: A Survey EgoSchema: A Diagnostic Benchmark for Very Long-form Video Language Understanding
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d3ebb7c6-b8f5-4554-81cc-2595dd0a4c2b · inbound
Movie Facts and Fibs (MF$^2$): A Benchmark for Long Movie Understanding EgoSchema: A Diagnostic Benchmark for Very Long-form Video Language Understanding
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9e01ba51-5333-48b2-be59-6a13003a0f9b · inbound
CausalVQA: A Physically Grounded Causal Reasoning Benchmark for Video Models EgoSchema: A Diagnostic Benchmark for Very Long-form Video Language Understanding
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 468d22ad-8ab8-4b9f-b59b-9d312cb0db29 · inbound
LaCo: Efficient Layer-wise Compression of Visual Tokens for Multimodal Large Language Models EgoSchema: A Diagnostic Benchmark for Very Long-form Video Language Understanding
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e0001093-1c8d-428c-8368-db9820138888 · inbound
Infinite Video Understanding EgoSchema: A Diagnostic Benchmark for Very Long-form Video Language Understanding
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 86434552-e4a1-40e7-b298-5d29ddd11a74 · inbound
HERMES: KV Cache as Hierarchical Memory for Efficient Streaming Video Understanding EgoSchema: A Diagnostic Benchmark for Very Long-form Video Language Understanding
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 05cf38b2-0bf8-4475-b41a-9d86e78056f9 · inbound
Adaptive Greedy Frame Selection for Long Video Understanding EgoSchema: A Diagnostic Benchmark for Very Long-form Video Language Understanding
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation e0e18d26-0bf9-486f-b1f2-9c81172d79f5 · inbound
MACS: Modality-Aware Capacity Scaling for Efficient Multimodal MoE Inference EgoSchema: A Diagnostic Benchmark for Very Long-form Video Language Understanding
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation c04ffac2-679e-424d-a546-0c9ed249b1ff · inbound
HumanNet: Scaling Human-centric Video Learning to One Million Hours EgoSchema: A Diagnostic Benchmark for Very Long-form Video Language Understanding
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation f0cc8be1-9168-4401-b590-c2240d6e1ad1 · inbound
EgoCoT-Bench: Benchmarking Grounded and Verifiable Operation-Centric Chain of Thought Reasoning for MLLMs EgoSchema: A Diagnostic Benchmark for Very Long-form Video Language Understanding
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation ccad6fc8-7167-48d0-a6bc-5f44830d7c13 · inbound
EgoProx: Evaluating MLLMs on Egocentric 3D Proximity Reasoning Across a Cognitive Hierarchy EgoSchema: A Diagnostic Benchmark for Very Long-form Video Language Understanding
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation dc68e372-e42d-4d31-9ffa-227accf43bb7 · inbound
TeachObs: A Human-Validated Benchmark for Multimodal Teaching Observation and Model Evaluation EgoSchema: A Diagnostic Benchmark for Very Long-form Video Language Understanding
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 09494012-2fb3-4efb-8910-3abc33393d82 · inbound
Pause and Think: A Dataset and Benchmark for Video-Grounded Assistive Action Suggestion EgoSchema: A Diagnostic Benchmark for Very Long-form Video Language Understanding
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation b38608ab-01fb-4bc1-b61a-febc772f53a0 · inbound
Pause and Think: A Dataset and Benchmark for Video-Grounded Assistive Action Suggestion EgoSchema: A Diagnostic Benchmark for Very Long-form Video Language Understanding
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation a35d243a-9411-42c6-a6eb-2415516408d5 · inbound
SuperMemory-VQA: An Egocentric Visual Question-Answering Benchmark for Long-Horizon Memory EgoSchema: A Diagnostic Benchmark for Very Long-form Video Language Understanding
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation b33479a2-f95b-47b1-bb50-245f20f7b627 · inbound
NextMotionQA: Benchmarking and Judging Human Motion Understanding with Vision-Language Models EgoSchema: A Diagnostic Benchmark for Very Long-form Video Language Understanding
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 2ceec0df-85e7-4ce4-8fab-9895482f6318 · inbound
MOSS-Video-Preview: Toward Real-Time Video Understanding via Cross-Attention EgoSchema: A Diagnostic Benchmark for Very Long-form Video Language Understanding
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 4f2851b0-9617-47dc-92f0-6fb59e21afad · inbound
EgoSAT: A Comprehensive Benchmark of Egocentric Streaming Interaction Understanding EgoSchema: A Diagnostic Benchmark for Very Long-form Video Language Understanding
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 365b3d41-34c0-4c51-bf09-e04b99c0b6c3 · inbound
From Accuracy to Visual Dependence: Auditing and Filtering Modality Collapse in Traffic VideoQA EgoSchema: A Diagnostic Benchmark for Very Long-form Video Language Understanding
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation dbf7d4f3-fae6-4381-8f1b-f91477ad19b1 · inbound
MEMORA: Embodied Action Memory from Egocentric Videos for Reasoning and Planning EgoSchema: A Diagnostic Benchmark for Very Long-form Video Language Understanding
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b7e99be9-fe47-41e4-a095-7abe1408cc75 · inbound
ObjectStream: Latent Objects as Memory Anchors for Streaming Video Understanding EgoSchema: A Diagnostic Benchmark for Very Long-form Video Language Understanding
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6e0885d0-099b-498c-9184-bfdcca6e8f75 · inbound
CRAFT: Compression via Recursive Adaptive Fusion of Video Tokens for Vision-Language Models EgoSchema: A Diagnostic Benchmark for Very Long-form Video Language Understanding
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.