Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 40 inbound Pith citation observations for arXiv:2408.14023.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T23:42:35.215228Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-04T09:39:46.765108Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation e8bbff39-2518-420d-8b3f-584957ae697c · inbound
MLVU: Benchmarking Multi-task Long Video Understanding Video-CCAM: Enhancing Video-Language Understanding with Causal Cross-Attention Masks for Short and Long Videos
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 8ce1ff7b-9ec1-483c-9524-5056486179ab · inbound
LongVILA: Scaling Long-Context Visual Language Models for Long Videos Video-CCAM: Enhancing Video-Language Understanding with Causal Cross-Attention Masks for Short and Long Videos
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 23d85641-bcfe-4bb7-b319-6c101585e709 · inbound
VideoChat-Flash: Hierarchical Compression for Long-Context Video Modeling Video-CCAM: Enhancing Video-Language Understanding with Causal Cross-Attention Masks for Short and Long Videos
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation a1e3ce7b-6cf8-4833-9414-91be823de78e · inbound
InternVideo2.5: Empowering Video MLLMs with Long and Rich Context Modeling Video-CCAM: Enhancing Video-Language Understanding with Causal Cross-Attention Masks for Short and Long Videos
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation fe4d3805-a816-4272-b56a-86137f4f33ec · inbound
VideoLLaMA 3: Frontier Multimodal Foundation Models for Image and Video Understanding Video-CCAM: Enhancing Video-Language Understanding with Causal Cross-Attention Masks for Short and Long Videos
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation ce17acaf-da91-44b8-a33a-65f9f5f53e0c · inbound
VideoEval-Pro: Robust and Realistic Long Video Understanding Evaluation Video-CCAM: Enhancing Video-Language Understanding with Causal Cross-Attention Masks for Short and Long Videos
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4c265e10-9dcf-4610-9ba4-825985f4bb4d · inbound
ViaRL: Adaptive Temporal Grounding via Visual Iterated Amplification Reinforcement Learning Video-CCAM: Enhancing Video-Language Understanding with Causal Cross-Attention Masks for Short and Long Videos
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 889779c6-2937-4cab-bb15-295b9ff8aeea · inbound
FlexSelect: Flexible Token Selection for Efficient Long Video Understanding Video-CCAM: Enhancing Video-Language Understanding with Causal Cross-Attention Masks for Short and Long Videos
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 97b1818a-42de-4fdb-820d-985568e79bc4 · inbound
LeanPO: Lean Preference Optimization for Likelihood Alignment in Video-LLMs Video-CCAM: Enhancing Video-Language Understanding with Causal Cross-Attention Masks for Short and Long Videos
Reference 59
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7952014f-ff33-40b3-bd24-7d4a7bfe2962 · inbound
SparseMM: Head Sparsity Emerges from Visual Concept Responses in MLLMs Video-CCAM: Enhancing Video-Language Understanding with Causal Cross-Attention Masks for Short and Long Videos
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d632ca7f-a4c0-4de8-9889-4d3d58abaa42 · inbound
Video-XL-2: Towards Very Long-Video Understanding Through Task-Aware KV Sparsification Video-CCAM: Enhancing Video-Language Understanding with Causal Cross-Attention Masks for Short and Long Videos
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 29079ea3-7f5a-45e3-b673-b96d566aed81 · inbound
Task-Aware KV Compression For Cost-Effective Long Video Understanding Video-CCAM: Enhancing Video-Language Understanding with Causal Cross-Attention Masks for Short and Long Videos
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 55a5ff53-44d3-4273-8475-82950400326c · inbound
Q-Frame: Query-aware Frame Selection and Multi-Resolution Adaptation for Video-LLMs Video-CCAM: Enhancing Video-Language Understanding with Causal Cross-Attention Masks for Short and Long Videos
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bb93b653-36b1-4e29-a744-433a9651ab04 · inbound
LongAnimation: Long Animation Generation with Dynamic Global-Local Memory Video-CCAM: Enhancing Video-Language Understanding with Causal Cross-Attention Masks for Short and Long Videos
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d8855dac-cc08-4c80-8e64-ec425fe59203 · inbound
VideoForest: Person-Anchored Hierarchical Reasoning for Cross-Video Question Answering Video-CCAM: Enhancing Video-Language Understanding with Causal Cross-Attention Masks for Short and Long Videos
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 444ed85b-6e62-4477-8c1d-a427b44b15fa · inbound
Video-MTR: Reinforced Multi-Turn Reasoning for Long Video Understanding Video-CCAM: Enhancing Video-Language Understanding with Causal Cross-Attention Masks for Short and Long Videos
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f875cfc8-e833-497e-87d3-73a3ddd9fb79 · inbound
LongVideo-R1: Smart Navigation for Low-cost Long Video Understanding Video-CCAM: Enhancing Video-Language Understanding with Causal Cross-Attention Masks for Short and Long Videos
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation ed57979b-9d67-4aca-9efa-d64112ba6835 · inbound
Seeing the Scene Matters: Revealing Forgetting in Video Understanding Models with a Scene-Aware Long-Video Benchmark Video-CCAM: Enhancing Video-Language Understanding with Causal Cross-Attention Masks for Short and Long Videos
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 1c61831f-6a19-4f9f-a654-df1feb208d45 · inbound
AdaSpark: Adaptive Sparsity for Efficient Long-Video Understanding Video-CCAM: Enhancing Video-Language Understanding with Causal Cross-Attention Masks for Short and Long Videos
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation f2f9886c-ae29-49aa-9743-76d23ee107a3 · inbound
One Token per Highly Selective Frame: Towards Extreme Compression for Long Video Understanding Video-CCAM: Enhancing Video-Language Understanding with Causal Cross-Attention Masks for Short and Long Videos
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation fdd603bd-a208-4c6a-bfda-351152e3382f · inbound
VideoRouter: Query-Adaptive Dual Routing for Efficient Long-Video Understanding Video-CCAM: Enhancing Video-Language Understanding with Causal Cross-Attention Masks for Short and Long Videos
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation e0758a7f-de1a-4979-95b9-fb764a784d0b · inbound
VideoRouter: Query-Adaptive Dual Routing for Efficient Long-Video Understanding Video-CCAM: Enhancing Video-Language Understanding with Causal Cross-Attention Masks for Short and Long Videos
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation f8cf6d9b-4edd-435d-8e9e-65ffefa10bf4 · inbound
Response-G1: Explicit Scene Graph Modeling for Proactive Streaming Video Understanding Video-CCAM: Enhancing Video-Language Understanding with Causal Cross-Attention Masks for Short and Long Videos
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 0a27ff8e-79be-4bdd-9476-ea62c810ea37 · inbound
Response-G1: Explicit Scene Graph Modeling for Proactive Streaming Video Understanding Video-CCAM: Enhancing Video-Language Understanding with Causal Cross-Attention Masks for Short and Long Videos
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 353b1242-ee1c-4f1f-8d1c-6fcd19ab717f · inbound
An Efficient Streaming Video Understanding Framework with Agentic Control Video-CCAM: Enhancing Video-Language Understanding with Causal Cross-Attention Masks for Short and Long Videos
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 4fd44adf-2d77-4b56-b812-061e54bb2403 · inbound
Lance: Unified Multimodal Modeling by Multi-Task Synergy Video-CCAM: Enhancing Video-Language Understanding with Causal Cross-Attention Masks for Short and Long Videos
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 60c46325-5641-4837-a59b-d9537c6f88d3 · inbound
Lance: Unified Multimodal Modeling by Multi-Task Synergy Video-CCAM: Enhancing Video-Language Understanding with Causal Cross-Attention Masks for Short and Long Videos
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 22a9fa5d-4c61-43d1-8f22-ee450655e785 · inbound
Swift Sampling: Selecting Temporal Surprises via Taylor Series Video-CCAM: Enhancing Video-Language Understanding with Causal Cross-Attention Masks for Short and Long Videos
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 433ae2e0-4eaa-49c3-a4f7-36f6b24270f2 · inbound
STORM: Internalized Modeling for Spatial-Temporal Reasoning in Video-Language Models Video-CCAM: Enhancing Video-Language Understanding with Causal Cross-Attention Masks for Short and Long Videos
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 172a8c6a-851a-49fb-8c07-0e34cf02566d · inbound
Towards Effective Long-Video Event Prediction via Multi-Level Event Semantics Mining Video-CCAM: Enhancing Video-Language Understanding with Causal Cross-Attention Masks for Short and Long Videos
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation a60abe6e-cf91-4903-b0b6-abe486c2ae41 · inbound
GOPAgen: Motion-Aware and Efficient Agentic Long-Video Understanding with Structural Memory and Hierarchical Reasoning Video-CCAM: Enhancing Video-Language Understanding with Causal Cross-Attention Masks for Short and Long Videos
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation b06a113c-fe5c-436d-904e-14032ee29aaa · inbound
Q-Fold: Query-Aware Focus-Context Spatio-Temporal Folding for Long Video Understanding Video-CCAM: Enhancing Video-Language Understanding with Causal Cross-Attention Masks for Short and Long Videos
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 612534ff-5173-43db-934e-1bda7fb6300c · inbound
InternVideo3: Agentify Foundation Models with Multimodal Contextual Reasoning Video-CCAM: Enhancing Video-Language Understanding with Causal Cross-Attention Masks for Short and Long Videos
Reference 249
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 1cefe5d0-9b92-4e34-afdc-c43f7e417913 · inbound
HPP: Hierarchical Programmatic Probing for Long Video Understanding by Decoupling Perception and Reasoning Video-CCAM: Enhancing Video-Language Understanding with Causal Cross-Attention Masks for Short and Long Videos
Reference 140
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation dbaf52bc-d04f-4e4f-a079-f604648b05b8 · inbound
CoVStream: Edge-Cloud Collaboration for Understanding of Long Video Streams Video-CCAM: Enhancing Video-Language Understanding with Causal Cross-Attention Masks for Short and Long Videos
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 8611e1e8-588d-4426-9146-fe6284f1944d · inbound
MuseBench: Benchmarking Intent-Level Audiovisual Arts Understanding in MLLMs Video-CCAM: Enhancing Video-Language Understanding with Causal Cross-Attention Masks for Short and Long Videos
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 440a1036-f026-47e1-8c1a-cf37f656f759 · inbound
TimeThink: Reasoning with Time for Video LLMs Video-CCAM: Enhancing Video-Language Understanding with Causal Cross-Attention Masks for Short and Long Videos
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2bbef040-3c0b-47ca-98da-46922382d900 · inbound
Efficient Frame Selection for Long Videos at Test Time with Attention-Based MLLM Selectors Video-CCAM: Enhancing Video-Language Understanding with Causal Cross-Attention Masks for Short and Long Videos
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dc306c73-3c39-4073-9433-30caa848ad63 · inbound
Reasoning with Memory: A Temporal Granularity-Adaptive Framework for Training-Free Long Video Understanding Video-CCAM: Enhancing Video-Language Understanding with Causal Cross-Attention Masks for Short and Long Videos
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2adaf088-b632-4f60-a978-61943e5b8a5f · inbound
Evidence-Driven Dynamic Visual Selector for Efficient Long Video Understanding Video-CCAM: Enhancing Video-Language Understanding with Causal Cross-Attention Masks for Short and Long Videos
Reference 60
Source-reported events for the cited work
Unavailable: canonical work link unavailable.