Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-16T10:50:29.935626Z
Paper Citation Record · LEDGER
As of 18 August 2026, this Paper Citation Record lists 53 of 53 outbound references and 2 inbound Pith citation observations for arXiv:2504.17828.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-16T10:50:29.935626Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T11:08:41.965687Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-13T21:18:17.187055Z
53 of 53 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation df1542c8-2c67-4624-a15c-24caeb6f6067 · outbound
VEU-Bench: Towards Comprehensive Understanding of Video Editing The anatomy of video editing: A dataset and benchmark suite for ai-assisted video editing
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation beb144f7-db10-4bfe-8dd2-a481b1edba75 · outbound
VEU-Bench: Towards Comprehensive Understanding of Video Editing Theory of film practice
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 5595545d-4e5a-44c4-ab37-17a59303d935 · outbound
VEU-Bench: Towards Comprehensive Understanding of Video Editing Reframe Anything: LLM Agent for Open World Video Reframing
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2a79a1f4-8a33-416d-a960-f09d74109829 · outbound
VEU-Bench: Towards Comprehensive Understanding of Video Editing Match cutting: Finding cuts with smooth visual transitions
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation c7dfa941-13be-4f45-b591-a253d06c7c01 · outbound
VEU-Bench: Towards Comprehensive Understanding of Video Editing ShareGPT4Video: Improving Video Understanding and Generation with Better Captions
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b3db770a-b655-4bb7-aa36-8e01199c7ae3 · outbound
VEU-Bench: Towards Comprehensive Understanding of Video Editing How Far Are We to GPT-4V? Closing the Gap to Commercial Multimodal Models with Open-Source Suites
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3351857e-fa43-432f-bef0-26ec0cc6b496 · outbound
VEU-Bench: Towards Comprehensive Understanding of Video Editing VideoLLaMA 2: Advancing Spatial-Temporal Modeling and Audio Understanding in Video-LLMs
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a754fa32-211a-45d2-8038-8f7b487133f4 · outbound
VEU-Bench: Towards Comprehensive Understanding of Video Editing The technique of film and video editing: history, theory, and practice
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 3a9f687c-a050-4582-a0ea-c9b0cf0357e5 · outbound
VEU-Bench: Towards Comprehensive Understanding of Video Editing Video-CCAM: Enhancing Video-Language Understanding with Causal Cross-Attention Masks for Short and Long Videos
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0c919074-4ccd-4001-8559-1770e74ba05e · outbound
VEU-Bench: Towards Comprehensive Understanding of Video Editing Automatic Non-Linear Video Editing Transfer
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation edd27ce0-1296-4e94-a87b-ef66ac6f0892 · outbound
VEU-Bench: Towards Comprehensive Understanding of Video Editing Video-MME: The First-Ever Comprehensive Evaluation Benchmark of Multi-modal LLMs in Video Analysis
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8ed59f42-0b90-4f8f-9ef0-30986f331394 · outbound
VEU-Bench: Towards Comprehensive Understanding of Video Editing Edit3K: Universal Representation Learning for Video Editing Components
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d4005da3-0840-4488-aed2-b1e65b38636b · outbound
VEU-Bench: Towards Comprehensive Understanding of Video Editing LoRA: Low-Rank Adaptation of Large Language Models
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e261ab51-bb24-4474-9371-e2408318bfbf · outbound
VEU-Bench: Towards Comprehensive Understanding of Video Editing Vtimellm: Empower llm to grasp video moments
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 73678a1a-6d14-422b-8d89-2a769e4d20a2 · outbound
VEU-Bench: Towards Comprehensive Understanding of Video Editing Movienet: A holistic dataset for movie understanding
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 68d33af1-7e79-4102-a671-fd5fade83331 · outbound
VEU-Bench: Towards Comprehensive Understanding of Video Editing Study of vari- ous video annotation techniques
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 4a8e3af3-20fc-4ea2-a60b-8b3d86b630f6 · outbound
VEU-Bench: Towards Comprehensive Understanding of Video Editing Automatic color scheme extraction from movies
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation e89586d6-630c-443f-b1bc-0021e5cc66f7 · outbound
VEU-Bench: Towards Comprehensive Understanding of Video Editing Frame Order Matters: A Temporal Sequence-Aware Model for Few-Shot Action Recognition
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4b0d1287-fdb6-4ba6-9f4e-21bb5ea600d7 · outbound
VEU-Bench: Towards Comprehensive Understanding of Video Editing LLaVA-OneVision: Easy Visual Task Transfer
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2d400170-9e9e-4a89-9a06-69b9794aa5bf · outbound
VEU-Bench: Towards Comprehensive Understanding of Video Editing Mvbench: A comprehensive multi-modal video understand- ing benchmark
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 8f0db0df-d6d7-4151-967e-e4e583c20c2c · outbound
VEU-Bench: Towards Comprehensive Understanding of Video Editing Video-LLaVA: Learning United Visual Representation by Alignment Before Projection
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7aea31df-fcd5-40dc-a327-bff6b98039b0 · outbound
VEU-Bench: Towards Comprehensive Understanding of Video Editing Vila: On pre-training for visual language models
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 73917e7a-15a5-4439-9230-725c43bc659a · outbound
VEU-Bench: Towards Comprehensive Understanding of Video Editing OmniCLIP: Adapting CLIP for Video Recognition with Spatial-Temporal Omni-Scale Feature Learning
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 468a9f86-0350-4cb1-827d-ae8a6ee09610 · outbound
VEU-Bench: Towards Comprehensive Understanding of Video Editing TempCompass: Do Video LLMs Really Understand Videos?
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c9a6248b-b904-44db-aff0-78109d8466af · outbound
VEU-Bench: Towards Comprehensive Understanding of Video Editing Oryx MLLM: On-Demand Spatial-Temporal Understanding at Arbitrary Resolution
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4d84cdba-c51d-4582-84fe-e8517857df54 · outbound
VEU-Bench: Towards Comprehensive Understanding of Video Editing Decoupled Weight Decay Regularization
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dc7bf747-f927-41c1-bf94-44379ec167ec · outbound
VEU-Bench: Towards Comprehensive Understanding of Video Editing Video-ChatGPT: Towards Detailed Video Understanding via Large Vision and Language Models
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6cd29212-5daf-4d76-96ad-008046d093cd · outbound
VEU-Bench: Towards Comprehensive Understanding of Video Editing Video-chatgpt: Towards detailed video understanding via large vision and language models
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation f7beb7ce-554f-4a46-aff1-e5cb0db0f250 · outbound
VEU-Bench: Towards Comprehensive Understanding of Video Editing Film language: A semiotics of the cinema
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 6fc9c3f4-7bf4-4c71-a0c4-9899432b3a79 · outbound
VEU-Bench: Towards Comprehensive Understanding of Video Editing Hello gpt-4o
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 428a71d4-79e1-4026-8e88-f4986202a2c6 · outbound
VEU-Bench: Towards Comprehensive Understanding of Video Editing Learning to cut by watch- ing movies
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 65b45f01-ef13-46c5-bce8-816fb3f72caf · outbound
VEU-Bench: Towards Comprehensive Understanding of Video Editing Moviecuts: A new dataset and benchmark for cut type recognition
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 541fefd0-a7db-4db1-859d-e0fc4c8a4870 · outbound
VEU-Bench: Towards Comprehensive Understanding of Video Editing Perception test: A diagnostic benchmark for multimodal video models
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 53ddda76-f0fc-4a95-a994-a3dc9da77f97 · outbound
VEU-Bench: Towards Comprehensive Understanding of Video Editing Autotran- sition: Learning to recommend video transition effects
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation f1e1463d-c381-4a44-ac2b-d25a5958e5ac · outbound
VEU-Bench: Towards Comprehensive Understanding of Video Editing Harnessing ai for augmenting creativity: Appli- cation to movie trailer creation
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 7ad4a3b3-9abe-4f14-ac99-62a99f5824c9 · outbound
VEU-Bench: Towards Comprehensive Understanding of Video Editing Diffusion Model-Based Video Editing: A Survey
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e8576dfe-7d15-40e7-adcc-918118619203 · outbound
VEU-Bench: Towards Comprehensive Understanding of Video Editing Gemini: A Family of Highly Capable Multimodal Models
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c53bf548-3e40-45fe-b2c8-ff5bc67b71da · outbound
VEU-Bench: Towards Comprehensive Understanding of Video Editing Converting video formats with ffmpeg
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation e317ea7e-28aa-4d83-87a0-e22badb7d63b · outbound
VEU-Bench: Towards Comprehensive Understanding of Video Editing Movie lens: Discovering and characterizing editing patterns in the anal- ysis of short movie sequences
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 83adc78b-320e-415b-9731-e83f95bff8ac · outbound
VEU-Bench: Towards Comprehensive Understanding of Video Editing Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cb1685b3-a5b2-45f8-b21b-0ab8809b8b54 · outbound
VEU-Bench: Towards Comprehensive Understanding of Video Editing LVBench: An Extreme Long Video Understanding Benchmark
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 45929c96-d7fa-48d0-ab14-aa397f109fc6 · outbound
VEU-Bench: Towards Comprehensive Understanding of Video Editing Editing techniques with final cut pro
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 8e86f818-dcd2-4741-a4ce-57d825c415fe · outbound
VEU-Bench: Towards Comprehensive Understanding of Video Editing LongVideoBench: A Benchmark for Long-context Interleaved Video-Language Understanding
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 16a12aa8-eff9-408a-b1d3-77ba01e12531 · outbound
VEU-Bench: Towards Comprehensive Understanding of Video Editing Zero-Shot Long-Form Video Understanding through Screenplay
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation c31df461-7bc2-43e2-a2a2-480719ea4c3b · outbound
VEU-Bench: Towards Comprehensive Understanding of Video Editing Video Repurposing from User Generated Content: A Large-scale Dataset and Benchmark
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 227d5305-fddf-43ff-aa02-e6f2529b05f7 · outbound
VEU-Bench: Towards Comprehensive Understanding of Video Editing Beyond Raw Videos: Understanding Edited Videos with Large Multimodal Model
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 18bda182-7892-4052-bb7a-c37ea413cccc · outbound
VEU-Bench: Towards Comprehensive Understanding of Video Editing LongVILA: Scaling Long-Context Visual Language Models for Long Videos
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2d406e3a-6e76-425a-aefd-ba39fb6a3d8c · outbound
VEU-Bench: Towards Comprehensive Understanding of Video Editing MiniCPM-V: A GPT-4V Level MLLM on Your Phone
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c4431211-cab5-45ba-86fa-5796e0e7adb9 · outbound
VEU-Bench: Towards Comprehensive Understanding of Video Editing mPLUG-Owl3: Towards Long Image-Sequence Understanding in Multi-Modal Large Language Models
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 436a70dd-7435-4c9d-af02-12d9b814f80d · outbound
VEU-Bench: Towards Comprehensive Understanding of Video Editing Video-LLaMA: An Instruction-tuned Audio-Visual Language Model for Video Understanding
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d761f01e-d6ef-4c4b-9afb-59c0bba72a75 · outbound
VEU-Bench: Towards Comprehensive Understanding of Video Editing InternLM-XComposer-2.5: A Versatile Large Vision Language Model Supporting Long-Contextual Input and Output
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7dee8977-a33c-49e8-9fc7-dd7533c02385 · outbound
VEU-Bench: Towards Comprehensive Understanding of Video Editing LLaVA-Video: Video Instruction Tuning With Synthetic Data
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8f12a558-3aa6-4041-9e8e-bf707c9c3e40 · outbound
VEU-Bench: Towards Comprehensive Understanding of Video Editing MLVU: Benchmarking Multi-task Long Video Understanding
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 86bf2cb7-94da-403f-8ed4-05db4472013b · inbound
RSVP: Reasoning Segmentation via Visual Prompting and Multi-modal Chain-of-Thought VEU-Bench: Towards Comprehensive Understanding of Video Editing
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d47f2d1d-dfd3-4cfa-b3df-33df2b972b46 · inbound
VERTIGO: Visual Preference Optimization for Cinematic Camera Trajectory Generation VEU-Bench: Towards Comprehensive Understanding of Video Editing
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.