Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-10T20:42:31.879391Z
Paper Citation Record · LEDGER
As of 16 August 2026, this Paper Citation Record lists 35 of 35 outbound references and 0 inbound Pith citation observations for arXiv:2501.07554.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-10T20:42:31.879391Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
35 of 35 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 767c366f-a965-4f64-9df7-96d07d65ab07 · outbound
SST-EM: Advanced Metrics for Evaluating Semantic, Spatial and Temporal Aspects in Video Editing Detectron2 object detection & manipulating images using cartooniza- tion
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 653c8f41-7a89-48b1-baa2-0b89e24d5a83 · outbound
SST-EM: Advanced Metrics for Evaluating Semantic, Spatial and Temporal Aspects in Video Editing A deep learning framework for quality assessment and restoration in video endoscopy
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 942d282c-5328-4b35-a70c-365d22da9a9a · outbound
SST-EM: Advanced Metrics for Evaluating Semantic, Spatial and Temporal Aspects in Video Editing Text2live: Text-driven layered image and video editing
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 3aabbcc7-f49a-4d2d-b84d-8476f907a3d5 · outbound
SST-EM: Advanced Metrics for Evaluating Semantic, Spatial and Temporal Aspects in Video Editing PaliGemma: A versatile 3B VLM for transfer
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 128ee8d3-6fdb-4d6e-a6e0-f45f95cd55a7 · outbound
SST-EM: Advanced Metrics for Evaluating Semantic, Spatial and Temporal Aspects in Video Editing Streaming Video Diffusion: Online Video Editing with Diffusion Models
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 094c37a2-89ea-44f0-bf68-1918eebe2c47 · outbound
SST-EM: Advanced Metrics for Evaluating Semantic, Spatial and Temporal Aspects in Video Editing Control-A-Video: Controllable Text-to-Video Diffusion Models with Motion Prior and Reward Feedback Learning
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3ea98bb1-999f-4318-9c07-99ef3c32ea2e · outbound
SST-EM: Advanced Metrics for Evaluating Semantic, Spatial and Temporal Aspects in Video Editing EditBoard: Towards a Comprehensive Evaluation Benchmark for Text-Based Video Editing Models
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 4f839925-9cd3-4f87-a370-b4f8dd1f9097 · outbound
SST-EM: Advanced Metrics for Evaluating Semantic, Spatial and Temporal Aspects in Video Editing Clip-adapter: Better vision-language models with fea- ture adapters
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 77dce725-1a90-42c1-8461-636a555e1769 · outbound
SST-EM: Advanced Metrics for Evaluating Semantic, Spatial and Temporal Aspects in Video Editing TokenFlow: Consistent Diffusion Features for Consistent Video Editing
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8e1e698a-bd57-4bc2-8507-4990874be4c2 · outbound
SST-EM: Advanced Metrics for Evaluating Semantic, Spatial and Temporal Aspects in Video Editing Enhancing the video editing capabilities of text-to-video generators using ddpm inversion
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation c0f74ddc-28c9-4946-aa9e-8a805ae7567c · outbound
SST-EM: Advanced Metrics for Evaluating Semantic, Spatial and Temporal Aspects in Video Editing Text-based Talking Video Editing with Cascaded Conditional Diffusion
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 7ee745d7-dfcc-4819-893c-bed2e4b7139e · outbound
SST-EM: Advanced Metrics for Evaluating Semantic, Spatial and Temporal Aspects in Video Editing ultralytics/yolov5: v6
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 521c2502-1795-4146-ad13-6ade0265f007 · outbound
SST-EM: Advanced Metrics for Evaluating Semantic, Spatial and Temporal Aspects in Video Editing Vilt: Vision- and-language transformer without convolution or region su- pervision
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation c7a53771-99f1-4548-a29b-6de66bc4ec26 · outbound
SST-EM: Advanced Metrics for Evaluating Semantic, Spatial and Temporal Aspects in Video Editing Segment any- thing
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 47f2576a-b26c-4b9b-ac2a-11040c167a98 · outbound
SST-EM: Advanced Metrics for Evaluating Semantic, Spatial and Temporal Aspects in Video Editing AnyV2V: A Tuning-Free Framework For Any Video-to-Video Editing Tasks
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 79454fb2-a119-498d-90c7-57f9f6afec31 · outbound
SST-EM: Advanced Metrics for Evaluating Semantic, Spatial and Temporal Aspects in Video Editing Shape-aware text-driven lay- ered video editing
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation beca871b-6f7f-4781-9cc8-7c2bc8a4c4a8 · outbound
SST-EM: Advanced Metrics for Evaluating Semantic, Spatial and Temporal Aspects in Video Editing Blip: Bootstrapping language-image pre-training for unified vision-language understanding and generation
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 20cc6552-8158-4efe-a95c-7fd4a5cce111 · outbound
SST-EM: Advanced Metrics for Evaluating Semantic, Spatial and Temporal Aspects in Video Editing Vidtome: Video token merging for zero-shot video editing
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation ae2c48f6-4d0a-431f-8163-7adc9f0db45b · outbound
SST-EM: Advanced Metrics for Evaluating Semantic, Spatial and Temporal Aspects in Video Editing Grounding dino: Marrying dino with grounded pre-training for open-set object detection
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 2aae9f3b-ea1c-4fbd-80a5-d965136920b3 · outbound
SST-EM: Advanced Metrics for Evaluating Semantic, Spatial and Temporal Aspects in Video Editing Video-p2p: Video editing with cross-attention control
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 7818f995-9564-4e6d-b8b9-1c6be52192c9 · outbound
SST-EM: Advanced Metrics for Evaluating Semantic, Spatial and Temporal Aspects in Video Editing TrailBlazer: Trajectory Control for Diffusion-Based Video Generation
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ee6f2988-171e-4fe0-9bea-86a3cafb8b7a · outbound
SST-EM: Advanced Metrics for Evaluating Semantic, Spatial and Temporal Aspects in Video Editing Gazed–gaze-guided cinematic editing of wide-angle monocular video recordings
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation ad541a63-3434-4ea7-a4c6-be9aabd57398 · outbound
SST-EM: Advanced Metrics for Evaluating Semantic, Spatial and Temporal Aspects in Video Editing Fatezero: Fus- ing attentions for zero-shot text-based video editing
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 2540c33d-c6d7-4004-afa9-625c574a117f · outbound
SST-EM: Advanced Metrics for Evaluating Semantic, Spatial and Temporal Aspects in Video Editing Enhanced end-to-end video editing: Adaptive customization of path, object, and motion dynam- ics
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 1b9d67d4-c46f-4683-805a-662d556dcbc5 · outbound
SST-EM: Advanced Metrics for Evaluating Semantic, Spatial and Temporal Aspects in Video Editing Pro- cedural crowd generation for semantically augmented virtual cities
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation c3c4c8ca-a9d7-4fce-b741-b6e35ebebe80 · outbound
SST-EM: Advanced Metrics for Evaluating Semantic, Spatial and Temporal Aspects in Video Editing Yolov7: Trainable bag-of-freebies sets new state-of-the-art for real-time object detectors
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bd7b1a1b-5fee-4b89-82a7-80d80ec6358f · outbound
SST-EM: Advanced Metrics for Evaluating Semantic, Spatial and Temporal Aspects in Video Editing Actionclip: Adapting language-image pretrained models for video action recognition
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 8ad1278e-5828-4f00-ad50-a55d8667cae8 · outbound
SST-EM: Advanced Metrics for Evaluating Semantic, Spatial and Temporal Aspects in Video Editing MedCLIP: Contrastive Learning from Unpaired Medical Images and Text
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c9d04de6-7bf2-4d94-9cac-776af417d893 · outbound
SST-EM: Advanced Metrics for Evaluating Semantic, Spatial and Temporal Aspects in Video Editing Tune-a-video: One-shot tuning of image diffusion models for text-to-video generation
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 93072393-8846-4e5d-b095-09bb36a68b05 · outbound
SST-EM: Advanced Metrics for Evaluating Semantic, Spatial and Temporal Aspects in Video Editing Florence-2: Advancing a unified representation for a variety of vision tasks
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 0e1a16af-b1a6-41a4-9fec-753fb24a3d02 · outbound
SST-EM: Advanced Metrics for Evaluating Semantic, Spatial and Temporal Aspects in Video Editing Temporally consistent semantic video editing
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 0f59556f-a314-4459-9846-2604fd3f692a · outbound
SST-EM: Advanced Metrics for Evaluating Semantic, Spatial and Temporal Aspects in Video Editing Context-aware talking-head video editing
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 72ac9507-7d40-4dc9-aa4a-0c4538349109 · outbound
SST-EM: Advanced Metrics for Evaluating Semantic, Spatial and Temporal Aspects in Video Editing Rethinking Human Evaluation Protocol for Text-to-Video Models: Enhancing Reliability,Reproducibility, and Practicality
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation d4388d50-4b4b-4850-94d6-cf47385e8094 · outbound
SST-EM: Advanced Metrics for Evaluating Semantic, Spatial and Temporal Aspects in Video Editing Motiondirector: Motion customization of text-to-video diffusion models
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation ae470dc1-653f-435b-b9e0-1552428baf80 · outbound
SST-EM: Advanced Metrics for Evaluating Semantic, Spatial and Temporal Aspects in Video Editing Deformable DETR: Deformable Transformers for End-to-End Object Detection
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.