Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-03T03:15:11.601779Z
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 9 of 9 outbound references and 4 inbound Pith citation observations for arXiv:2602.08711.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-03T03:15:11.601779Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-01T10:02:03.266407Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-07-04T17:40:00.898211Z
9 of 9 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 4c776e3f-1171-4350-aae9-9bc360072212 · outbound
TimeChat-Captioner: Scripting Multi-Scene Videos with Time-Aware and Structural Audio-Visual Captions • Each minute usually contains4–5 segments, but prioritize the video’s logic over strict numbers
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c7a12329-78df-4c02-a3f0-2b912ae77545 · outbound
TimeChat-Captioner: Scripting Multi-Scene Videos with Time-Aware and Structural Audio-Visual Captions • Includecharacters, actions, objects, emotions, and scene detailswhere relevant
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ebf37857-5121-4598-a9ae-5277186fafd9 · outbound
TimeChat-Captioner: Scripting Multi-Scene Videos with Time-Aware and Structural Audio-Visual Captions • Timestamp format:minutes:seconds(e.g.,0:00 - 0:11)
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ed901d18-4d82-40c8-a04b-510d88aaa265 · outbound
TimeChat-Captioner: Scripting Multi-Scene Videos with Time-Aware and Structural Audio-Visual Captions Each GT caption defines the rough boundaries of a segment
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c71922d9-2cfc-4bca-b50e-9b56dc80e979 · outbound
TimeChat-Captioner: Scripting Multi-Scene Videos with Time-Aware and Structural Audio-Visual Captions Use the video itself toexpand with details: •Characters: actions, gestures, facial expressions, emotions
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 163a77b1-b06a-42c1-9fe1-1480feb004d2 · outbound
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a576c31c-d18f-463f-b3fd-044dad84c7e8 · outbound
TimeChat-Captioner: Scripting Multi-Scene Videos with Time-Aware and Structural Audio-Visual Captions Unresolved cited work
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8bbc8a29-9465-4f13-b4cf-9b98fa517a21 · outbound
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 95218483-f9a3-4ebb-b2e0-62ea836414ee · outbound
TimeChat-Captioner: Scripting Multi-Scene Videos with Time-Aware and Structural Audio-Visual Captions HumanOmniV2: From Understanding to Omni-Modal Reasoning with Context
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 52056ccb-249e-40b0-a81d-75779896f93b · inbound
OmniScript: Towards Audio-Visual Script Generation for Long-Form Cinematic Video TimeChat-Captioner: Scripting Multi-Scene Videos with Time-Aware and Structural Audio-Visual Captions
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 0a8d3301-f527-4e13-bd04-03d87c9e37d8 · inbound
DiffCap-Bench: A Comprehensive, Challenging, Robust Benchmark for Image Difference Captioning TimeChat-Captioner: Scripting Multi-Scene Videos with Time-Aware and Structural Audio-Visual Captions
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ae5e7b81-915d-4acf-ac1f-cc23263dbccd · inbound
CineCap: Structured Reasoning with Spatio-Temporal Anchors for Cinematographic Video Captioning TimeChat-Captioner: Scripting Multi-Scene Videos with Time-Aware and Structural Audio-Visual Captions
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 7deff422-86ba-4f42-a199-15919089ae58 · inbound
PercepCap: Video Captioner with Structured Spatio-Temporal Perception TimeChat-Captioner: Scripting Multi-Scene Videos with Time-Aware and Structural Audio-Visual Captions
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.