Pith. sign in

Paper Citation Record · LEDGER

TALC: Time-Aligned Captions for Multi-Scene Text-to-Video Generation

As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 5 inbound Pith citation observations for arXiv:2405.04682.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2405.04682 v4

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 5 of 5 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 5 of 5 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T16:39:20.837113Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-20T11:34:37.806610Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 62631cea-690f-4cc0-87d4-eed86fd5b573 · inbound

VideoPhy: Evaluating Physical Commonsense for Video Generation cites this paper.

VideoPhy: Evaluating Physical Commonsense for Video Generation TALC: Time-Aligned Captions for Multi-Scene Text-to-Video Generation

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-20T11:34:37.808511Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-20T11:34:37.599691Z digest=sha256:7510a9cf6c746ed803debe53b4a405e2bdc1721bfa6b514515b54985cd6b90cd

Observation 0417c5fc-31fa-4223-9446-89fc25712067 · inbound

LoViC: Efficient Long Video Generation with Context Compression cites this paper.

LoViC: Efficient Long Video Generation with Context Compression TALC: Time-Aligned Captions for Multi-Scene Text-to-Video Generation

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T16:39:20.837113Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:39:20.837113Z digest=sha256:54575cad6ad3cd927f9287b6c88637f64d651d857049c2ef94325c4a5316e353

Observation b1ab58d7-6992-4fd4-ba12-1760b5b788ca · inbound

Rolling Forcing: Autoregressive Long Video Diffusion in Real Time cites this paper.

Rolling Forcing: Autoregressive Long Video Diffusion in Real Time TALC: Time-Aligned Captions for Multi-Scene Text-to-Video Generation

Reference 57

Resolution
verified exact
arxiv_id, observed 2026-05-16T11:15:29.260004Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-16T11:15:29.102090Z digest=sha256:2c2609f17b02ba13e6c66e64f5a4b894c501be130ac942ae7e9b545fc77b20db

Observation cdde0bd5-d641-4a8c-bfaa-0d05ea1d7fb5 · inbound

TS-Attn: Temporal-wise Separable Attention for Multi-Event Video Generation cites this paper.

TS-Attn: Temporal-wise Separable Attention for Multi-Event Video Generation TALC: Time-Aligned Captions for Multi-Scene Text-to-Video Generation

Reference 1

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T02:53:29.905801Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T02:45:10.577070Z digest=sha256:d79bffdab89dab00d8de432f1afb9c889482b3f53e055cad01ad8eac9aee0396

Observation 0c35e817-0dce-493e-a1cf-ddfd7bc8507c · inbound

CausalCine: Real-Time Autoregressive Generation for Multi-Shot Video Narratives cites this paper.

CausalCine: Real-Time Autoregressive Generation for Multi-Shot Video Narratives TALC: Time-Aligned Captions for Multi-Scene Text-to-Video Generation

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-05-13T05:42:21.395852Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T05:38:21.587653Z digest=sha256:66fc7099136886a06b48a004281d59a648e8853c194955c3fc68fb218a6f3caf