Pith. sign in

Paper Citation Record · LEDGER

Long Video Generation with Time-Agnostic VQGAN and Time-Sensitive Transformer

As of 11 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 11 inbound Pith citation observations for arXiv:2204.03638.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2204.03638 v4

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 11 of 11 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 11 of 11 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-10T19:48:08.135556Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T05:27:40.599122Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 67597821-2c1c-4a13-bc11-49e7516a0a3f · inbound

CogVideo: Large-scale Pretraining for Text-to-Video Generation via Transformers cites this paper.

CogVideo: Large-scale Pretraining for Text-to-Video Generation via Transformers Long Video Generation with Time-Agnostic VQGAN and Time-Sensitive Transformer

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-11T12:24:30.181165Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-11T12:24:30.071822Z digest=sha256:4b160ddae48cec8d74030f85b970db6835bde8445ff0db260df1d26a93303405

Observation bf101d3b-eb4d-475f-845f-be1e5a5b6298 · inbound

Latent Video Diffusion Models for High-Fidelity Long Video Generation cites this paper.

Latent Video Diffusion Models for High-Fidelity Long Video Generation Long Video Generation with Time-Agnostic VQGAN and Time-Sensitive Transformer

Reference 6

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T04:27:43.184025Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-15T04:27:43.019940Z digest=sha256:58439e07110eea2ad5530a506dd7043ce63db4fc976d64ff8a477c78551ac78d

Observation ba484fe8-d9f4-41a7-aac1-6934ac3fe784 · inbound

DragNUWA: Fine-grained Control in Video Generation by Integrating Text, Image, and Trajectory cites this paper.

DragNUWA: Fine-grained Control in Video Generation by Integrating Text, Image, and Trajectory Long Video Generation with Time-Agnostic VQGAN and Time-Sensitive Transformer

Reference 257

Resolution
verified exact
arxiv_id, observed 2026-05-20T13:03:58.193602Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-20T13:03:57.828598Z digest=sha256:10015a63ebad921f305da3d57bdc02ea71c3bda79a04c710e28565bc63d10e84

Observation b843b63e-f2d6-4f91-82ec-6f8eff33911f · inbound

Learnings from Scaling Visual Tokenizers for Reconstruction and Generation cites this paper.

Learnings from Scaling Visual Tokenizers for Reconstruction and Generation Long Video Generation with Time-Agnostic VQGAN and Time-Sensitive Transformer

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-10T19:48:08.135556Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T19:48:08.135556Z digest=sha256:8ea72b94c4d0524a334398c290b238c1f58d46a641a3952e2a92eca30551aaf0

Observation 08a1bc57-c39a-47f3-943f-1c15b1c1d9e3 · inbound

Separate Motion from Appearance: Customizing Motion via Customizing Text-to-Video Diffusion Models cites this paper.

Separate Motion from Appearance: Customizing Motion via Customizing Text-to-Video Diffusion Models Long Video Generation with Time-Agnostic VQGAN and Time-Sensitive Transformer

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-10T11:17:24.751477Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T11:17:24.751477Z digest=sha256:429e00b4e43f45c984569d7eadc1999a3bb03fba7fc1ad74e6e2fb178f2cce76

Observation 63c1c70b-3ba1-4d43-a700-cf71dd1f3ea8 · inbound

Masked Generative Nested Transformers with Decode Time Scaling cites this paper.

Masked Generative Nested Transformers with Decode Time Scaling Long Video Generation with Time-Agnostic VQGAN and Time-Sensitive Transformer

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-09T19:21:56.208010Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T19:21:56.208010Z digest=sha256:0d7ecd7c46130aaccec0021d766cc1cc2f1f551c1ebd83a9a6d27c5ce72b7dd7

Observation 29ece65a-06f1-4bd6-829d-197a19a08a39 · inbound

ELT: Elastic Looped Transformers for Visual Generation cites this paper.

ELT: Elastic Looped Transformers for Visual Generation Long Video Generation with Time-Agnostic VQGAN and Time-Sensitive Transformer

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-05-11T07:06:00.145124Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T17:19:22.543462Z digest=sha256:1f7000f26d0635a9b2b937df8b360b1483a689dffa75e8b56635c2401337d38b

Observation 1ec01d9a-d528-431f-9ab0-a04bfdc268dc · inbound

ELT: Elastic Looped Transformers for Visual Generation cites this paper.

ELT: Elastic Looped Transformers for Visual Generation Long Video Generation with Time-Agnostic VQGAN and Time-Sensitive Transformer

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-02T16:35:03.177296Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T16:35:03.177296Z digest=sha256:ba60723ab2a8b1f632ce162e34ea1e083b2f88be8064e87b0e7fa700ea1a15ac

Observation f0612a31-cdc8-4585-8221-891bffb55bc5 · inbound

IDEAL: In-DEpth ALignment Makes A Discrete Representation AutoEncoder cites this paper.

IDEAL: In-DEpth ALignment Makes A Discrete Representation AutoEncoder Long Video Generation with Time-Agnostic VQGAN and Time-Sensitive Transformer

Reference 13

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T05:27:40.600866Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-27T13:10:14.308216Z digest=sha256:e3a40aa4c4b472169f79790336d83a79e6ca63d1d502d0a970d36b6d0bd5fd59

Observation 0f3d40c5-5f23-40b6-ad25-193b00edfbd8 · inbound

VideoRAE: Taming Video Foundation Models for Generative Modeling via Representation Autoencoders cites this paper.

VideoRAE: Taming Video Foundation Models for Generative Modeling via Representation Autoencoders Long Video Generation with Time-Agnostic VQGAN and Time-Sensitive Transformer

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-02T02:51:44.235751Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T02:51:44.235751Z digest=sha256:0b3105d4e6504321b574d9c75e489b9bf33e2f443cd833b984d4c68b716469d5

Observation 6662caa6-708a-44db-99c8-e9d3f65da8b8 · inbound

Reinforcement Learning: From Algorithms To Foundation Models cites this paper.

Reinforcement Learning: From Algorithms To Foundation Models Long Video Generation with Time-Agnostic VQGAN and Time-Sensitive Transformer

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-01T17:45:00.981903Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T17:45:00.981903Z digest=sha256:57007ef348d228920d010aa53ef4c3aaadb43221cd15f9695823a0c6131788a1