Pith. sign in

Paper Citation Record · LEDGER

LARP: Tokenizing Videos with a Learned Autoregressive Generative Prior

As of 21 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 13 inbound Pith citation observations for arXiv:2410.21264.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2410.21264 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 13 of 13 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00

measured 13 of 13 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-15T20:38:30.329076Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T20:28:55.548089Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 5e73021d-f44d-4ac8-8af3-0535260537c7 · inbound

Efficient Long Video Tokenization via Coordinate-based Patch Reconstruction cites this paper.

Efficient Long Video Tokenization via Coordinate-based Patch Reconstruction LARP: Tokenizing Videos with a Learned Autoregressive Generative Prior

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-12T15:00:31.779652Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:00:31.779652Z digest=sha256:978ac86551bb3f872f985fc08c6158a634c98f8e1e20c2a0d325fee7e4540462

Observation 8fabbf4d-909f-490c-900e-0d38893bd01c · inbound

Divot: Diffusion Powers Video Tokenizer for Comprehension and Generation cites this paper.

Divot: Diffusion Powers Video Tokenizer for Comprehension and Generation LARP: Tokenizing Videos with a Learned Autoregressive Generative Prior

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-11T21:30:08.494766Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T21:30:08.494766Z digest=sha256:198ba71e858c69000b67a20b2d63f5f23bb6d682b5dffe056a900e73991e950b

Observation 40f906bf-7435-48b5-99db-ea375161fe36 · inbound

SweetTok: Semantic-Aware Spatial-Temporal Tokenizer for Compact Video Discretization cites this paper.

SweetTok: Semantic-Aware Spatial-Temporal Tokenizer for Compact Video Discretization LARP: Tokenizing Videos with a Learned Autoregressive Generative Prior

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-11T17:59:36.661509Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T17:59:36.661509Z digest=sha256:2fd46369f79a898bcf11a1a530d62cd41cad16f73aa2e18cf1ebb3e48662c91a

Observation 33d2fd40-5049-4518-a1e4-e766afda89cb · inbound

Learnings from Scaling Visual Tokenizers for Reconstruction and Generation cites this paper.

Learnings from Scaling Visual Tokenizers for Reconstruction and Generation LARP: Tokenizing Videos with a Learned Autoregressive Generative Prior

Reference 2008

Resolution
unresolved
no resolver link, observed 2026-08-10T19:48:08.219120Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T19:48:08.219120Z digest=sha256:fa52fafad6d479cc8c8ef977c2cc1ea3b2b8881d9dd741a217a2585927eb0efa

Observation 950114c0-1f65-481b-82ab-c6ec39114caf · inbound

Video-GPT via Next Clip Diffusion cites this paper.

Video-GPT via Next Clip Diffusion LARP: Tokenizing Videos with a Learned Autoregressive Generative Prior

Reference 75

Resolution
unresolved
no resolver link, observed 2026-08-15T20:38:30.329076Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:38:30.329076Z digest=sha256:628dd321b4701808fb76d56786a58e00ccc16dababcec3e71868a936f571448e

Observation 4bbb3b56-c0b8-4258-bde4-a770d0ebed82 · inbound

Vision as a Dialect: Unifying Visual Understanding and Generation via Text-Aligned Representations cites this paper.

Vision as a Dialect: Unifying Visual Understanding and Generation via Text-Aligned Representations LARP: Tokenizing Videos with a Learned Autoregressive Generative Prior

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-15T18:46:10.844791Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T18:46:10.844791Z digest=sha256:6058798fd09abe92bbfc8f428229fe12f70177961357b08bab8ddc245ff0f912

Observation 5d22831f-0601-4971-85ce-c9510b584922 · inbound

KeyframeFace: Language-Driven Facial Animation via Semantic Keyframes cites this paper.

KeyframeFace: Language-Driven Facial Animation via Semantic Keyframes LARP: Tokenizing Videos with a Learned Autoregressive Generative Prior

Reference 34

Resolution
verified exact
arxiv_id, observed 2026-05-16T22:48:38.138963Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-16T22:46:51.954117Z digest=sha256:e42fdea0c1a760746366e86532db78321b6b64ec439639f92caf3be345c25501

Observation 279410c1-ee40-4d46-97ca-00152d5f588c · inbound

Autoregressive Visual Generation Needs a Prologue cites this paper.

Autoregressive Visual Generation Needs a Prologue LARP: Tokenizing Videos with a Learned Autoregressive Generative Prior

Reference 47

Resolution
verified exact
arxiv_id, observed 2026-05-11T18:51:06.123795Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-08T13:52:42.834440Z digest=sha256:120e80d7ad63ffc1e3ff4c0414e9b94ceac643fe89712da408f29ccdd1a1fae9

Observation 9397069a-2688-4461-b4d7-2bc2f727b23d · inbound

Autoregressive Visual Generation Needs a Prologue cites this paper.

Autoregressive Visual Generation Needs a Prologue LARP: Tokenizing Videos with a Learned Autoregressive Generative Prior

Reference 47

Resolution
verified exact
arxiv_id, observed 2026-07-01T13:15:46.137396Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-30T23:36:51.148162Z digest=sha256:3215e4f1c3b5768128179c649f0a8fe22981b6944a18fa9013d08a5cdb367320

Observation b939ef1f-3d11-4877-814d-3e94c29ee631 · inbound

Diffusing in the Right Space: A Systematic Study of Latent Diffusability cites this paper.

Diffusing in the Right Space: A Systematic Study of Latent Diffusability LARP: Tokenizing Videos with a Learned Autoregressive Generative Prior

Reference 39

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T02:46:27.641875Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-06-28T10:44:24.318786Z digest=sha256:140abecbb0f12d6e8805940ac5b5e07c4771913df78978e4bf44d2ddb5f76d62

Observation 487082ec-3b6a-4076-8152-4ed14b3ca30d · inbound

Adaptive Tokenisation Via Temporal Redundancy Masking And Latent Inpainting cites this paper.

Adaptive Tokenisation Via Temporal Redundancy Masking And Latent Inpainting LARP: Tokenizing Videos with a Learned Autoregressive Generative Prior

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-07-02T11:46:56.089570Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-28T02:55:00.586061Z digest=sha256:65b1e6cff6f0279b5086403700d874df991c658f9b3b1592371a62b4b70fb485

Observation f46ea481-0bc4-47e3-a2a2-407cecf4b184 · inbound

TivTok: Broadcasting Time-Invariant Tokens for Scalable Video Tokenization cites this paper.

TivTok: Broadcasting Time-Invariant Tokens for Scalable Video Tokenization LARP: Tokenizing Videos with a Learned Autoregressive Generative Prior

Reference 92

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T20:28:55.549434Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-06-27T01:20:32.508409Z digest=sha256:a42f9a63d7e7195e110eaa0791b20bc251411a965414d7274cf8a0249fd24c38

Observation 6385974f-1d7d-427e-9122-1a17fa270dae · inbound

VideoRAE: Taming Video Foundation Models for Generative Modeling via Representation Autoencoders cites this paper.

VideoRAE: Taming Video Foundation Models for Generative Modeling via Representation Autoencoders LARP: Tokenizing Videos with a Learned Autoregressive Generative Prior

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-02T02:51:46.641264Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T02:51:46.641264Z digest=sha256:666fbf006293a5f435661245ca3740b588b16482b8cfa45db1e0c866e6acee27