Pith. sign in

Paper Citation Record · LEDGER

LARP: Tokenizing Videos with a Learned Autoregressive Generative Prior

As of 21 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 13 inbound Pith citation observations for arXiv:2410.21264.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2410.21264 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 13 of 13 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00

measured 13 of 13 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-15T20:38:30.329076Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T20:28:55.548089Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 5e73021d-f44d-4ac8-8af3-0535260537c7 · inbound

Efficient Long Video Tokenization via Coordinate-based Patch Reconstruction cites this paper.

Efficient Long Video Tokenization via Coordinate-based Patch Reconstruction LARP: Tokenizing Videos with a Learned Autoregressive Generative Prior

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-12T15:00:31.779652Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:00:31.779652Z digest=sha256:978ac86551bb3f872f985fc08c6158a634c98f8e1e20c2a0d325fee7e4540462

Observation 8fabbf4d-909f-490c-900e-0d38893bd01c · inbound

Divot: Diffusion Powers Video Tokenizer for Comprehension and Generation cites this paper.

Divot: Diffusion Powers Video Tokenizer for Comprehension and Generation LARP: Tokenizing Videos with a Learned Autoregressive Generative Prior

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-11T21:30:08.494766Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T21:30:08.494766Z digest=sha256:198ba71e858c69000b67a20b2d63f5f23bb6d682b5dffe056a900e73991e950b

Observation 40f906bf-7435-48b5-99db-ea375161fe36 · inbound

SweetTok: Semantic-Aware Spatial-Temporal Tokenizer for Compact Video Discretization cites this paper.

SweetTok: Semantic-Aware Spatial-Temporal Tokenizer for Compact Video Discretization LARP: Tokenizing Videos with a Learned Autoregressive Generative Prior

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-11T17:59:36.661509Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T17:59:36.661509Z digest=sha256:2fd46369f79a898bcf11a1a530d62cd41cad16f73aa2e18cf1ebb3e48662c91a

Observation 33d2fd40-5049-4518-a1e4-e766afda89cb · inbound

Learnings from Scaling Visual Tokenizers for Reconstruction and Generation cites this paper.

Learnings from Scaling Visual Tokenizers for Reconstruction and Generation LARP: Tokenizing Videos with a Learned Autoregressive Generative Prior

Reference 2008

Resolution
unresolved
no resolver link, observed 2026-08-10T19:48:08.219120Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T19:48:08.219120Z digest=sha256:fa52fafad6d479cc8c8ef977c2cc1ea3b2b8881d9dd741a217a2585927eb0efa

Observation 950114c0-1f65-481b-82ab-c6ec39114caf · inbound

Video-GPT via Next Clip Diffusion cites this paper.

Video-GPT via Next Clip Diffusion LARP: Tokenizing Videos with a Learned Autoregressive Generative Prior

Reference 75

Resolution
unresolved
no resolver link, observed 2026-08-15T20:38:30.329076Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:38:30.329076Z digest=sha256:628dd321b4701808fb76d56786a58e00ccc16dababcec3e71868a936f571448e

Observation 4bbb3b56-c0b8-4258-bde4-a770d0ebed82 · inbound

Vision as a Dialect: Unifying Visual Understanding and Generation via Text-Aligned Representations cites this paper.

Vision as a Dialect: Unifying Visual Understanding and Generation via Text-Aligned Representations LARP: Tokenizing Videos with a Learned Autoregressive Generative Prior

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-15T18:46:10.844791Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T18:46:10.844791Z digest=sha256:6058798fd09abe92bbfc8f428229fe12f70177961357b08bab8ddc245ff0f912

Observation 5d22831f-0601-4971-85ce-c9510b584922 · inbound

KeyframeFace: Language-Driven Facial Animation via Semantic Keyframes cites this paper.

KeyframeFace: Language-Driven Facial Animation via Semantic Keyframes LARP: Tokenizing Videos with a Learned Autoregressive Generative Prior

Reference 34

Resolution
verified exact
arxiv_id, observed 2026-05-16T22:48:38.138963Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-16T22:46:51.954117Z digest=sha256:0110a3623b1c60e26deee276a94a0f3c8f57d3e5911bfb2aa0431ae0d6a7f8b6

Observation 279410c1-ee40-4d46-97ca-00152d5f588c · inbound

Autoregressive Visual Generation Needs a Prologue cites this paper.

Autoregressive Visual Generation Needs a Prologue LARP: Tokenizing Videos with a Learned Autoregressive Generative Prior

Reference 47

Resolution
verified exact
arxiv_id, observed 2026-05-11T18:51:06.123795Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-08T13:52:42.834440Z digest=sha256:403184282e67da8c079fa6444a91e5a6512dec48216f50c460b05b0e9a4bb07e

Observation 9397069a-2688-4461-b4d7-2bc2f727b23d · inbound

Autoregressive Visual Generation Needs a Prologue cites this paper.

Autoregressive Visual Generation Needs a Prologue LARP: Tokenizing Videos with a Learned Autoregressive Generative Prior

Reference 47

Resolution
verified exact
arxiv_id, observed 2026-07-01T13:15:46.137396Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-30T23:36:51.148162Z digest=sha256:6d18950f21a26e155d3bf1cf2df3e0ebf01eab2aa90a7071271eeb4ef162a2da

Observation b939ef1f-3d11-4877-814d-3e94c29ee631 · inbound

Diffusing in the Right Space: A Systematic Study of Latent Diffusability cites this paper.

Diffusing in the Right Space: A Systematic Study of Latent Diffusability LARP: Tokenizing Videos with a Learned Autoregressive Generative Prior

Reference 39

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T02:46:27.641875Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-06-28T10:44:24.318786Z digest=sha256:2379b9cbedafcc60d7c8845337eb24fc0f2654dbfab8646460ca111b5999112a

Observation 487082ec-3b6a-4076-8152-4ed14b3ca30d · inbound

Adaptive Tokenisation Via Temporal Redundancy Masking And Latent Inpainting cites this paper.

Adaptive Tokenisation Via Temporal Redundancy Masking And Latent Inpainting LARP: Tokenizing Videos with a Learned Autoregressive Generative Prior

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-07-02T11:46:56.089570Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-28T02:55:00.586061Z digest=sha256:de863a8fedc46358229d28d0af1a9752a950f36b4178bbe7bb35ab1fcc061544

Observation f46ea481-0bc4-47e3-a2a2-407cecf4b184 · inbound

TivTok: Broadcasting Time-Invariant Tokens for Scalable Video Tokenization cites this paper.

TivTok: Broadcasting Time-Invariant Tokens for Scalable Video Tokenization LARP: Tokenizing Videos with a Learned Autoregressive Generative Prior

Reference 92

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T20:28:55.549434Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-06-27T01:20:32.508409Z digest=sha256:1bae5ce670adecda8bd9366b0b9eaa4a9bd22c620afabd2bdc6639e83dbb79a5

Observation 6385974f-1d7d-427e-9122-1a17fa270dae · inbound

VideoRAE: Taming Video Foundation Models for Generative Modeling via Representation Autoencoders cites this paper.

VideoRAE: Taming Video Foundation Models for Generative Modeling via Representation Autoencoders LARP: Tokenizing Videos with a Learned Autoregressive Generative Prior

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-02T02:51:46.641264Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T02:51:46.641264Z digest=sha256:666fbf006293a5f435661245ca3740b588b16482b8cfa45db1e0c866e6acee27