Pith. sign in

Paper Citation Record · LEDGER

MAGVIT: Masked Generative Video Transformer

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 7 inbound Pith citation observations for arXiv:2212.05199.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2212.05199 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 7 of 7 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 7 of 7 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T11:54:57.347834Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation b8ac0df3-9642-4073-aa10-fa8153dde81c · inbound

FAST: Efficient Action Tokenization for Vision-Language-Action Models cites this paper.

FAST: Efficient Action Tokenization for Vision-Language-Action Models MAGVIT: Masked Generative Video Transformer

Reference 67

Resolution
verified exact
arxiv_id, observed 2026-05-11T08:52:32.112802Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-11T08:52:31.686474Z digest=sha256:3c6a8e1b6af2e9aaf359619a8a59fc6b3d167e036b85bf2052de56ea29c80cae

Observation 82a3eb92-7055-4bd8-9543-6ce37290fdde · inbound

Humanoid World Models: Open World Foundation Models for Humanoid Robotics cites this paper.

Humanoid World Models: Open World Foundation Models for Humanoid Robotics MAGVIT: Masked Generative Video Transformer

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-07T11:54:57.347834Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:54:57.347834Z digest=sha256:64c76467eb9425cff79f6f28fa7f3735b306d69f3652e216e17619ba7d715d51

Observation 9dda5626-d89f-474a-92d8-7620a5c62f59 · inbound

Infinite Video Understanding cites this paper.

Infinite Video Understanding MAGVIT: Masked Generative Video Transformer

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-06T18:10:14.507734Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:10:14.507734Z digest=sha256:74fd391c9cd659031b669f83bac439d80b9c51a08822bb2fefcf8c51e53cb42c

Observation 144296a1-b1b6-42ab-a870-9fc2f61ec084 · inbound

Can Your Model Separate Yolks with a Water Bottle? Benchmarking Physical Commonsense Understanding in Video Generation Models cites this paper.

Can Your Model Separate Yolks with a Water Bottle? Benchmarking Physical Commonsense Understanding in Video Generation Models MAGVIT: Masked Generative Video Transformer

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-06T15:27:03.235158Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:27:03.235158Z digest=sha256:5bc9007d5acb50d12ac6cb03a598df3d4f26a3b9f8f15dfce540b4f6932b6d72

Observation 6321bf00-d4cd-4399-b8f1-2cfa6976d55c · inbound

World Action Models: The Next Frontier in Embodied AI cites this paper.

World Action Models: The Next Frontier in Embodied AI MAGVIT: Masked Generative Video Transformer

Reference 288

Resolution
verified exact
arxiv_id, observed 2026-05-13T05:07:18.096556Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-13T05:01:16.802019Z digest=sha256:5f27c4991439d9b8442d91673ae9c5b488a65a2acd70418be9416285552237b4

Observation 8b4f0df9-7aac-41d6-a8f9-6183e02d0ad2 · inbound

Wavelet as Tokenizer: Preliminary Results on a Shared Wavelet Token Schema for Natural Signals cites this paper.

Wavelet as Tokenizer: Preliminary Results on a Shared Wavelet Token Schema for Natural Signals MAGVIT: Masked Generative Video Transformer

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-07-01T20:46:13.363676Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-28T18:05:11.392504Z digest=sha256:649f2921a02c070c1190b496a180a59efe31ee83f316d8f3c7c7034258b19815

Observation a2d94116-368c-42cb-8088-809127e1ff68 · inbound

BioVid: Autoregressive Video Generation with Biological Behavior Semantic Comprehension cites this paper.

BioVid: Autoregressive Video Generation with Biological Behavior Semantic Comprehension MAGVIT: Masked Generative Video Transformer

Reference 10

Resolution
metadata mismatch
arxiv_id, observed 2026-06-27T18:41:08.034995Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-27T18:35:14.331659Z digest=sha256:b82ed02a20731e9f74770529e6dafb211cc721aa2a0292ad441a73bec09384b6