Pith. sign in

Paper Citation Record · LEDGER

MarDini: Masked Autoregressive Diffusion for Video Generation at Scale

As of 18 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 23 inbound Pith citation observations for arXiv:2410.20280.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2410.20280 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 23 of 23 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 23 of 23 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T05:17:50.337647Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-01T19:16:00.498245Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 5f009bd5-a877-460a-884d-dbbf479fee0f · inbound

Towards Precise Scaling Laws for Video Diffusion Transformers cites this paper.

Towards Precise Scaling Laws for Video Diffusion Transformers MarDini: Masked Autoregressive Diffusion for Video Generation at Scale

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-12T12:58:57.210336Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:58:57.210336Z digest=sha256:04b4024029ef0bfaae541ce0a4e0faeff88f8f903f35a1390a4baa81154077c3

Observation bab0be1d-6279-4510-8f4c-9b040b3495a8 · inbound

Continuous Speech Tokens Makes LLMs Robust Multi-Modality Learners cites this paper.

Continuous Speech Tokens Makes LLMs Robust Multi-Modality Learners MarDini: Masked Autoregressive Diffusion for Video Generation at Scale

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-11T21:12:24.041379Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T21:12:24.041379Z digest=sha256:c6bf085279725770fa43a1eea3f0a83c09055811ea2c32d35c915d5b3f75bdc7

Observation 8613954c-0e30-40fc-ba28-3d48f7efcca2 · inbound

Video Diffusion Transformers are In-Context Learners cites this paper.

Video Diffusion Transformers are In-Context Learners MarDini: Masked Autoregressive Diffusion for Video Generation at Scale

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-11T15:40:19.908050Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:40:19.908050Z digest=sha256:8a49dea21df73ccbe5ebcad3c6dfbb9ec2a2c93df8d73c324dcac4983a155558

Observation 8ca0697e-ed35-4eb9-9ec6-684dbe8fc8e7 · inbound

Ingredients: Blending Custom Photos with Video Diffusion Transformers cites this paper.

Ingredients: Blending Custom Photos with Video Diffusion Transformers MarDini: Masked Autoregressive Diffusion for Video Generation at Scale

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-10T22:25:25.708376Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:25:25.708376Z digest=sha256:12d00884efb5e3a3eb15def57dd5268948dff400c24c7da91f3f1b2017b07a91

Observation ac4003c4-34fb-4fda-b546-b44da45e8074 · inbound

Cosmos World Foundation Model Platform for Physical AI cites this paper.

Cosmos World Foundation Model Platform for Physical AI MarDini: Masked Autoregressive Diffusion for Video Generation at Scale

Reference 124

Resolution
verified exact
arxiv_id, observed 2026-05-10T23:38:45.698349Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-10T23:38:44.933410Z digest=sha256:c427d837240feb00e61a8d3f7c5c252c4cadd77aad21b6fc54f57baea7091cf6

Observation 46b0d591-b1c3-4e32-84f2-33358c34b27f · inbound

CAT Pruning: Cluster-Aware Token Pruning For Text-to-Image Diffusion Models cites this paper.

CAT Pruning: Cluster-Aware Token Pruning For Text-to-Image Diffusion Models MarDini: Masked Autoregressive Diffusion for Video Generation at Scale

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-09T19:05:52.909908Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T19:05:52.909908Z digest=sha256:01ebde8b4055de55218413435ea1d3b1a0e414bacf29f1ff1957dbb2ba640555

Observation 64c95c26-a4bf-4210-9e5b-e8b1d7145cfe · inbound

Learning Real-World Action-Video Dynamics with Heterogeneous Masked Autoregression cites this paper.

Learning Real-World Action-Video Dynamics with Heterogeneous Masked Autoregression MarDini: Masked Autoregressive Diffusion for Video Generation at Scale

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-08T22:55:57.775016Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T22:55:57.775016Z digest=sha256:40309a6d25b11eb6b2cacfd0f465973e80357cac7286d64cdd423364d2b221f4

Observation 02da69e5-797e-4332-b0f8-745f0591f461 · inbound

Capturing Conditional Dependence via Auto-regressive Diffusion Models cites this paper.

Capturing Conditional Dependence via Auto-regressive Diffusion Models MarDini: Masked Autoregressive Diffusion for Video Generation at Scale

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-16T05:17:50.337647Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T05:17:50.337647Z digest=sha256:a4e6e118e3cc111cadd087ec57605cfcd183ab27b57fd49c80fce911cdbb0eae

Observation 8b3ecda6-e615-44b9-9f25-9e7273423986 · inbound

MARRS: Masked Autoregressive Unit-based Reaction Synthesis cites this paper.

MARRS: Masked Autoregressive Unit-based Reaction Synthesis MarDini: Masked Autoregressive Diffusion for Video Generation at Scale

Reference 37

Resolution
verified exact
arxiv_id, observed 2026-05-22T14:51:42.644310Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-22T14:48:35.211157Z digest=sha256:6a818b6949c03c5b7ca5cbb71c70d142118922b9627448acca4d991926699cbd

Observation 141567d3-d732-4661-8fad-0912920fecb9 · inbound

Controllable Coupled Image Generation via Diffusion Models cites this paper.

Controllable Coupled Image Generation via Diffusion Models MarDini: Masked Autoregressive Diffusion for Video Generation at Scale

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T05:54:41.695472Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:54:41.695472Z digest=sha256:d564ad2b0c0048a0b3dc2b91401539bfc86e72fc40762bc9852d48502b038c27

Observation 2338f9ea-a91d-416f-9df8-7d2b94a55ea3 · inbound

Self Forcing: Bridging the Train-Test Gap in Autoregressive Video Diffusion cites this paper.

Self Forcing: Bridging the Train-Test Gap in Autoregressive Video Diffusion MarDini: Masked Autoregressive Diffusion for Video Generation at Scale

Reference 50

Resolution
verified exact
arxiv_id, observed 2026-05-11T01:36:53.310698Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-11T01:36:53.029590Z digest=sha256:d845a613af09f4cefa1773977fbecb9f2b12c1316511fa0a2877f80a54abf6cd

Observation 46ed2823-dacb-472d-84b9-81e9823608e1 · inbound

Show-o2: Improved Native Unified Multimodal Models cites this paper.

Show-o2: Improved Native Unified Multimodal Models MarDini: Masked Autoregressive Diffusion for Video Generation at Scale

Reference 73

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T18:51:16.056458Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-12T18:51:15.428692Z digest=sha256:8c2fd5cf00716035969cacec2d1d50fff349aea2aae521fbdbcd12fe04f9dac9

Observation f92cfa37-e3ce-47c6-8180-4db4227a202c · inbound

Geometry-aware 4D Video Generation for Robot Manipulation cites this paper.

Geometry-aware 4D Video Generation for Robot Manipulation MarDini: Masked Autoregressive Diffusion for Video Generation at Scale

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-22T00:14:27.886735Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-22T00:12:39.787489Z digest=sha256:a6cf651a7359c49624802f82390d84d9b4a3e5f7808fcb1163becfeac52ba146

Observation 7c8cce81-4771-4331-91bc-6ce333b3473f · inbound

Rolling Forcing: Autoregressive Long Video Diffusion in Real Time cites this paper.

Rolling Forcing: Autoregressive Long Video Diffusion in Real Time MarDini: Masked Autoregressive Diffusion for Video Generation at Scale

Reference 81

Resolution
verified exact
arxiv_id, observed 2026-05-16T11:15:29.166529Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-16T11:15:29.102090Z digest=sha256:a408763c910d50f739ce604434537ae2fcc6d983b1f273b830364567596d3f2c

Observation acf7b787-2bc1-42c4-8d67-1c71a9893f3a · inbound

End-to-End Training for Autoregressive Video Diffusion via Self-Resampling cites this paper.

End-to-End Training for Autoregressive Video Diffusion via Self-Resampling MarDini: Masked Autoregressive Diffusion for Video Generation at Scale

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-03T15:46:09.704906Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T15:46:09.704906Z digest=sha256:ff3a1a20cf4b39efaf4b38040cd75d9ef509becb1488569e27a90873f7cd1f0e

Observation f2da29bc-6b10-4924-bd66-25a2ca30ed61 · inbound

Stream-DiffVSR: Low-Latency Streamable Video Super-Resolution via Auto-Regressive Diffusion cites this paper.

Stream-DiffVSR: Low-Latency Streamable Video Super-Resolution via Auto-Regressive Diffusion MarDini: Masked Autoregressive Diffusion for Video Generation at Scale

Reference 39

Resolution
verified exact
arxiv_id, observed 2026-05-16T19:18:19.588304Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-16T19:14:15.240218Z digest=sha256:bc7dc307d559d114fefc2d2737ae60083d07ea286ee229e830406af4b512d6af

Observation df3c1475-da9c-4580-a188-f18f9f80fc9d · inbound

Rolling Sink: Bridging Limited-Horizon Training and Open-Ended Testing in Autoregressive Video Diffusion cites this paper.

Rolling Sink: Bridging Limited-Horizon Training and Open-Ended Testing in Autoregressive Video Diffusion MarDini: Masked Autoregressive Diffusion for Video Generation at Scale

Reference 62

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T07:07:29.882982Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-16T07:02:38.876518Z digest=sha256:600e6c7766559349208fe457c22cc4924523ca7f5014d845e9c77d4335f1b310

Observation a3ef76b0-435a-4d36-a9de-05a1dfcc15b9 · inbound

INSPATIO-WORLD: A Real-Time 4D World Simulator via Spatiotemporal Autoregressive Modeling cites this paper.

INSPATIO-WORLD: A Real-Time 4D World Simulator via Spatiotemporal Autoregressive Modeling MarDini: Masked Autoregressive Diffusion for Video Generation at Scale

Reference 55

Resolution
verified exact
arxiv_id, observed 2026-05-10T23:25:53.352993Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-10T19:08:56.588282Z digest=sha256:5970679bee6fa4399ade0ecb1176b2d6bbfca7c69c2a8af09548fa2b894f6d46

Observation b30dd2ac-47dc-404d-af0f-14835633f96d · inbound

Image-to-Video Diffusion: From Foundations to Open Frontiers cites this paper.

Image-to-Video Diffusion: From Foundations to Open Frontiers MarDini: Masked Autoregressive Diffusion for Video Generation at Scale

Reference 48

Resolution
verified exact
arxiv_id, observed 2026-05-20T15:08:25.020789Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-20T15:06:02.084336Z digest=sha256:65fd2f3b305432b39329488f5796ad0158d0ddde1b24812adb4ccd9036d328bc

Observation c68a26c1-c320-4993-a9e9-7abdbe70360f · inbound

Lance: Unified Multimodal Modeling by Multi-Task Synergy cites this paper.

Lance: Unified Multimodal Modeling by Multi-Task Synergy MarDini: Masked Autoregressive Diffusion for Video Generation at Scale

Reference 73

Resolution
verified exact
arxiv_id, observed 2026-05-20T11:48:14.943518Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-20T11:46:52.658984Z digest=sha256:fcf08da330bfe42bf81695605354091234f58ef658dd4b01242a40b1b569c760

Observation 534419d3-5744-4d85-99f0-0f81aa9bbb25 · inbound

Lance: Unified Multimodal Modeling by Multi-Task Synergy cites this paper.

Lance: Unified Multimodal Modeling by Multi-Task Synergy MarDini: Masked Autoregressive Diffusion for Video Generation at Scale

Reference 74

Resolution
verified exact
arxiv_id, observed 2026-05-21T07:59:50.644203Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-21T07:56:34.034047Z digest=sha256:35ebcabd6b6642035829123b9625cbda3cc8ef0c9a353cd7776a4fbea66fda27

Observation 0aa31732-9ad9-48e7-bdf1-2a0eeeff900c · inbound

Lumos-Nexus: Efficient Frequency Bridging with Homogeneous Latent Space for Video Unified Models cites this paper.

Lumos-Nexus: Efficient Frequency Bridging with Homogeneous Latent Space for Video Unified Models MarDini: Masked Autoregressive Diffusion for Video Generation at Scale

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-07-01T19:16:00.500488Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-28T22:56:21.783415Z digest=sha256:1bf1e46df3307a9566c076db39c90cb8a8fb5121eaa745b342aface6f98ab375

Observation 3dfe1268-1b4d-4e56-8308-e43fecba378b · inbound

Stream Forcing: Constructing Unified Training Trajectory for Robust Streaming Video Generation cites this paper.

Stream Forcing: Constructing Unified Training Trajectory for Robust Streaming Video Generation MarDini: Masked Autoregressive Diffusion for Video Generation at Scale

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-15T14:26:02.527602Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T14:26:02.527602Z digest=sha256:78ba7a6ed9e27b519fb829d5d69ac3aa2d1809c1305a61b88944a6d59d2d4b37