Pith. sign in

Paper Citation Record · LEDGER

MonoFormer: One Transformer for Both Diffusion and Autoregression

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 12 inbound Pith citation observations for arXiv:2409.16280.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2409.16280 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 12 of 12 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 12 of 12 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-08T19:32:32.073558Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-06-29T23:14:01.731405Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation f1970a41-a2c8-44f1-bc4d-66b0716a5c3f · inbound

Janus-Pro: Unified Multimodal Understanding and Generation with Data and Model Scaling cites this paper.

Janus-Pro: Unified Multimodal Understanding and Generation with Data and Model Scaling MonoFormer: One Transformer for Both Diffusion and Autoregression

Reference 54

Resolution
verified exact
arxiv_id, observed 2026-05-11T08:14:53.043514Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-11T08:14:52.890145Z digest=sha256:61746790a54ea37cf7982aa8d044b7ba5e122ef23d504b2f56f028ba060b9331

Observation cdafda51-e48f-43b0-8445-fe23c13e09e6 · inbound

UniCMs: A Unified Consistency Model For Efficient Multimodal Generation and Understanding cites this paper.

UniCMs: A Unified Consistency Model For Efficient Multimodal Generation and Understanding MonoFormer: One Transformer for Both Diffusion and Autoregression

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-08T19:32:32.073558Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T19:32:32.073558Z digest=sha256:3d3ba2628808cf9b7382ff96db5f647afbdeb8dbf12673ad28ed8bcb47b2ca27

Observation 6e649b9f-8e7c-4028-8a62-a889466678e9 · inbound

HybridVLA: Collaborative Diffusion and Autoregression in a Unified Vision-Language-Action Model cites this paper.

HybridVLA: Collaborative Diffusion and Autoregression in a Unified Vision-Language-Action Model MonoFormer: One Transformer for Both Diffusion and Autoregression

Reference 68

Resolution
verified exact
arxiv_id, observed 2026-05-15T22:00:48.747273Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-15T22:00:48.667428Z digest=sha256:1b61936753d0289194d37f7dd1e2b0108c08f0e9a1fed299260c0246be1ca2b8

Observation a6476030-a421-4f75-9e33-c9d7a477eed6 · inbound

ComfyMind: Toward General-Purpose Generation via Tree-Based Planning and Reactive Feedback cites this paper.

ComfyMind: Toward General-Purpose Generation via Tree-Based Planning and Reactive Feedback MonoFormer: One Transformer for Both Diffusion and Autoregression

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T14:42:13.772497Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:42:13.772497Z digest=sha256:ef15430310109d49c78a0301f81e2cc2d8029dff1bf4f9b7b7f91b21a1f446c3

Observation bb647602-edc0-4160-b56c-70180aab2871 · inbound

Muddit: Liberating Generation Beyond Text-to-Image with a Unified Discrete Diffusion Model cites this paper.

Muddit: Liberating Generation Beyond Text-to-Image with a Unified Discrete Diffusion Model MonoFormer: One Transformer for Both Diffusion and Autoregression

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-05-19T13:02:18.406221Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-19T12:59:31.454155Z digest=sha256:c0b41433392f708020b13a681543e69cb32dd29f248d7bfe3c156552882f76a4

Observation 69d98c81-9564-48fe-8a07-aeefee81b03d · inbound

MADFormer: Mixed Autoregressive and Diffusion Transformers for Continuous Image Generation cites this paper.

MADFormer: Mixed Autoregressive and Diffusion Transformers for Continuous Image Generation MonoFormer: One Transformer for Both Diffusion and Autoregression

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T05:26:02.300549Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:26:02.300549Z digest=sha256:4cef0a9efbe20144c0513181ae8fa37cbb339fe6e2c693db7a952b94260e9055

Observation 6739e870-947f-49f0-be19-e358def20ee3 · inbound

Show-o2: Improved Native Unified Multimodal Models cites this paper.

Show-o2: Improved Native Unified Multimodal Models MonoFormer: One Transformer for Both Diffusion and Autoregression

Reference 146

Resolution
verified exact
arxiv_id, observed 2026-05-12T18:51:15.977705Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-12T18:51:15.428692Z digest=sha256:6959b9e0f4ac9a581ce725dfcce0ed177846a535080178ed638d406aaa92e1bb

Observation 3b9d40d5-b6bd-490f-9d3f-24a5d8907f3b · inbound

SRUM: Fine-Grained Self-Rewarding for Unified Multimodal Models cites this paper.

SRUM: Fine-Grained Self-Rewarding for Unified Multimodal Models MonoFormer: One Transformer for Both Diffusion and Autoregression

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-04T09:54:26.483487Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:54:26.483487Z digest=sha256:ca555c16f156c58c8bec006d65c12abd330fdc5fd63493ea8fa4a02c922c1b3b

Observation 467c2d02-f499-40a0-a494-5885d7e57832 · inbound

Generative AI Meets 6G and Beyond: Diffusion Models for Semantic Communications cites this paper.

Generative AI Meets 6G and Beyond: Diffusion Models for Semantic Communications MonoFormer: One Transformer for Both Diffusion and Autoregression

Reference 112

Resolution
verified exact
arxiv_id, observed 2026-05-17T23:40:31.129030Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-17T23:40:01.988138Z digest=sha256:7cfd7241a76f29edda81c684789e2888e5e420bfd87b6dd4c248896ab852bdb9

Observation fc17375e-21ef-4e2a-a95d-511875bac497 · inbound

LLaMo: Scaling Pretrained Language Models for Unified Motion Understanding and Generation with Continuous Autoregressive Tokens cites this paper.

LLaMo: Scaling Pretrained Language Models for Unified Motion Understanding and Generation with Continuous Autoregressive Tokens MonoFormer: One Transformer for Both Diffusion and Autoregression

Reference 92

Resolution
verified exact
arxiv_id, observed 2026-05-16T05:02:19.635308Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-16T05:01:11.880003Z digest=sha256:0b4cb9d1be1628f9996133ab7d2d7c915dc99d655431f707c587d6991eaa1121

Observation 82518d8e-1e50-4256-aadb-1cd7e7d593ff · inbound

MONET: A Massive, Open, Non-redundant and Enriched Text-to-image dataset cites this paper.

MONET: A Massive, Open, Non-redundant and Enriched Text-to-image dataset MonoFormer: One Transformer for Both Diffusion and Autoregression

Reference 110

Resolution
verified exact
arxiv_id, observed 2026-05-21T05:23:58.465588Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T05:21:18.369534Z digest=sha256:1eeb895bd75fa2dd529a3b3ca7e1bc2f74150f432e9e7c5f5177d7ac23e3a568

Observation 2400eb6a-2ba1-489c-9c58-8f3faab8de1e · inbound

DIVA: Harnessing the Representation Divergence in Unified Multimodal Models for Mutual Reinforcement cites this paper.

DIVA: Harnessing the Representation Divergence in Unified Multimodal Models for Mutual Reinforcement MonoFormer: One Transformer for Both Diffusion and Autoregression

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-06-29T23:14:01.733544Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-29T23:08:57.793923Z digest=sha256:4856d8d7a1dda6ba34f79a2c2b621eeccd5fb50b3d1b72443fa312aafd5ec885