Pith. sign in

Paper Citation Record · LEDGER

CogView3: Finer and Faster Text-to-Image Generation via Relay Diffusion

As of 13 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 12 inbound Pith citation observations for arXiv:2403.05121.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2403.05121 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 12 of 12 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-13T06:32:02.005865+00:00

measured 12 of 12 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-12T13:33:37.081737Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-19T16:37:39.706914Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 755c6773-817d-49a2-b51c-6e579d3eb58a · inbound

CogVideoX: Text-to-Video Diffusion Models with An Expert Transformer cites this paper.

CogVideoX: Text-to-Video Diffusion Models with An Expert Transformer CogView3: Finer and Faster Text-to-Image Generation via Relay Diffusion

Reference 113

Resolution
verified exact
arxiv_id, observed 2026-05-10T18:26:22.373165Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-05-10T18:26:22.224924Z digest=sha256:0e2a5c8f19c2f918b086d3d817b0a62f24d3b300f2ca05767dee7d128ca3de20

Observation 6cd1de61-ea4e-4deb-b001-b7c970360934 · inbound

Text-to-Image Synthesis: A Decade Survey cites this paper.

Text-to-Image Synthesis: A Decade Survey CogView3: Finer and Faster Text-to-Image Generation via Relay Diffusion

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-12T13:33:37.081737Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T13:33:37.081737Z digest=sha256:a5c4f9fa2072dd0b6faf33c2e76e70ca6a7850e4e8049f93669c1c2a9f56333c

Observation 5f1d241a-6087-4577-a6fc-1a5c6ee886a9 · inbound

Self-Cross Diffusion Guidance for Text-to-Image Synthesis of Similar Subjects cites this paper.

Self-Cross Diffusion Guidance for Text-to-Image Synthesis of Similar Subjects CogView3: Finer and Faster Text-to-Image Generation via Relay Diffusion

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-12T10:46:55.052802Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:46:55.052802Z digest=sha256:41405b2e3eb0fc946933d6629790cc421ecbdd46a06120136c7c51e8871f64dc

Observation 6146ddc6-b2b9-4022-a09e-97fa9ce131df · inbound

Owl-1: Omni World Model for Consistent Long Video Generation cites this paper.

Owl-1: Omni World Model for Consistent Long Video Generation CogView3: Finer and Faster Text-to-Image Generation via Relay Diffusion

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-11T16:58:05.615131Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:58:05.615131Z digest=sha256:fb1cbb72dbeac451f5aeaee508d97c5645675d012dd7bd1d2b2224569e0623ce

Observation d04c625c-b45c-402b-b027-90853af58386 · inbound

From Noise to Nuance: Advances in Deep Generative Image Models cites this paper.

From Noise to Nuance: Advances in Deep Generative Image Models CogView3: Finer and Faster Text-to-Image Generation via Relay Diffusion

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-11T17:33:17.758804Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T17:33:17.758804Z digest=sha256:cd57a613f72a272aa79d130c1ed595290f9dfb11116990bb0186575d8650bf27

Observation ef2769b1-060a-4613-bb64-aae6b48c1fff · inbound

SafeCFG: Controlling Harmful Features with Dynamic Safe Guidance for Safe Generation cites this paper.

SafeCFG: Controlling Harmful Features with Dynamic Safe Guidance for Safe Generation CogView3: Finer and Faster Text-to-Image Generation via Relay Diffusion

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-11T10:56:14.414958Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T10:56:14.414958Z digest=sha256:bc8aa76a3bfabdbd2faa2f20fcb1d880f8332bce020bdfe683efe27b682dd19a

Observation 1dd032ed-7d5e-43ad-9f67-9266eec42b7f · inbound

MMIG-Bench: Towards Comprehensive and Explainable Evaluation of Multi-Modal Image Generation Models cites this paper.

MMIG-Bench: Towards Comprehensive and Explainable Evaluation of Multi-Modal Image Generation Models CogView3: Finer and Faster Text-to-Image Generation via Relay Diffusion

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-07T14:17:37.891666Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:17:37.891666Z digest=sha256:1d527e701fd3be6bc86683dd26dfc1ecd0d88623fe55376a8369cd29ea0cd5c0

Observation 0ea1fe66-6904-44fb-9d3c-61110a361a00 · inbound

Z-Image: An Efficient Image Generation Foundation Model with Single-Stream Diffusion Transformer cites this paper.

Z-Image: An Efficient Image Generation Foundation Model with Single-Stream Diffusion Transformer CogView3: Finer and Faster Text-to-Image Generation via Relay Diffusion

Reference 97

Resolution
verified exact
arxiv_id, observed 2026-05-11T14:08:37.346049Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-11T14:08:36.801359Z digest=sha256:3d88edea5c06bf8c10e2ad46e82fccf5d8ff181b08f687d83a79511d556929e2

Observation 5cc1cd5d-82f7-4f75-8817-df4ba1599eb1 · inbound

Z-Image: An Efficient Image Generation Foundation Model with Single-Stream Diffusion Transformer cites this paper.

Z-Image: An Efficient Image Generation Foundation Model with Single-Stream Diffusion Transformer CogView3: Finer and Faster Text-to-Image Generation via Relay Diffusion

Reference 97

Resolution
unresolved
no resolver link, observed 2026-08-03T19:47:32.968338Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T19:47:32.968338Z digest=sha256:2e8495f19084a7d14fbe82bbd9f6379095a9e876436195fe7f8323b95192482a

Observation b7f42422-ca9b-4d61-9356-e8e00118026e · inbound

DynT2I-Eval: A Dynamic Evaluation Framework for Text-to-Image Models cites this paper.

DynT2I-Eval: A Dynamic Evaluation Framework for Text-to-Image Models CogView3: Finer and Faster Text-to-Image Generation via Relay Diffusion

Reference 49

Resolution
verified exact
arxiv_id, observed 2026-05-11T18:46:10.475551Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-08T13:54:00.141439Z digest=sha256:291da78fe423e85ee2ec57bf3293783587fda8bba121e5600ee5d21712028b95

Observation 495e1f8a-d1cf-4c69-8b9a-06ab5cf47ff8 · inbound

ImageAttributionBench: How Far Are We from Generalizable Attribution? cites this paper.

ImageAttributionBench: How Far Are We from Generalizable Attribution? CogView3: Finer and Faster Text-to-Image Generation via Relay Diffusion

Reference 86

Resolution
verified exact
arxiv_id, observed 2026-05-14T19:39:23.894406Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-14T19:38:41.659261Z digest=sha256:7681cdb06bed8e483ec99529c1b324560d774f9e89e59d1890a662198ac4a8a1

Observation 664e705a-63c5-455d-a0bc-6feb23e7adfe · inbound

Unlocking Complex Visual Generation via Closed-Loop Verified Reasoning cites this paper.

Unlocking Complex Visual Generation via Closed-Loop Verified Reasoning CogView3: Finer and Faster Text-to-Image Generation via Relay Diffusion

Reference 60

Resolution
verified exact
arxiv_id, observed 2026-05-19T16:37:39.708562Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-19T16:35:43.166697Z digest=sha256:64984f871acc699afd5b3daf34944b9549c24c571b7abaec6d451611ff42317e