Pith. sign in

Paper Citation Record · LEDGER

PixArt-\Sigma: Weak-to-Strong Training of Diffusion Transformer for 4K Text-to-Image Generation

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 39 inbound Pith citation observations for arXiv:2403.04692.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2403.04692 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 39 of 39 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 39 of 39 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-08T01:04:45.691154Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation c1bf6e93-ff1f-4a8c-ac56-7194297bd8e7 · inbound

PipeFusion: Patch-level Pipeline Parallelism for Diffusion Transformers Inference cites this paper.

PipeFusion: Patch-level Pipeline Parallelism for Diffusion Transformers Inference PixArt-\Sigma: Weak-to-Strong Training of Diffusion Transformer for 4K Text-to-Image Generation

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-24T01:18:42.436471Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-24T01:17:11.261301Z digest=sha256:4ef79c352b0c2e7fe318bb93ca91dc6428f75c5e8a1956dbad6d978a09cf5590

Observation fd326b99-339d-41ef-9657-0f1857e97335 · inbound

Emu3: Next-Token Prediction is All You Need cites this paper.

Emu3: Next-Token Prediction is All You Need PixArt-\Sigma: Weak-to-Strong Training of Diffusion Transformer for 4K Text-to-Image Generation

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-11T10:56:07.071520Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-11T10:56:06.418360Z digest=sha256:489fbb1a33f0a1b250370f07363d82a2085575a013e888c454fb1609ad00809c

Observation 8728af1f-2f07-4c39-bdda-51192b67a42b · inbound

Representation Alignment for Generation: Training Diffusion Transformers Is Easier Than You Think cites this paper.

Representation Alignment for Generation: Training Diffusion Transformers Is Easier Than You Think PixArt-\Sigma: Weak-to-Strong Training of Diffusion Transformer for 4K Text-to-Image Generation

Reference 121

Resolution
verified exact
arxiv_id, observed 2026-05-12T15:09:37.106792Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-12T15:09:36.982610Z digest=sha256:f2964b3d398288ea65d80c5cf530f70178199671fd11b474e2694a7035689cbd

Observation 1dd59143-9730-4ee8-84a2-cc731d0854b8 · inbound

SANA: Efficient High-Resolution Image Synthesis with Linear Diffusion Transformers cites this paper.

SANA: Efficient High-Resolution Image Synthesis with Linear Diffusion Transformers PixArt-\Sigma: Weak-to-Strong Training of Diffusion Transformer for 4K Text-to-Image Generation

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-15T00:56:50.061833Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-15T00:56:50.009149Z digest=sha256:022466cd3d9afc823958896a28fcd6e4d1655e55e0d9a2c7db7097d1074b2bad

Observation 5efd29ea-e9f0-4d9b-bbb7-c04c01814d08 · inbound

Open-Sora Plan: Open-Source Large Video Generation Model cites this paper.

Open-Sora Plan: Open-Source Large Video Generation Model PixArt-\Sigma: Weak-to-Strong Training of Diffusion Transformer for 4K Text-to-Image Generation

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-23T08:42:45.242218Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-23T08:38:27.946746Z digest=sha256:cb7274220c3ab61d8d3570c2fb51d75eb9abe687e51313bcc97e17aff53a8f91

Observation 5f7af95b-3c7a-4a80-8bbe-962ac750425f · inbound

Inference-Time Scaling for Diffusion Models beyond Scaling Denoising Steps cites this paper.

Inference-Time Scaling for Diffusion Models beyond Scaling Denoising Steps PixArt-\Sigma: Weak-to-Strong Training of Diffusion Transformer for 4K Text-to-Image Generation

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-05-20T11:45:17.556019Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-20T11:45:17.473970Z digest=sha256:7074c93bd10396f13659ec8fe0e592c1ff95d9f02f4edc05060bd4045ab55de5

Observation 3a998c9d-486f-477f-9ba3-fe809f89580b · inbound

Janus-Pro: Unified Multimodal Understanding and Generation with Data and Model Scaling cites this paper.

Janus-Pro: Unified Multimodal Understanding and Generation with Data and Model Scaling PixArt-\Sigma: Weak-to-Strong Training of Diffusion Transformer for 4K Text-to-Image Generation

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-11T08:14:53.079659Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-11T08:14:52.890145Z digest=sha256:536b5818d42ccaf138355e4dfac98a9a94455a7425a9ed5dd69b41dab696bec0

Observation b4466de2-93a3-412d-bd41-62fb0e516af2 · inbound

Mogao: An Omni Foundation Model for Interleaved Multi-Modal Generation cites this paper.

Mogao: An Omni Foundation Model for Interleaved Multi-Modal Generation PixArt-\Sigma: Weak-to-Strong Training of Diffusion Transformer for 4K Text-to-Image Generation

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-17T07:24:04.781066Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-17T07:24:04.460276Z digest=sha256:1ed7dcddc4706e9fc6fc9901c44fe874ddf98f5d95700249907a0ede6ebf1c3e

Observation feaf3134-0e15-4747-8289-d5020fd4d328 · inbound

Differentiable Solver Search for Fast Diffusion Sampling cites this paper.

Differentiable Solver Search for Fast Diffusion Sampling PixArt-\Sigma: Weak-to-Strong Training of Diffusion Transformer for 4K Text-to-Image Generation

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-07T13:46:44.443170Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:46:44.443170Z digest=sha256:2f4cbe75f0fb717d7b95c4a4514c4aebb5b6118d072173294df131ea63a07e1b

Observation 6aa7c631-c78f-4fbc-a423-e9be382acc5f · inbound

Normalized Attention Guidance: Universal Negative Guidance for Diffusion Models cites this paper.

Normalized Attention Guidance: Universal Negative Guidance for Diffusion Models PixArt-\Sigma: Weak-to-Strong Training of Diffusion Transformer for 4K Text-to-Image Generation

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T13:42:36.858642Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:42:36.858642Z digest=sha256:96edc0a3967d8e6274a102c7a646bcf9792249f211dd0ea530ad9424f40058b8

Observation cc33ce15-5a5a-40cb-91cd-f8ffe9b554a6 · inbound

OpenUni: A Simple Baseline for Unified Multimodal Understanding and Generation cites this paper.

OpenUni: A Simple Baseline for Unified Multimodal Understanding and Generation PixArt-\Sigma: Weak-to-Strong Training of Diffusion Transformer for 4K Text-to-Image Generation

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T12:44:13.780066Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:44:13.780066Z digest=sha256:265f23dd1f00b98368bd6898b7076b3e6af83a3bf6e3f6b7dfa3242317bc6f3f

Observation a479a1dc-1cf6-4de7-ad80-b1f77a206296 · inbound

Multi-Group Proportional Representation for Text-to-Image Models cites this paper.

Multi-Group Proportional Representation for Text-to-Image Models PixArt-\Sigma: Weak-to-Strong Training of Diffusion Transformer for 4K Text-to-Image Generation

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T12:43:32.426651Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:43:32.426651Z digest=sha256:5f1496800e0559e1695b140a1965253cd678070f1adb706aae815812b2622bf2

Observation 37a2c438-6550-49a2-a6d7-e56fa3cfdf6a · inbound

How Much To Guide: Revisiting Adaptive Guidance in Classifier-Free Guidance Text-to-Vision Diffusion Models cites this paper.

How Much To Guide: Revisiting Adaptive Guidance in Classifier-Free Guidance Text-to-Vision Diffusion Models PixArt-\Sigma: Weak-to-Strong Training of Diffusion Transformer for 4K Text-to-Image Generation

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T05:18:56.551512Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:18:56.551512Z digest=sha256:e16104cdf08217b13c28aee57e4fdc64a63f47189153e47aa086337ac6af4550

Observation 00f7d635-966c-4406-b3de-4ab59a097c14 · inbound

Re-Thinking the Automatic Evaluation of Image-Text Alignment in Text-to-Image Models cites this paper.

Re-Thinking the Automatic Evaluation of Image-Text Alignment in Text-to-Image Models PixArt-\Sigma: Weak-to-Strong Training of Diffusion Transformer for 4K Text-to-Image Generation

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T05:14:40.852888Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:14:40.852888Z digest=sha256:b72c217722dbaed85715fb5a10cf924bc4d5b5c1f7166491bc18f801646f5e56

Observation 85473dbf-336f-4058-a13a-706dc983d539 · inbound

ShareGPT-4o-Image: Aligning Multimodal Models with GPT-4o-Level Image Generation cites this paper.

ShareGPT-4o-Image: Aligning Multimodal Models with GPT-4o-Level Image Generation PixArt-\Sigma: Weak-to-Strong Training of Diffusion Transformer for 4K Text-to-Image Generation

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T23:28:03.229878Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:28:03.229878Z digest=sha256:53cffb5c107d48b96127eb55e2d9492a56c3b40fca544585449366f355632fcf

Observation 1c6e5d5b-eb49-4408-9bf2-6a90c31c6534 · inbound

UniMC: Taming Diffusion Transformer for Unified Keypoint-Guided Multi-Class Image Generation cites this paper.

UniMC: Taming Diffusion Transformer for Unified Keypoint-Guided Multi-Class Image Generation PixArt-\Sigma: Weak-to-Strong Training of Diffusion Transformer for 4K Text-to-Image Generation

Reference 2017

Resolution
unresolved
no resolver link, observed 2026-08-06T20:26:37.842944Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:26:37.842944Z digest=sha256:bd656cfe258818a20be644c4be3dd8e7c36c4edfdd50848f512e9a67c6c5b8d9

Observation 7a5cdab3-5c44-4c2d-857f-bdbbecbb83dc · inbound

Structured Captions Improve Prompt Adherence in Text-to-Image Models (Re-LAION-Caption 19M) cites this paper.

Structured Captions Improve Prompt Adherence in Text-to-Image Models (Re-LAION-Caption 19M) PixArt-\Sigma: Weak-to-Strong Training of Diffusion Transformer for 4K Text-to-Image Generation

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-06T19:50:20.126209Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:50:20.126209Z digest=sha256:5f9257887b96aa3c912b92579ad5e6686813e2ddfd6eaded6bab56a620c7be73

Observation 674c6596-15e6-40f1-8fcf-43b62960a7f5 · inbound

Inversion-DPO: Precise and Efficient Post-Training for Diffusion Models cites this paper.

Inversion-DPO: Precise and Efficient Post-Training for Diffusion Models PixArt-\Sigma: Weak-to-Strong Training of Diffusion Transformer for 4K Text-to-Image Generation

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T17:54:45.724628Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:54:45.724628Z digest=sha256:0dad41b139919b135384e2508bd56cc0a70445979ebdbef35df4525010cca306

Observation ae74d6d3-7395-4556-ac5d-55236b283dec · inbound

DMQ: Dissecting Outliers of Diffusion Models for Post-Training Quantization cites this paper.

DMQ: Dissecting Outliers of Diffusion Models for Post-Training Quantization PixArt-\Sigma: Weak-to-Strong Training of Diffusion Transformer for 4K Text-to-Image Generation

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T16:42:11.259659Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:42:11.259659Z digest=sha256:7eda22c44d397721af5935e8bbd2188810436ab0e2593c5ba6ef9e4f2787aefd

Observation faab2c4a-7a9a-433d-91d0-b5731022b0c1 · inbound

APT: Improving Diffusion Models for High Resolution Image Generation with Adaptive Path Tracing cites this paper.

APT: Improving Diffusion Models for High Resolution Image Generation with Adaptive Path Tracing PixArt-\Sigma: Weak-to-Strong Training of Diffusion Transformer for 4K Text-to-Image Generation

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T12:37:05.366376Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:37:05.366376Z digest=sha256:28ec27de6e4f92ee1b1f2d15028147264491d867a026b788f38bcfe967d0a4e8

Observation f5421e55-af37-428a-8e1a-7ea89fc9132b · inbound

SnapGen++: Unleashing Diffusion Transformers for Efficient High-Fidelity Image Generation on Edge Devices cites this paper.

SnapGen++: Unleashing Diffusion Transformers for Efficient High-Fidelity Image Generation on Edge Devices PixArt-\Sigma: Weak-to-Strong Training of Diffusion Transformer for 4K Text-to-Image Generation

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-03T10:56:11.852224Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T10:56:11.852224Z digest=sha256:3df1dd7dc550a6fb550060d533baf0fdadd63ae30b241932b487ac3ec87e8bad

Observation e9eb2faf-cc5d-45fc-91b7-2bba5ead58f8 · inbound

TIQA: Human-Aligned Perceptual Text Quality Assessment in Generated Images cites this paper.

TIQA: Human-Aligned Perceptual Text Quality Assessment in Generated Images PixArt-\Sigma: Weak-to-Strong Training of Diffusion Transformer for 4K Text-to-Image Generation

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-15T14:35:55.881028Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-15T14:32:59.720416Z digest=sha256:494159a852379553a67becbff7bdabc531e47927cd3721d9c49ff577a87237a5

Observation 68bd0e14-cad3-4529-935b-429b68459ca0 · inbound

SURF: Signature-Retained Fast Video Generation cites this paper.

SURF: Signature-Retained Fast Video Generation PixArt-\Sigma: Weak-to-Strong Training of Diffusion Transformer for 4K Text-to-Image Generation

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-21T18:14:17.450007Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T18:11:39.642701Z digest=sha256:b92256829f1fb91842777649005e07dbd1e557cf56d8e6479f0903233557a382

Observation 47178d1a-d3e3-41b4-a887-bac27138d552 · inbound

Personalizing Text-to-Image Generation to Individual Taste cites this paper.

Personalizing Text-to-Image Generation to Individual Taste PixArt-\Sigma: Weak-to-Strong Training of Diffusion Transformer for 4K Text-to-Image Generation

Reference 8

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T05:26:02.359239Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T18:07:31.236729Z digest=sha256:0ccc56ed4c43c23c97c757edd9f4ee7b98368c9dd6a7fee0bb69ce0a553c882c

Observation 384fa532-d39a-4972-b90f-006eee372352 · inbound

FlowGuard: Towards Lightweight In-Generation Safety Detection for Diffusion Models via Linear Latent Decoding cites this paper.

FlowGuard: Towards Lightweight In-Generation Safety Detection for Diffusion Models via Linear Latent Decoding PixArt-\Sigma: Weak-to-Strong Training of Diffusion Transformer for 4K Text-to-Image Generation

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-11T07:41:00.462920Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T17:03:35.420451Z digest=sha256:da606ae5b53bca68279bd68703e7aeba0079e09a9497401ecbea9f3e2e979917

Observation 3da6f15e-deb5-4bd2-9cde-70ed7e28b979 · inbound

BiasIG: Benchmarking Multi-dimensional Social Biases in Text-to-Image Models cites this paper.

BiasIG: Benchmarking Multi-dimensional Social Biases in Text-to-Image Models PixArt-\Sigma: Weak-to-Strong Training of Diffusion Transformer for 4K Text-to-Image Generation

Reference 45

Resolution
verified exact
arxiv_id, observed 2026-05-11T10:21:00.664118Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T15:33:15.025940Z digest=sha256:73efcffa1b2050b817aceb3178efd33648a647a00fe5f4f39413b819b50f5671

Observation d4d838cf-c640-47a3-a5dd-b30f6936948c · inbound

ACPO: Anchor-Constrained Perceptual Optimization for Diffusion Models with No-Reference Quality Guidance cites this paper.

ACPO: Anchor-Constrained Perceptual Optimization for Diffusion Models with No-Reference Quality Guidance PixArt-\Sigma: Weak-to-Strong Training of Diffusion Transformer for 4K Text-to-Image Generation

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-05-09T03:04:47.616327Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-07T14:05:05.560294Z digest=sha256:fb937ddf5823c64e7b2e3140ff1a3cfd1ddcf99d98ed24cce061ac9f85d9a43f

Observation c2f42710-215e-44fc-85df-39fb931310a2 · inbound

Advancing Aesthetic Image Generation via Composition Transfer cites this paper.

Advancing Aesthetic Image Generation via Composition Transfer PixArt-\Sigma: Weak-to-Strong Training of Diffusion Transformer for 4K Text-to-Image Generation

Reference 4

Resolution
metadata mismatch
arxiv_id, observed 2026-05-09T06:30:44.333026Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-08T18:23:12.827094Z digest=sha256:aa7f37b6778db724ecd7de900ef9d986bcc6ae79a99b93a814416e1646865ad7

Observation baf9c3aa-7f7d-4542-b5ac-c3f9f62bc7a7 · inbound

ElasticDiT: Efficient Diffusion Transformers via Elastic Architecture and Sparse Attention for High-Resolution Image Generation on Mobile Devices cites this paper.

ElasticDiT: Efficient Diffusion Transformers via Elastic Architecture and Sparse Attention for High-Resolution Image Generation on Mobile Devices PixArt-\Sigma: Weak-to-Strong Training of Diffusion Transformer for 4K Text-to-Image Generation

Reference 33

Resolution
metadata mismatch
arxiv_id, observed 2026-05-20T20:03:43.842993Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-20T20:00:27.987481Z digest=sha256:9a97a171e5cec083a17c2fa2e38da03ae2a7dfa4192b48248db7df3292e3cbad

Observation 5dfc211d-86f5-47fd-abdc-2723cb4a028d · inbound

MONET: A Massive, Open, Non-redundant and Enriched Text-to-image dataset cites this paper.

MONET: A Massive, Open, Non-redundant and Enriched Text-to-image dataset PixArt-\Sigma: Weak-to-Strong Training of Diffusion Transformer for 4K Text-to-Image Generation

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-21T05:23:58.469175Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T05:21:18.369534Z digest=sha256:5c1e0873282109da372cf53ef02e3c8ecdb66df1e71cdad4d6b737cee01a8a73

Observation 10e89b33-dd90-4ceb-b175-44480b64073b · inbound

ORBIS: Output-Guided Token Reduction with Distribution-Aware Matching for Video Diffusion Acceleration cites this paper.

ORBIS: Output-Guided Token Reduction with Distribution-Aware Matching for Video Diffusion Acceleration PixArt-\Sigma: Weak-to-Strong Training of Diffusion Transformer for 4K Text-to-Image Generation

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-22T07:24:43.433506Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-22T07:21:16.861996Z digest=sha256:467cfc1992901738271adf96bb75c0c8d7a8d2be7081e0fb295da6cd345d895f

Observation 8d5bd4ae-0825-4893-890a-711a59f87a9b · inbound

DiSC: Resolution-Scalable Acceleration of Diffusion Models by Exploiting Sparsity and Cached Token Reuse with Hash-based Distribution cites this paper.

DiSC: Resolution-Scalable Acceleration of Diffusion Models by Exploiting Sparsity and Cached Token Reuse with Hash-based Distribution PixArt-\Sigma: Weak-to-Strong Training of Diffusion Transformer for 4K Text-to-Image Generation

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-06-29T19:33:53.900908Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-29T19:33:05.248105Z digest=sha256:ed04f8ffcb1679860d7fdc84a56e8873ca81fabb7e9c9f573fc3992fcead3861

Observation 0b450218-c9b7-4c4c-bd1f-ffc49161699b · inbound

FocusDiT: Masking Queries in Diffusion Transformers for Fine-grained Image Generation cites this paper.

FocusDiT: Masking Queries in Diffusion Transformers for Fine-grained Image Generation PixArt-\Sigma: Weak-to-Strong Training of Diffusion Transformer for 4K Text-to-Image Generation

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-07-01T22:26:17.436674Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-28T15:25:40.039365Z digest=sha256:049784d800c9b9c3066b7297fa7730303ad6df97e747f631e310e7d83802c98a

Observation d9c33576-307d-4a8c-8c2e-59afbd9c511a · inbound

Forged Calamity: Benchmark for Cross-Domain Synthetic Disaster Detection in the Age of Diffusion cites this paper.

Forged Calamity: Benchmark for Cross-Domain Synthetic Disaster Detection in the Age of Diffusion PixArt-\Sigma: Weak-to-Strong Training of Diffusion Transformer for 4K Text-to-Image Generation

Reference 3

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T23:59:07.406676Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-26T21:32:27.296146Z digest=sha256:8e81c09c4e097596fc64b7b41972828a3e8146bca40a371e0e5184db8171f019

Observation 361c2a0a-68aa-4f4f-84c3-eef4ceb1430c · inbound

Cross-Space Distillation: Teaching One-Step Students with Modern Diffusion Teachers cites this paper.

Cross-Space Distillation: Teaching One-Step Students with Modern Diffusion Teachers PixArt-\Sigma: Weak-to-Strong Training of Diffusion Transformer for 4K Text-to-Image Generation

Reference 3

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T10:25:41.126840Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-07-01T05:35:48.896721Z digest=sha256:b6e2eccfcc4ff38413bc3e77bf3a2ebf56f4277567ac2545a112afa17ac9bb23

Observation 950cf92b-89e3-43d0-9d20-41e1c08ae5cd · inbound

Unified Backbone Refinement for Diffusion Models via Internal-Latent Analysis cites this paper.

Unified Backbone Refinement for Diffusion Models via Internal-Latent Analysis PixArt-\Sigma: Weak-to-Strong Training of Diffusion Transformer for 4K Text-to-Image Generation

Reference 5

Resolution
unresolved
no resolver link, observed 2026-07-14T16:28:17.428855Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T16:28:17.428855Z digest=sha256:6c9bf5fc35bb86c8fd05faf26f461580c4c670efaa54baacdd64d0626f6b965e

Observation 9a6d307a-9be5-43f5-9ede-8532edf70965 · inbound

MixDiffusion: Mixing Diffusion-based Uni-condition Text-to-Image Generation Models for Multi-condition Image Synthesis cites this paper.

MixDiffusion: Mixing Diffusion-based Uni-condition Text-to-Image Generation Models for Multi-condition Image Synthesis PixArt-\Sigma: Weak-to-Strong Training of Diffusion Transformer for 4K Text-to-Image Generation

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-01T17:30:31.211432Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T17:30:31.211432Z digest=sha256:6cb1496313f975838d69019d04a92d9ee9a0892804834dc80b9533b693157db4

Observation 4b1322ca-60f7-43af-95db-debf5566edb5 · inbound

AI-generated Images Challenge Visual Trust in High-risk Scenarios cites this paper.

AI-generated Images Challenge Visual Trust in High-risk Scenarios PixArt-\Sigma: Weak-to-Strong Training of Diffusion Transformer for 4K Text-to-Image Generation

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-01T08:52:47.034867Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T08:52:47.034867Z digest=sha256:7af3c542bda56c7575d63e3f29c21d21fd2107c4b873f3b2cf5c3b9b3e0b28e3

Observation 5b74a80e-9e74-4a5b-82fa-53b6f5ab358f · inbound

TASQ: Temporal-Adaptive Bit Sparsification Quantization for Diffusion Models cites this paper.

TASQ: Temporal-Adaptive Bit Sparsification Quantization for Diffusion Models PixArt-\Sigma: Weak-to-Strong Training of Diffusion Transformer for 4K Text-to-Image Generation

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-08T01:04:45.691154Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T01:04:45.691154Z digest=sha256:28a995adff280c1e00eea4bab7b621239a06a576e8b282037a8cc1f6c079dbf3