Pith. sign in

Paper Citation Record · LEDGER

PixArt-\Sigma: Weak-to-Strong Training of Diffusion Transformer for 4K Text-to-Image Generation

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 39 inbound Pith citation observations for arXiv:2403.04692.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2403.04692 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 39 of 39 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 39 of 39 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-08T01:04:45.691154Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation c1bf6e93-ff1f-4a8c-ac56-7194297bd8e7 · inbound

PipeFusion: Patch-level Pipeline Parallelism for Diffusion Transformers Inference cites this paper.

PipeFusion: Patch-level Pipeline Parallelism for Diffusion Transformers Inference PixArt-\Sigma: Weak-to-Strong Training of Diffusion Transformer for 4K Text-to-Image Generation

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-24T01:18:42.436471Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-24T01:17:11.261301Z digest=sha256:c92198c6f063cd84a7b4ff77b12f35d780c10c1ee4c6f5344ce235396c594ccd

Observation fd326b99-339d-41ef-9657-0f1857e97335 · inbound

Emu3: Next-Token Prediction is All You Need cites this paper.

Emu3: Next-Token Prediction is All You Need PixArt-\Sigma: Weak-to-Strong Training of Diffusion Transformer for 4K Text-to-Image Generation

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-11T10:56:07.071520Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-11T10:56:06.418360Z digest=sha256:1a99fa4bcded138dfa0f561cc556f53e6e43749e2661c7a16b20cde883d59677

Observation 8728af1f-2f07-4c39-bdda-51192b67a42b · inbound

Representation Alignment for Generation: Training Diffusion Transformers Is Easier Than You Think cites this paper.

Representation Alignment for Generation: Training Diffusion Transformers Is Easier Than You Think PixArt-\Sigma: Weak-to-Strong Training of Diffusion Transformer for 4K Text-to-Image Generation

Reference 121

Resolution
verified exact
arxiv_id, observed 2026-05-12T15:09:37.106792Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-12T15:09:36.982610Z digest=sha256:7fdd78157339d7f3e55f71313625f8d02c77c322658cf46871dd071bff677d64

Observation 1dd59143-9730-4ee8-84a2-cc731d0854b8 · inbound

SANA: Efficient High-Resolution Image Synthesis with Linear Diffusion Transformers cites this paper.

SANA: Efficient High-Resolution Image Synthesis with Linear Diffusion Transformers PixArt-\Sigma: Weak-to-Strong Training of Diffusion Transformer for 4K Text-to-Image Generation

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-15T00:56:50.061833Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-15T00:56:50.009149Z digest=sha256:3810ddd36d58eac63677690ede698701d525f8b531a358d9071733ae8d393083

Observation 5efd29ea-e9f0-4d9b-bbb7-c04c01814d08 · inbound

Open-Sora Plan: Open-Source Large Video Generation Model cites this paper.

Open-Sora Plan: Open-Source Large Video Generation Model PixArt-\Sigma: Weak-to-Strong Training of Diffusion Transformer for 4K Text-to-Image Generation

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-23T08:42:45.242218Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-23T08:38:27.946746Z digest=sha256:039eee86fa9ea8f87c5ba4c4b8e66da0395155e00b22ddeeadabc2f46195e0a1

Observation 5f7af95b-3c7a-4a80-8bbe-962ac750425f · inbound

Inference-Time Scaling for Diffusion Models beyond Scaling Denoising Steps cites this paper.

Inference-Time Scaling for Diffusion Models beyond Scaling Denoising Steps PixArt-\Sigma: Weak-to-Strong Training of Diffusion Transformer for 4K Text-to-Image Generation

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-05-20T11:45:17.556019Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-20T11:45:17.473970Z digest=sha256:a5ee4d206457ec6940f6763508acef60d505b8175c255ad8e7ad1dfdf39d7522

Observation 3a998c9d-486f-477f-9ba3-fe809f89580b · inbound

Janus-Pro: Unified Multimodal Understanding and Generation with Data and Model Scaling cites this paper.

Janus-Pro: Unified Multimodal Understanding and Generation with Data and Model Scaling PixArt-\Sigma: Weak-to-Strong Training of Diffusion Transformer for 4K Text-to-Image Generation

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-11T08:14:53.079659Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-11T08:14:52.890145Z digest=sha256:81d88707a83a4c2acef508c24fbfa3da7f4791ed770f6cdd701551b3ae2fa989

Observation b4466de2-93a3-412d-bd41-62fb0e516af2 · inbound

Mogao: An Omni Foundation Model for Interleaved Multi-Modal Generation cites this paper.

Mogao: An Omni Foundation Model for Interleaved Multi-Modal Generation PixArt-\Sigma: Weak-to-Strong Training of Diffusion Transformer for 4K Text-to-Image Generation

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-17T07:24:04.781066Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-17T07:24:04.460276Z digest=sha256:555bbf0a03af1f1ee383c6df001061a13935521b0587ae9b0d2b3e8a1d7ab201

Observation feaf3134-0e15-4747-8289-d5020fd4d328 · inbound

Differentiable Solver Search for Fast Diffusion Sampling cites this paper.

Differentiable Solver Search for Fast Diffusion Sampling PixArt-\Sigma: Weak-to-Strong Training of Diffusion Transformer for 4K Text-to-Image Generation

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-07T13:46:44.443170Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:46:44.443170Z digest=sha256:2f4cbe75f0fb717d7b95c4a4514c4aebb5b6118d072173294df131ea63a07e1b

Observation 6aa7c631-c78f-4fbc-a423-e9be382acc5f · inbound

Normalized Attention Guidance: Universal Negative Guidance for Diffusion Models cites this paper.

Normalized Attention Guidance: Universal Negative Guidance for Diffusion Models PixArt-\Sigma: Weak-to-Strong Training of Diffusion Transformer for 4K Text-to-Image Generation

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T13:42:36.858642Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:42:36.858642Z digest=sha256:96edc0a3967d8e6274a102c7a646bcf9792249f211dd0ea530ad9424f40058b8

Observation cc33ce15-5a5a-40cb-91cd-f8ffe9b554a6 · inbound

OpenUni: A Simple Baseline for Unified Multimodal Understanding and Generation cites this paper.

OpenUni: A Simple Baseline for Unified Multimodal Understanding and Generation PixArt-\Sigma: Weak-to-Strong Training of Diffusion Transformer for 4K Text-to-Image Generation

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T12:44:13.780066Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:44:13.780066Z digest=sha256:265f23dd1f00b98368bd6898b7076b3e6af83a3bf6e3f6b7dfa3242317bc6f3f

Observation a479a1dc-1cf6-4de7-ad80-b1f77a206296 · inbound

Multi-Group Proportional Representation for Text-to-Image Models cites this paper.

Multi-Group Proportional Representation for Text-to-Image Models PixArt-\Sigma: Weak-to-Strong Training of Diffusion Transformer for 4K Text-to-Image Generation

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T12:43:32.426651Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:43:32.426651Z digest=sha256:5f1496800e0559e1695b140a1965253cd678070f1adb706aae815812b2622bf2

Observation 37a2c438-6550-49a2-a6d7-e56fa3cfdf6a · inbound

How Much To Guide: Revisiting Adaptive Guidance in Classifier-Free Guidance Text-to-Vision Diffusion Models cites this paper.

How Much To Guide: Revisiting Adaptive Guidance in Classifier-Free Guidance Text-to-Vision Diffusion Models PixArt-\Sigma: Weak-to-Strong Training of Diffusion Transformer for 4K Text-to-Image Generation

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T05:18:56.551512Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:18:56.551512Z digest=sha256:e16104cdf08217b13c28aee57e4fdc64a63f47189153e47aa086337ac6af4550

Observation 00f7d635-966c-4406-b3de-4ab59a097c14 · inbound

Re-Thinking the Automatic Evaluation of Image-Text Alignment in Text-to-Image Models cites this paper.

Re-Thinking the Automatic Evaluation of Image-Text Alignment in Text-to-Image Models PixArt-\Sigma: Weak-to-Strong Training of Diffusion Transformer for 4K Text-to-Image Generation

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T05:14:40.852888Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:14:40.852888Z digest=sha256:b72c217722dbaed85715fb5a10cf924bc4d5b5c1f7166491bc18f801646f5e56

Observation 85473dbf-336f-4058-a13a-706dc983d539 · inbound

ShareGPT-4o-Image: Aligning Multimodal Models with GPT-4o-Level Image Generation cites this paper.

ShareGPT-4o-Image: Aligning Multimodal Models with GPT-4o-Level Image Generation PixArt-\Sigma: Weak-to-Strong Training of Diffusion Transformer for 4K Text-to-Image Generation

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T23:28:03.229878Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:28:03.229878Z digest=sha256:53cffb5c107d48b96127eb55e2d9492a56c3b40fca544585449366f355632fcf

Observation 1c6e5d5b-eb49-4408-9bf2-6a90c31c6534 · inbound

UniMC: Taming Diffusion Transformer for Unified Keypoint-Guided Multi-Class Image Generation cites this paper.

UniMC: Taming Diffusion Transformer for Unified Keypoint-Guided Multi-Class Image Generation PixArt-\Sigma: Weak-to-Strong Training of Diffusion Transformer for 4K Text-to-Image Generation

Reference 2017

Resolution
unresolved
no resolver link, observed 2026-08-06T20:26:37.842944Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:26:37.842944Z digest=sha256:bd656cfe258818a20be644c4be3dd8e7c36c4edfdd50848f512e9a67c6c5b8d9

Observation 7a5cdab3-5c44-4c2d-857f-bdbbecbb83dc · inbound

Structured Captions Improve Prompt Adherence in Text-to-Image Models (Re-LAION-Caption 19M) cites this paper.

Structured Captions Improve Prompt Adherence in Text-to-Image Models (Re-LAION-Caption 19M) PixArt-\Sigma: Weak-to-Strong Training of Diffusion Transformer for 4K Text-to-Image Generation

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-06T19:50:20.126209Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:50:20.126209Z digest=sha256:5f9257887b96aa3c912b92579ad5e6686813e2ddfd6eaded6bab56a620c7be73

Observation 674c6596-15e6-40f1-8fcf-43b62960a7f5 · inbound

Inversion-DPO: Precise and Efficient Post-Training for Diffusion Models cites this paper.

Inversion-DPO: Precise and Efficient Post-Training for Diffusion Models PixArt-\Sigma: Weak-to-Strong Training of Diffusion Transformer for 4K Text-to-Image Generation

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T17:54:45.724628Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:54:45.724628Z digest=sha256:0dad41b139919b135384e2508bd56cc0a70445979ebdbef35df4525010cca306

Observation ae74d6d3-7395-4556-ac5d-55236b283dec · inbound

DMQ: Dissecting Outliers of Diffusion Models for Post-Training Quantization cites this paper.

DMQ: Dissecting Outliers of Diffusion Models for Post-Training Quantization PixArt-\Sigma: Weak-to-Strong Training of Diffusion Transformer for 4K Text-to-Image Generation

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T16:42:11.259659Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:42:11.259659Z digest=sha256:7eda22c44d397721af5935e8bbd2188810436ab0e2593c5ba6ef9e4f2787aefd

Observation faab2c4a-7a9a-433d-91d0-b5731022b0c1 · inbound

APT: Improving Diffusion Models for High Resolution Image Generation with Adaptive Path Tracing cites this paper.

APT: Improving Diffusion Models for High Resolution Image Generation with Adaptive Path Tracing PixArt-\Sigma: Weak-to-Strong Training of Diffusion Transformer for 4K Text-to-Image Generation

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T12:37:05.366376Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:37:05.366376Z digest=sha256:28ec27de6e4f92ee1b1f2d15028147264491d867a026b788f38bcfe967d0a4e8

Observation f5421e55-af37-428a-8e1a-7ea89fc9132b · inbound

SnapGen++: Unleashing Diffusion Transformers for Efficient High-Fidelity Image Generation on Edge Devices cites this paper.

SnapGen++: Unleashing Diffusion Transformers for Efficient High-Fidelity Image Generation on Edge Devices PixArt-\Sigma: Weak-to-Strong Training of Diffusion Transformer for 4K Text-to-Image Generation

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-03T10:56:11.852224Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T10:56:11.852224Z digest=sha256:3df1dd7dc550a6fb550060d533baf0fdadd63ae30b241932b487ac3ec87e8bad

Observation e9eb2faf-cc5d-45fc-91b7-2bba5ead58f8 · inbound

TIQA: Human-Aligned Perceptual Text Quality Assessment in Generated Images cites this paper.

TIQA: Human-Aligned Perceptual Text Quality Assessment in Generated Images PixArt-\Sigma: Weak-to-Strong Training of Diffusion Transformer for 4K Text-to-Image Generation

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-15T14:35:55.881028Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-15T14:32:59.720416Z digest=sha256:a84b719c2a05d5c6388a546323f6adf0c8d6a642b7081b4ec8c23e6ac65843ac

Observation 68bd0e14-cad3-4529-935b-429b68459ca0 · inbound

SURF: Signature-Retained Fast Video Generation cites this paper.

SURF: Signature-Retained Fast Video Generation PixArt-\Sigma: Weak-to-Strong Training of Diffusion Transformer for 4K Text-to-Image Generation

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-21T18:14:17.450007Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-21T18:11:39.642701Z digest=sha256:3d578a658e420661c0546a405b72e56fa200e96eedbef52743210fa3ab0f545b

Observation 47178d1a-d3e3-41b4-a887-bac27138d552 · inbound

Personalizing Text-to-Image Generation to Individual Taste cites this paper.

Personalizing Text-to-Image Generation to Individual Taste PixArt-\Sigma: Weak-to-Strong Training of Diffusion Transformer for 4K Text-to-Image Generation

Reference 8

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T05:26:02.359239Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T18:07:31.236729Z digest=sha256:7751d73583a9b055e994f64d89f3b4be201aff0337a01bb759cf5035b8005b87

Observation 384fa532-d39a-4972-b90f-006eee372352 · inbound

FlowGuard: Towards Lightweight In-Generation Safety Detection for Diffusion Models via Linear Latent Decoding cites this paper.

FlowGuard: Towards Lightweight In-Generation Safety Detection for Diffusion Models via Linear Latent Decoding PixArt-\Sigma: Weak-to-Strong Training of Diffusion Transformer for 4K Text-to-Image Generation

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-11T07:41:00.462920Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T17:03:35.420451Z digest=sha256:02abe48bdf5b5bdc1a054a247e6450b698665fd19badef6ec25895bdca77744f

Observation 3da6f15e-deb5-4bd2-9cde-70ed7e28b979 · inbound

BiasIG: Benchmarking Multi-dimensional Social Biases in Text-to-Image Models cites this paper.

BiasIG: Benchmarking Multi-dimensional Social Biases in Text-to-Image Models PixArt-\Sigma: Weak-to-Strong Training of Diffusion Transformer for 4K Text-to-Image Generation

Reference 45

Resolution
verified exact
arxiv_id, observed 2026-05-11T10:21:00.664118Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T15:33:15.025940Z digest=sha256:30a9876d0183598442780c16b2da2bb8b9400042f0f08e42f25cd0f5125b7344

Observation d4d838cf-c640-47a3-a5dd-b30f6936948c · inbound

ACPO: Anchor-Constrained Perceptual Optimization for Diffusion Models with No-Reference Quality Guidance cites this paper.

ACPO: Anchor-Constrained Perceptual Optimization for Diffusion Models with No-Reference Quality Guidance PixArt-\Sigma: Weak-to-Strong Training of Diffusion Transformer for 4K Text-to-Image Generation

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-05-09T03:04:47.616327Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-07T14:05:05.560294Z digest=sha256:a76b9701821924bf71221256758148ce58e0a271975e50d7b177546a4ef5bea8

Observation c2f42710-215e-44fc-85df-39fb931310a2 · inbound

Advancing Aesthetic Image Generation via Composition Transfer cites this paper.

Advancing Aesthetic Image Generation via Composition Transfer PixArt-\Sigma: Weak-to-Strong Training of Diffusion Transformer for 4K Text-to-Image Generation

Reference 4

Resolution
metadata mismatch
arxiv_id, observed 2026-05-09T06:30:44.333026Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-08T18:23:12.827094Z digest=sha256:b0395c40746a6bbe59882f34ff931fbd20ff8aec2376d30636bc2fab9b15883f

Observation baf9c3aa-7f7d-4542-b5ac-c3f9f62bc7a7 · inbound

ElasticDiT: Efficient Diffusion Transformers via Elastic Architecture and Sparse Attention for High-Resolution Image Generation on Mobile Devices cites this paper.

ElasticDiT: Efficient Diffusion Transformers via Elastic Architecture and Sparse Attention for High-Resolution Image Generation on Mobile Devices PixArt-\Sigma: Weak-to-Strong Training of Diffusion Transformer for 4K Text-to-Image Generation

Reference 33

Resolution
metadata mismatch
arxiv_id, observed 2026-05-20T20:03:43.842993Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-20T20:00:27.987481Z digest=sha256:41807e77c40a7d1d0629ef1d7b322b9ab3c94845c47335defd18d897cb994562

Observation 5dfc211d-86f5-47fd-abdc-2723cb4a028d · inbound

MONET: A Massive, Open, Non-redundant and Enriched Text-to-image dataset cites this paper.

MONET: A Massive, Open, Non-redundant and Enriched Text-to-image dataset PixArt-\Sigma: Weak-to-Strong Training of Diffusion Transformer for 4K Text-to-Image Generation

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-21T05:23:58.469175Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-21T05:21:18.369534Z digest=sha256:4ec59ab9d2f3cd8222201f29a7ff79a73f5ad32f5e4bb3b81d46d558daedb793

Observation 10e89b33-dd90-4ceb-b175-44480b64073b · inbound

ORBIS: Output-Guided Token Reduction with Distribution-Aware Matching for Video Diffusion Acceleration cites this paper.

ORBIS: Output-Guided Token Reduction with Distribution-Aware Matching for Video Diffusion Acceleration PixArt-\Sigma: Weak-to-Strong Training of Diffusion Transformer for 4K Text-to-Image Generation

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-22T07:24:43.433506Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-22T07:21:16.861996Z digest=sha256:22ec27663024d648d36e9a1253ff29abf0d1be3ec48d4547699748908ae3f9e6

Observation 8d5bd4ae-0825-4893-890a-711a59f87a9b · inbound

DiSC: Resolution-Scalable Acceleration of Diffusion Models by Exploiting Sparsity and Cached Token Reuse with Hash-based Distribution cites this paper.

DiSC: Resolution-Scalable Acceleration of Diffusion Models by Exploiting Sparsity and Cached Token Reuse with Hash-based Distribution PixArt-\Sigma: Weak-to-Strong Training of Diffusion Transformer for 4K Text-to-Image Generation

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-06-29T19:33:53.900908Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-29T19:33:05.248105Z digest=sha256:f2d247f7ef4d22c8a54000b5195f89b52bf30379e42140fc071da7508570e78f

Observation 0b450218-c9b7-4c4c-bd1f-ffc49161699b · inbound

FocusDiT: Masking Queries in Diffusion Transformers for Fine-grained Image Generation cites this paper.

FocusDiT: Masking Queries in Diffusion Transformers for Fine-grained Image Generation PixArt-\Sigma: Weak-to-Strong Training of Diffusion Transformer for 4K Text-to-Image Generation

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-07-01T22:26:17.436674Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-28T15:25:40.039365Z digest=sha256:707e935b93385290d09aa400cf8fd94c70d747aa77a44327d0c536936901a460

Observation d9c33576-307d-4a8c-8c2e-59afbd9c511a · inbound

Forged Calamity: Benchmark for Cross-Domain Synthetic Disaster Detection in the Age of Diffusion cites this paper.

Forged Calamity: Benchmark for Cross-Domain Synthetic Disaster Detection in the Age of Diffusion PixArt-\Sigma: Weak-to-Strong Training of Diffusion Transformer for 4K Text-to-Image Generation

Reference 3

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T23:59:07.406676Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-26T21:32:27.296146Z digest=sha256:1a93b000402fb5f4729877908e33821d6b1d2ba156c174c24ad1e9214e12b129

Observation 361c2a0a-68aa-4f4f-84c3-eef4ceb1430c · inbound

Cross-Space Distillation: Teaching One-Step Students with Modern Diffusion Teachers cites this paper.

Cross-Space Distillation: Teaching One-Step Students with Modern Diffusion Teachers PixArt-\Sigma: Weak-to-Strong Training of Diffusion Transformer for 4K Text-to-Image Generation

Reference 3

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T10:25:41.126840Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-07-01T05:35:48.896721Z digest=sha256:350533e2e643b7940e74281d5e19604ac922aa018933e564196a52b6d5b424ab

Observation 950cf92b-89e3-43d0-9d20-41e1c08ae5cd · inbound

Unified Backbone Refinement for Diffusion Models via Internal-Latent Analysis cites this paper.

Unified Backbone Refinement for Diffusion Models via Internal-Latent Analysis PixArt-\Sigma: Weak-to-Strong Training of Diffusion Transformer for 4K Text-to-Image Generation

Reference 5

Resolution
unresolved
no resolver link, observed 2026-07-14T16:28:17.428855Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T16:28:17.428855Z digest=sha256:6c9bf5fc35bb86c8fd05faf26f461580c4c670efaa54baacdd64d0626f6b965e

Observation 9a6d307a-9be5-43f5-9ede-8532edf70965 · inbound

MixDiffusion: Mixing Diffusion-based Uni-condition Text-to-Image Generation Models for Multi-condition Image Synthesis cites this paper.

MixDiffusion: Mixing Diffusion-based Uni-condition Text-to-Image Generation Models for Multi-condition Image Synthesis PixArt-\Sigma: Weak-to-Strong Training of Diffusion Transformer for 4K Text-to-Image Generation

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-01T17:30:31.211432Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T17:30:31.211432Z digest=sha256:6cb1496313f975838d69019d04a92d9ee9a0892804834dc80b9533b693157db4

Observation 4b1322ca-60f7-43af-95db-debf5566edb5 · inbound

AI-generated Images Challenge Visual Trust in High-risk Scenarios cites this paper.

AI-generated Images Challenge Visual Trust in High-risk Scenarios PixArt-\Sigma: Weak-to-Strong Training of Diffusion Transformer for 4K Text-to-Image Generation

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-01T08:52:47.034867Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T08:52:47.034867Z digest=sha256:7af3c542bda56c7575d63e3f29c21d21fd2107c4b873f3b2cf5c3b9b3e0b28e3

Observation 5b74a80e-9e74-4a5b-82fa-53b6f5ab358f · inbound

TASQ: Temporal-Adaptive Bit Sparsification Quantization for Diffusion Models cites this paper.

TASQ: Temporal-Adaptive Bit Sparsification Quantization for Diffusion Models PixArt-\Sigma: Weak-to-Strong Training of Diffusion Transformer for 4K Text-to-Image Generation

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-08T01:04:45.691154Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T01:04:45.691154Z digest=sha256:28a995adff280c1e00eea4bab7b621239a06a576e8b282037a8cc1f6c079dbf3