Pith. sign in

REVIEW 4 cited by

A Closer Look at Time Steps is Worthy of Triple Speed-Up for Diffusion Model Training

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2405.17403 v3 pith:WSUWUA7P submitted 2024-05-27 cs.LG cs.AI

classification cs.LGcs.AI
keywords stepsdiffusiontimetrainingconvergencemodelaccelerationapproach
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Training diffusion models is always a computation-intensive task. In this paper, we introduce a novel speed-up method for diffusion model training, called, which is based on a closer look at time steps. Our key findings are: i) Time steps can be empirically divided into acceleration, deceleration, and convergence areas based on the process increment. ii) These time steps are imbalanced, with many concentrated in the convergence area. iii) The concentrated steps provide limited benefits for diffusion training. To address this, we design an asymmetric sampling strategy that reduces the frequency of steps from the convergence area while increasing the sampling probability for steps from other areas. Additionally, we propose a weighting strategy to emphasize the importance of time steps with rapid-change process increments. As a plug-and-play and architecture-agnostic approach, SpeeD consistently achieves 3-times acceleration across various diffusion architectures, datasets, and tasks. Notably, due to its simple design, our approach significantly reduces the cost of diffusion model training with minimal overhead. Our research enables more researchers to train diffusion models at a lower cost.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 4 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. TASQ: Temporal-Adaptive Bit Sparsification Quantization for Diffusion Models

    cs.CV 2026-08 conditional novelty 7.0 of 10

    A temporal-spatial LSB mask over one shared weight buffer lets diffusion models use lower bit precision in less sensitive denoising stages, cutting compute by 25-50% on bit-serial hardware with no loss in image quality.

  2. Test-Time Scaling of Diffusion Models via Noise Trajectory Search

    cs.LG 2025-05 conditional novelty 6.0 of 10

    An epsilon-greedy search over per-step noise trajectories improves proxy rewards in diffusion image generation without retraining.

  3. REPA Works Until It Doesn't: Early-Stopped, Holistic Alignment Supercharges Diffusion Training

    cs.CV 2025-05 conditional novelty 6.0 of 10

    HASTE trains diffusion transformers faster by aligning student features and attention maps with a DINOv2 teacher early in training and then switching the alignment off, matching vanilla SiT quality on ImageNet 256x256...

  4. Improved Immiscible Diffusion: Accelerate Diffusion Training by Reducing Its Miscibility

    cs.CV 2025-05 conditional novelty 5.0 of 10

    Diffusion models can be trained up to roughly 4x faster in the paper's experiments by reducing trajectory miscibility via KNN noise selection or image scaling, though the mechanism is not fully isolated.

Pith tools