REVIEW 4 cited by
Rectified Diffusion: Straightness Is Not Your Need in Rectified Flow
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
abstract
Diffusion models have greatly improved visual generation but are hindered by slow generation speed due to the computationally intensive nature of solving generative ODEs. Rectified flow, a widely recognized solution, improves generation speed by straightening the ODE path. Its key components include: 1) using the diffusion form of flow-matching, 2) employing $\boldsymbol v$-prediction, and 3) performing rectification (a.k.a. reflow). In this paper, we argue that the success of rectification primarily lies in using a pretrained diffusion model to obtain matched pairs of noise and samples, followed by retraining with these matched noise-sample pairs. Based on this, components 1) and 2) are unnecessary. Furthermore, we highlight that straightness is not an essential training target for rectification; rather, it is a specific case of flow-matching models. The more critical training target is to achieve a first-order approximate ODE path, which is inherently curved for models like DDPM and Sub-VP. Building on this insight, we propose Rectified Diffusion, which generalizes the design space and application scope of rectification to encompass the broader category of diffusion models, rather than being restricted to flow-matching models. We validate our method on Stable Diffusion v1-5 and Stable Diffusion XL. Our method not only greatly simplifies the training procedure of rectified flow-based previous works (e.g., InstaFlow) but also achieves superior performance with even lower training cost. Our code is available at https://github.com/G-U-N/Rectified-Diffusion.
Forward citations
Cited by 4 Pith papers
-
Adversarial Distribution Matching for Diffusion Distillation Towards Efficient Image and Video Synthesis
A new adversarial distribution matching loss for diffusion distillation gives one-step and few-step generators that match or exceed prior distillation methods on SDXL, SD3, and CogVideoX.
-
Beyond Optimal Transport: Model-Aligned Coupling for Flow Matching
MAC improves few-step flow-matching generation by up-weighting couplings with the lowest prediction error under the current model.
-
Multi-User Generative Semantic Communication with Intent-Aware Semantic-Splitting Multiple Access
A framework that broadcasts common semantic road maps and personalized text prompts, then jointly optimizes beamforming and semantic extraction with a CLIP/LPIPS-based efficiency score using PPO.
-
AudioTurbo: Fast Text-to-Audio Generation with Rectified Diffusion
AudioTurbo fine-tunes a diffusion model on deterministic noise-audio pairs created by the pretrained Auffusion model, achieving strong text-to-audio results in 10 inference steps.
Discussion (0). Continue with ORCID to comment.