REVIEW 5 cited by
Refining Generative Process with Discriminator Guidance in Score-based Diffusion Models
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
Signed reviews
read the original abstract
The proposed method, Discriminator Guidance, aims to improve sample generation of pre-trained diffusion models. The approach introduces a discriminator that gives explicit supervision to a denoising sample path whether it is realistic or not. Unlike GANs, our approach does not require joint training of score and discriminator networks. Instead, we train the discriminator after score training, making discriminator training stable and fast to converge. In sample generation, we add an auxiliary term to the pre-trained score to deceive the discriminator. This term corrects the model score to the data score at the optimal discriminator, which implies that the discriminator helps better score estimation in a complementary way. Using our algorithm, we achive state-of-the-art results on ImageNet 256x256 with FID 1.83 and recall 0.64, similar to the validation data's FID (1.68) and recall (0.66). We release the code at https://github.com/alsdudrla10/DG.
Forward citations
Cited by 5 Pith papers
-
Unifying Generative Models with Path Integrals
A one-loop correction, computed from two auxiliary ODEs, brings deterministic generative samplers close to the stochastic reference (53% error reduced to 1.6% on a cubic drift), within a path-integral framework that u...
-
Improving Compositional Generation with Diffusion Models Using Lift Scores
CompLift accepts or rejects generated samples by computing lift scores from conditional and unconditional denoising errors, improving compositional alignment without retraining.
-
Generating time-consistent dynamics with discriminator-guided image diffusion models
A time-consistency discriminator guides a pretrained image diffusion model at inference time to generate realistic, stable spatiotemporal sequences without finetuning the diffusion model.
-
Integration Flow Models
Integration Flow learns the integrated denoising map of an ODE generative model and reports competitive one-step FID on CIFAR-10 and ImageNet for VE diffusion, rectified flow, and PFGM++.
-
Visual Generation Without Guidance
GFT trains a single β-conditioned network that reproduces Classifier-Free Guidance's sampling distribution, matching CFG FID scores across five model families with half the inference cost.
Discussion (0). Continue with ORCID to comment.