REVIEW 11 cited by
Image Restoration with Mean-Reverting Stochastic Differential Equations
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
This paper presents a stochastic differential equation (SDE) approach for general-purpose image restoration. The key construction consists in a mean-reverting SDE that transforms a high-quality image into a degraded counterpart as a mean state with fixed Gaussian noise. Then, by simulating the corresponding reverse-time SDE, we are able to restore the origin of the low-quality image without relying on any task-specific prior knowledge. Crucially, the proposed mean-reverting SDE has a closed-form solution, allowing us to compute the ground truth time-dependent score and learn it with a neural network. Moreover, we propose a maximum likelihood objective to learn an optimal reverse trajectory that stabilizes the training and improves the restoration results. The experiments show that our proposed method achieves highly competitive performance in quantitative comparisons on image deraining, deblurring, and denoising, setting a new state-of-the-art on two deraining datasets. Finally, the general applicability of our approach is further demonstrated via qualitative results on image super-resolution, inpainting, and dehazing. Code is available at https://github.com/Algolzw/image-restoration-sde.
Forward citations
Cited by 11 Pith papers
-
ScaleResfusion: Residual Rectified Flow based on Residual Vector Field
ScaleResfusion modifies rectified flow to start from a noisy low-quality image and learn only a residual velocity field, enabling 4-step image restoration with LoRA fine-tuning of billion-scale text-to-image models.
-
Decoupling Cross-Modality Manifold Discrepancy: Leveraging Visible Diffusion Priors for Infrared Super-Resolution
Shift-IISR steers a frozen visible-pretrained diffusion model toward the infrared manifold via global representation modulation and local Sobel-edge refinement, improving infrared super-resolution consistency.
-
Geometric Analysis of Magnetic Labyrinthine Stripe Evolution via Deep Learning Segmentation
U-Net segmentation of magneto-optical images combined with skeletonization and graph analysis quantifies the transition from quenched to annealed states in magnetic labyrinthine stripes and identifies two field-polari...
-
TAP: Parameter-efficient Task-Aware Prompting for Adverse Weather Removal
A two-stage prompt-tuning method with low-rank and contrastive prompt enhancement claims all-in-one adverse weather removal at 2.75M parameters.
-
UniDB: A Unified Diffusion Bridge Framework via Stochastic Optimal Control
A stochastic optimal control formulation of diffusion bridges, where Doob's h-transform is the infinite-penalty limit and a finite penalty yields a tunable detail-preserving bridge.
-
Degradation-Consistent Learning via Bidirectional Diffusion for Low-Light Image Enhancement
Training a diffusion model on both enhancement and degradation paths, with a shared encoder and a reflection-aware correction module, yields state-of-the-art low-light enhancement on multiple benchmarks.
-
DiffNMR3: Advancing NMR Resolution Beyond Instrumental Limits
A conditional diffusion model (MSSR) reconstructs high-resolution NMR spectra from Gaussian-blurred low-resolution spectra across multiple upscaling factors, but only synthetic degradation is tested.
-
Diffusion-Based Limited-Angle CT Reconstruction under Noisy Conditions
Adding a noise-matched rectification step (RNSD+) to mean-reverting SDE sinogram inpainting improves limited-angle CT reconstruction under 5-15 dB Gaussian noise on the synthetic ChromSTET2025 dataset.
-
Low-Light Enhancement via Encoder-Decoder Network with Illumination Guidance
EDNIG, a U-Net with illumination guidance, spatial pyramid pooling, and a GAN-based loss, reports 21.512 dB PSNR and 0.8313 SSIM on the LOL validation set.
-
Frequency-Domain Fusion Transformer for Image Inpainting
Dabformer combines wavelet and Gabor filtered attention with an FFT gating network, but the reported gains over prior methods are inconsistent across datasets.
-
Demystifying the Visual Quality Paradox in Multimodal Large Language Models
Multimodal LLM accuracy can improve on visually degraded images, and a lightweight test-time tuning module that modulates input quality yields small accuracy gains on some benchmarks.
Discussion (0). Continue with ORCID to comment.