REVIEW 3 cited by
DeSRA: Detect and Delete the Artifacts of GAN-based Real-World Super-Resolution Models
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
Image super-resolution (SR) with generative adversarial networks (GAN) has achieved great success in restoring realistic details. However, it is notorious that GAN-based SR models will inevitably produce unpleasant and undesirable artifacts, especially in practical scenarios. Previous works typically suppress artifacts with an extra loss penalty in the training phase. They only work for in-distribution artifact types generated during training. When applied in real-world scenarios, we observe that those improved methods still generate obviously annoying artifacts during inference. In this paper, we analyze the cause and characteristics of the GAN artifacts produced in unseen test data without ground-truths. We then develop a novel method, namely, DeSRA, to Detect and then Delete those SR Artifacts in practice. Specifically, we propose to measure a relative local variance distance from MSE-SR results and GAN-SR results, and locate the problematic areas based on the above distance and semantic-aware thresholds. After detecting the artifact regions, we develop a finetune procedure to improve GAN-based SR models with a few samples, so that they can deal with similar types of artifacts in more unseen real data. Equipped with our DeSRA, we can successfully eliminate artifacts from inference and improve the ability of SR models to be applied in real-world scenarios. The code will be available at https://github.com/TencentARC/DeSRA.
Forward citations
Cited by 3 Pith papers
-
RAGSR: Regional Attention Guided Diffusion for Image Super-Resolution
RAGSR combines region-level vision-language captions with regional attention masks to improve fine-grained detail generation in diffusion-based super-resolution.
-
DroneSR: Rethinking Few-shot Thermal Image Super-Resolution from Drone-based Perspective
Randomized Gaussian-guided quantization of thermal images reduces overfitting for few-shot diffusion super-resolution on a new drone infrared benchmark, but baseline comparisons are not training-matched and the diffus...
-
Incorporating Uncertainty-Guided and Top-k Codebook Matching for Real-World Blind Image Super-Resolution
UGTSR improves codebook-based blind super-resolution by combining uncertainty-guided loss weighting, top-3 codebook matching, and an align-attention module for fusing low- and high-quality features.
Discussion (0). Continue with ORCID to comment.