REVIEW 4 cited by
TextSR: Content-Aware Text Super-Resolution Guided by Recognition
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
Scene text recognition has witnessed rapid development with the advance of convolutional neural networks. Nonetheless, most of the previous methods may not work well in recognizing text with low resolution which is often seen in natural scene images. An intuitive solution is to introduce super-resolution techniques as pre-processing. However, conventional super-resolution methods in the literature mainly focus on reconstructing the detailed texture of natural images, which typically do not work well for text due to the unique characteristics of text. To tackle these problems, in this work, we propose a content-aware text super-resolution network to generate the information desired for text recognition. In particular, we design an end-to-end network that can perform super-resolution and text recognition simultaneously. Different from previous super-resolution methods, we use the loss of text recognition as the Text Perceptual Loss to guide the training of the super-resolution network, and thus it pays more attention to the text content, rather than the irrelevant background area. Extensive experiments on several challenging benchmarks demonstrate the effectiveness of our proposed method in restoring a sharp high-resolution image from a small blurred one, and show that the recognition performance clearly boosts up the performance of text recognizer. To our knowledge, this is the first work focusing on text super-resolution. Code will be released in https://github.com/xieenze/TextSR.
Forward citations
Cited by 4 Pith papers
-
Coupled Continuous-Discrete Generation for Scene Text Image Super-Resolution
A shared transformer trained with continuous flow matching for images and discrete diffusion for text jointly restores scene text images and reads out their characters, removing the external OCR prior.
-
Text-Aware Image Restoration with Diffusion Models
A diffusion restoration model jointly trained with a text-spotting module and prompted by its own recognized text improves text recognition accuracy on restored images compared with general-purpose restoration methods.
-
TextSR: Diffusion Super-Resolution with Multilingual OCR Guidance
TextSR super-resolves multilingual scene text by conditioning a diffusion model on UTF-8-encoded OCR characters, achieving top OCR accuracy on TextZoom and on small/medium text in a self-defined TextVQA evaluation.
-
Task-driven real-world super-resolution of document scans
Task-driven SR with OCR feature losses improves text-detection IoU on real scans but lowers PSNR, SSIM, and LPIPS relative to bicubic interpolation.
Discussion (0). Continue with ORCID to comment.