REVIEW 2 cited by
JND-Based Perceptual Optimization For Learned Image Compression
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
Signed reviews
read the original abstract
Recently, learned image compression schemes have achieved remarkable improvements in image fidelity (e.g., PSNR and MS-SSIM) compared to conventional hybrid image coding ones due to their high-efficiency non-linear transform, end-to-end optimization frameworks, etc. However, few of them take the Just Noticeable Difference (JND) characteristic of the Human Visual System (HVS) into account and optimize learned image compression towards perceptual quality. To address this issue, a JND-based perceptual quality loss is proposed. Considering that the amounts of distortion in the compressed image at different training epochs under different Quantization Parameters (QPs) are different, we develop a distortion-aware adjustor. After combining them together, we can better assign the distortion in the compressed image with the guidance of JND to preserve the high perceptual quality. All these designs enable the proposed method to be flexibly applied to various learned image compression schemes with high scalability and plug-and-play advantages. Experimental results on the Kodak dataset demonstrate that the proposed method has led to better perceptual quality than the baseline model under the same bit rate.
Forward citations
Cited by 2 Pith papers
-
Customizable ROI-Based Deep Image Compression
A text-prompted ROI image coder with a user-controlled quality knob and latent mask attention achieves strong RD and machine-vision results, though headline numbers use ground-truth masks.
-
DT-JRD: Deep Transformer based Just Recognizable Difference Prediction Model for Video Coding for Machines
A transformer-based model predicts the minimum detectable distortion for machine vision and uses it to reduce video coding bitrate by about 30% without measured loss in object-detection accuracy.
Discussion (0). Continue with ORCID to comment.