REVIEW 8 cited by
Quality-Aware Image-Text Alignment for Opinion-Unaware Image Quality Assessment
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
No-Reference Image Quality Assessment (NR-IQA) focuses on designing methods to measure image quality in alignment with human perception when a high-quality reference image is unavailable. Most state-of-the-art NR-IQA approaches are opinion-aware, i.e. they require human annotations for training. This dependency limits their scalability and broad applicability. To overcome this limitation, we propose QualiCLIP (Quality-aware CLIP), a CLIP-based self-supervised opinion-unaware approach that does not require human opinions. In particular, we introduce a quality-aware image-text alignment strategy to make CLIP generate quality-aware image representations. Starting from pristine images, we synthetically degrade them with increasing levels of intensity. Then, we train CLIP to rank these degraded images based on their similarity to quality-related antonym text prompts. At the same time, we force CLIP to generate consistent representations for images with similar content and the same level of degradation. Our experiments show that the proposed method improves over existing opinion-unaware approaches across multiple datasets with diverse distortion types. Moreover, despite not requiring human annotations, QualiCLIP achieves excellent performance against supervised opinion-aware methods in cross-dataset experiments, thus demonstrating remarkable generalization capabilities. The code and the model are publicly available at https://github.com/miccunifi/QualiCLIP.
Forward citations
Cited by 8 Pith papers
-
Impulse-to-Peak-Output Norm Optimal State-Feedback Control of Linear PDEs
Using PIE representations and Lyapunov LMIs with strong duality, the authors give provable I2P-norm bounds and constructive optimal state-feedback for linear PDEs.
-
Spectral and Trajectory Regularization for Diffusion Transformer Super-Resolution
Asymmetric adversarial distillation plus frequency distribution matching lets DiT models perform one-step Real-ISR without the grid-like periodic artifacts that plague prior one-step DiT distillations.
-
From Global to Granular: Revealing IQA Model Performance via Correlation Surface
GMC maps an IQA model's agreement with human scores across the quality-level and quality-difference landscape, exposing local strengths that global PLCC/SRCC hide.
-
The Devil is in the Darkness: Diffusion-Based Nighttime Dehazing Anchored in Brightness Perception
DiffND combines a depth- and sky-guided data synthesis pipeline with a diffusion model gated by a brightness perception network to achieve nighttime dehazing with day-level brightness.
-
Image Intrinsic Scale Assessment: Bridging the Gap Between Quality and Resolution
A new task and dataset for predicting the scale at which perceived image quality peaks, with a weak-label method that improves several no-reference quality models.
-
HiRQA: Hierarchical Ranking and Quality Alignment for Opinion-Unaware Image Quality Assessment
HiRQA is a self-supervised NR-IQA framework trained on synthetic distortions, using a higher-order ranking loss, embedding distance loss, and text-guided contrastive alignment, claimed to generalize to authentic distortions.
-
Modeling Beyond MOS: Quality Assessment Models Must Integrate Context, Reasoning, and Multimodality
A position paper contending that multimedia quality assessment should move beyond scalar Mean Opinion Score toward context-aware, explainable, and multimodal modeling.
-
A Lightweight Ensemble-Based Face Image Quality Assessment Method with Correlation-Aware Loss
An ensemble of MobileNetV3-Small and ShuffleNetV2 with a correlation-aware loss and test-time augmentation reaches SRCC 0.9829 and PLCC 0.9894 on the VQualA FIQA validation set.
Discussion (0). Continue with ORCID to comment.