REVIEW 14 cited by
Domain Generalization with MixStyle
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
Though convolutional neural networks (CNNs) have demonstrated remarkable ability in learning discriminative features, they often generalize poorly to unseen domains. Domain generalization aims to address this problem by learning from a set of source domains a model that is generalizable to any unseen domain. In this paper, a novel approach is proposed based on probabilistically mixing instance-level feature statistics of training samples across source domains. Our method, termed MixStyle, is motivated by the observation that visual domain is closely related to image style (e.g., photo vs.~sketch images). Such style information is captured by the bottom layers of a CNN where our proposed style-mixing takes place. Mixing styles of training instances results in novel domains being synthesized implicitly, which increase the domain diversity of the source domains, and hence the generalizability of the trained model. MixStyle fits into mini-batch training perfectly and is extremely easy to implement. The effectiveness of MixStyle is demonstrated on a wide range of tasks including category classification, instance retrieval and reinforcement learning.
Forward citations
Cited by 14 Pith papers
-
Quantization Meets OOD: Generalizable Quantization-aware Training from a Flatness Perspective
Quantization-aware training degrades out-of-distribution accuracy, and a flatness-aware method with gradient-disorder freezing, FQAT, partially recovers it.
-
MoEcho: Exploiting Side-Channel Attacks to Compromise User Privacy in Mixture-of-Experts LLMs
MoEcho claims to compromise user privacy in MoE LLMs and VLMs via four CPU and GPU side channels, but the provided manuscript body contains no supporting content.
-
CorrMoE: Mixture of Experts with De-stylization Learning for Cross-Scene and Cross-Domain Correspondence Pruning
CorrMoE combines Progressive Mixstyle de-stylization with a Bi-Fusion Mixture-of-Experts module to improve two-view correspondence pruning in cross-domain and cross-scene settings.
-
CTA: Cross-Task Alignment for Better Test Time Training
CTA aligns a SimCLR-trained encoder to a frozen supervised encoder, then adapts only the self-supervised encoder at test time, improving corrupted-image classification over prior test-time training methods.
-
DAM: Domain-Aware Module for Multi-Domain Dataset Condensation
A training-time module with learnable spatial masks and frequency-based pseudo-domain labels improves dataset condensation on multi-domain data without increasing images per class.
-
PIER: Physics-Informed Environmental Retrieval for Time-Series Modeling
PIER augments embedding-based retrieval for lake modeling with a physics-aware stream scored by local verifiers, improving water temperature and dissolved oxygen prediction across 356 lakes.
-
Towards Realistic Hand-Object Interaction with Gravity-Field Based Diffusion Bridge
A gravity-field and diffusion-based optimization refines hand-object contacts to reduce interpenetration and gaps while LLM text prompts guide contact regions.
-
Learning Semantic Directions for Feature Augmentation in Domain-Generalized Medical Segmentation
A feature-augmentation framework with a learned channel selector and covariance-guided intensity sampler that reports the best average domain-generalized segmentation on Prostate and Fundus benchmarks.
-
Positive Style Accumulation: A Style Screening and Continuous Utilization Framework for Federated DG-ReID
SSCU improves federated domain-generalizable person re-identification by screening styles with round-over-round Rank-1 gains and continuously training on the memorized positive styles.
-
Boosting Domain Generalized and Adaptive Detection with Diffusion Models: Fitness, Generalization, and Transferability
A single-step diffusion feature extractor with an object-masked auxiliary branch and consistency loss improves domain-generalized and adaptive detection accuracy and speed.
-
Adaptive Knowledge Distillation using a Device-Aware Teacher for Low-Complexity Acoustic Scene Classification
A low-complexity scene classifier trained by two-teacher knowledge distillation and device-specific fine-tuning reaches 57.93% accuracy on the DCASE 2025 development set.
-
Mix, Align, Distil: Reliable Cross-Domain Atypical Mitosis Classification
A DenseNet-121 trained with MixStyle, CBAM-based feature alignment, and EMA-teacher distillation achieves 0.8762 balanced accuracy on the MIDOG 2025 Task 2 atypical mitosis classification leaderboard.
-
Fully Automated SAM for Single-source Domain Generalization in Medical Image Segmentation
FA-SAM automates SAM-based medical segmentation across domains by generating prompt boxes with an uncertainty-enhanced network and fusing image and prompt embeddings.
-
Fourier Asymmetric Attention on Domain Generalization for Pan-Cancer Drug Response Prediction
FourierDrug uses bulk cell-line expression with adversarial domain generalization and a Fourier asymmetric attention constraint to predict drug response in unseen cancer types, single cells, and patients.
Discussion (0). Continue with ORCID to comment.