REVIEW 23 cited by
In Search of Lost Domain Generalization
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
The goal of domain generalization algorithms is to predict well on distributions different from those seen during training. While a myriad of domain generalization algorithms exist, inconsistencies in experimental conditions -- datasets, architectures, and model selection criteria -- render fair and realistic comparisons difficult. In this paper, we are interested in understanding how useful domain generalization algorithms are in realistic settings. As a first step, we realize that model selection is non-trivial for domain generalization tasks. Contrary to prior work, we argue that domain generalization algorithms without a model selection strategy should be regarded as incomplete. Next, we implement DomainBed, a testbed for domain generalization including seven multi-domain datasets, nine baseline algorithms, and three model selection criteria. We conduct extensive experiments using DomainBed and find that, when carefully implemented, empirical risk minimization shows state-of-the-art performance across all datasets. Looking forward, we hope that the release of DomainBed, along with contributions from fellow researchers, will streamline reproducible and rigorous research in domain generalization.
Forward citations
Cited by 23 Pith papers
-
Quantization Meets OOD: Generalizable Quantization-aware Training from a Flatness Perspective
Quantization-aware training degrades out-of-distribution accuracy, and a flatness-aware method with gradient-disorder freezing, FQAT, partially recovers it.
-
Face4FairShifts: A Large Image Benchmark for Fairness and Robust Learning across Visual Domains
A new face benchmark with four visual domains and fairness-sensitive labels provides larger measured distribution shifts and lower baseline performance than existing fairness datasets.
-
OpFlow: Learning Opportunity-Conditioned Choice Potentials for Robust OD Flow Prediction
Robust OD flow prediction comes from learning row-centered exposure-to-choice potentials and reconstructing counts as scale times allocation, not from raw-count supervision.
-
Exposure is not manifestation: measurement target and output resolution jointly determine which behavioural-faithfulness evaluator wins
Small hyperbolic models (146M–3B) report 100% creative-seed preference, 90.7% compliance-gap detection, and a selective-gating skeleton–wallpaper memory pilot as a companion-AI stack.
-
Rethink Domain Generalization in Heterogeneous Sequence MRI Segmentation
A semi-supervised pretraining method improves cross-sequence pancreas segmentation Dice from 43.55% to 70.39% (NU) and from 35.62% to 66.61% (IH) on the new PancreasDG benchmark.
-
A Supervised Machine Learning Framework for Multipactor Breakdown Prediction in High-Power Radio Frequency Devices and Accelerator Components: A Case Study in Planar Geometry
Random Forest and Extra Trees trained on PIC data predict planar multipactor susceptibility maps about as well as a Monte Carlo benchmark, while neural networks generalize more poorly to unseen materials.
-
Simulate, Refocus and Ensemble: An Attention-Refocusing Scheme for Domain Generalization
SRE improves CLIP's domain generalization by training an attention-refocuser on simulated target domains and ensembling the most attention-consistent checkpoints.
-
Subgraph Generation for Generalizing on Out-of-Distribution Links
FLEX is a generative framework that synthesizes counterfactual subgraphs with a semi-implicit graph VAE and adversarially co-trains a GNN to improve out-of-distribution link prediction.
-
Harmonizing and Merging Source Models for CLIP-based Domain Generalization
HAM trains per-domain CLIP encoders, enriches them with confident cross-domain samples, aligns their update directions, and merges them with redundancy trimming, reaching 79.0% average accuracy on five DG benchmarks w...
-
Moment Alignment: Unifying Gradient and Hessian Matching for Domain Generalization
A unified moment-alignment theory bounds target-domain error by cross-domain differences in loss derivatives, and the new CMA algorithm implements exact gradient and Hessian matching in closed form.
-
Hidden-Domain Routing for All-Type Audio Deepfake Detection
A router-then-specialist audio deepfake detector, which classifies audio type first and then applies type-specific models and thresholds, achieved 96.10% Macro-F1 and first place on AT-ADD Track2.
-
OpenEvoShield: Dual Non-Stationary Continual Defense for Open-World Multi-Agent System Attacks
OpenEvoShield claims to defend LLM multi-agent systems against evolving and novel attacks by combining asymmetric-rate continual learning with energy-based OOD detection.
-
Towards Domain-Generalized Open-Vocabulary Object Detection: A Progressive Domain-invariant Cross-modal Alignment Method
A progressive curriculum that trains open-vocabulary detectors on low-ambiguity, high-signal cross-modal alignments first improves robustness to visual domain shifts, with modest, test-tuned gains.
-
Causal Transfer in Medical Image Analysis
Causal Transfer Learning unifies structural causal models, invariant risk minimisation and counterfactuals with transfer learning to produce domain-robust medical image models.
-
$\Delta \mathrm{Energy}$: Optimizing Energy Change During Vision-Language Alignment Improves both OOD Detection and OOD Generalization
ΔEnergy, an energy-change OOD score for CLIP, and its EBM fine-tuning loss simultaneously improve OOD detection and covariate-shift generalization.
-
Domain-Shift-Aware Conformal Prediction for Large Language Models
DS-CP reweights calibration scores via embedding-based density ratios to improve conformal coverage under domain shift, but its stated guarantee requires a weight condition that the default configuration violates.
-
MorphGen: Morphology-Guided Representation Learning for Robust Single-Domain Generalization in Histopathological Cancer Classification
MorphGen uses supervised contrastive learning to align histopathology images with nuclear masks and applies SWA, reporting improved out-of-domain cancer classification accuracy on CAMELYON17, BCSS, and OCELOT.
-
Robust Invariant Representation Learning by Distribution Extrapolation
A new IRMv1 penalty based on extrapolated per-sample gradient magnitudes is proposed, with reported gains over IRM baselines on synthetic and vision benchmarks.
-
Simple Domain Generalization for Strong Pixel-Level Image Tampering Detection in Modern VLMs
A simple domain-generalization training recipe (balanced real/tampered batches, late injection of a companion VLM domain, low learning rate) yields large pixel-level tampering-localization gains on out-of-distribution...
-
Saving for the future: Enhancing generalization via partial logic regularization
PL-Reg adds a trainable mask and a defined/undefined classification loss to logic-based regularization, improving unknown-class accuracy across GCD, mDG+GCD, and CIL benchmarks.
-
Calibrated and Robust Foundation Models for Vision-Language and Medical Image Tasks Under Distribution Shift
StaRFM reuses the authors' earlier CalShift penalties, extends them to 3D medical segmentation with patch-wise and voxel-wise variants, and claims large gains that are not consistently supported by the paper's own tables.
-
Generalizing vision-language models to novel domains: A comprehensive survey
A survey of VLM generalization literature organized by transferred module, with benchmark tables and a review of multimodal LLMs.
-
Data Heterogeneity Modeling for Trustworthy Machine Learning
A survey that frames heterogeneity-aware machine learning as a paradigm spanning data collection, training, evaluation, and deployment, drawing mostly on the authors' prior results.
Discussion (0). Sign in to comment.