REVIEW 4 cited by
Identifying and Compensating for Feature Deviation in Imbalanced Deep Learning
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
Classifiers trained with class-imbalanced data are known to perform poorly on test data of the "minor" classes, of which we have insufficient training data. In this paper, we investigate learning a ConvNet classifier under such a scenario. We found that a ConvNet significantly over-fits the minor classes, which is quite opposite to traditional machine learning algorithms that often under-fit minor classes. We conducted a series of analysis and discovered the feature deviation phenomenon -- the learned ConvNet generates deviated features between the training and test data of minor classes -- which explains how over-fitting happens. To compensate for the effect of feature deviation which pushes test data toward low decision value regions, we propose to incorporate class-dependent temperatures (CDT) in training a ConvNet. CDT simulates feature deviation in the training phase, forcing the ConvNet to enlarge the decision values for minor-class data so that it can overcome real feature deviation in the test phase. We validate our approach on benchmark datasets and achieve promising performance. We hope that our insights can inspire new ways of thinking in resolving class-imbalanced deep learning.
Forward citations
Cited by 4 Pith papers
-
Lessons and Open Questions from a Unified Study of Camera-Trap Species Recognition Over Time
At a fixed camera-trap site, naively updating a recognition model on newly observed data frequently drops accuracy below the zero-shot baseline; LoRA with balanced softmax mostly fixes it, and post-processing closes m...
-
Addressing Imbalanced Domain-Incremental Learning through Dual-Balance Collaborative Experts
DCE trains frequency-aware experts with complementary losses plus a Gaussian-sampled dynamic selector, reporting SOTA accuracy on four imbalanced domain-incremental benchmarks.
-
Measuring Weak-to-Strong Legibility of Reasoning Models
Strong reasoning models need traces weaker models can digest; current efficiency metrics miss thoroughness and understate this weak-to-strong legibility requirement.
-
A Comprehensive Survey on Imbalanced Data Learning
A structured survey and benchmark that groups imbalanced data learning methods into data re-balancing, feature representation, training strategy, and ensemble learning.
Discussion (0). Continue with ORCID to comment.