Pith. sign in

REVIEW 4 cited by

Identifying and Compensating for Feature Deviation in Imbalanced Deep Learning

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2001.01385 v4 pith:CGZRRLHE submitted 2020-01-06 cs.LG cs.CVstat.ML

classification cs.LGcs.CVstat.ML
keywords dataconvnetdeviationfeatureclasseslearningminortest
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Classifiers trained with class-imbalanced data are known to perform poorly on test data of the "minor" classes, of which we have insufficient training data. In this paper, we investigate learning a ConvNet classifier under such a scenario. We found that a ConvNet significantly over-fits the minor classes, which is quite opposite to traditional machine learning algorithms that often under-fit minor classes. We conducted a series of analysis and discovered the feature deviation phenomenon -- the learned ConvNet generates deviated features between the training and test data of minor classes -- which explains how over-fitting happens. To compensate for the effect of feature deviation which pushes test data toward low decision value regions, we propose to incorporate class-dependent temperatures (CDT) in training a ConvNet. CDT simulates feature deviation in the training phase, forcing the ConvNet to enlarge the decision values for minor-class data so that it can overcome real feature deviation in the test phase. We validate our approach on benchmark datasets and achieve promising performance. We hope that our insights can inspire new ways of thinking in resolving class-imbalanced deep learning.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 4 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Lessons and Open Questions from a Unified Study of Camera-Trap Species Recognition Over Time

    cs.CV 2026-03 conditional novelty 6.0 of 10

    At a fixed camera-trap site, naively updating a recognition model on newly observed data frequently drops accuracy below the zero-shot baseline; LoRA with balanced softmax mostly fixes it, and post-processing closes m...

  2. Addressing Imbalanced Domain-Incremental Learning through Dual-Balance Collaborative Experts

    cs.LG 2025-07 conditional novelty 5.0 of 10

    DCE trains frequency-aware experts with complementary losses plus a Gaussian-sampled dynamic selector, reporting SOTA accuracy on four imbalanced domain-incremental benchmarks.

  3. Measuring Weak-to-Strong Legibility of Reasoning Models

    cs.MA 2026-03 unverdicted novelty 4.0 of 10

    Strong reasoning models need traces weaker models can digest; current efficiency metrics miss thoroughness and understate this weak-to-strong legibility requirement.

  4. A Comprehensive Survey on Imbalanced Data Learning

    cs.LG 2025-02 conditional novelty 3.0 of 10

    A structured survey and benchmark that groups imbalanced data learning methods into data re-balancing, feature representation, training strategy, and ensemble learning.

Pith tools