Pith. sign in

REVIEW 1 cited by

FAIR-Ensemble: When Fairness Naturally Emerges From Deep Ensembling

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2303.00586 v2 pith:VJSTI6M6 submitted 2023-03-01 stat.ML cs.AIcs.CVcs.CYcs.LG

classification stat.MLcs.AIcs.CVcs.CYcs.LG
keywords ensemblingevenfairnessmodelssimpledeepdnnsemerges
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
abstract

Ensembling multiple Deep Neural Networks (DNNs) is a simple and effective way to improve top-line metrics and to outperform a larger single model. In this work, we go beyond top-line metrics and instead explore the impact of ensembling on subgroup performances. Surprisingly, we observe that even with a simple homogeneous ensemble -- all the individual DNNs share the same training set, architecture, and design choices -- the minority group performance disproportionately improves with the number of models compared to the majority group, i.e. fairness naturally emerges from ensembling. Even more surprising, we find that this gain keeps occurring even when a large number of models is considered, e.g. $20$, despite the fact that the average performance of the ensemble plateaus with fewer models. Our work establishes that simple DNN ensembles can be a powerful tool for alleviating disparate impact from DNN classifiers, thus curbing algorithmic harm. We also explore why this is the case. We find that even in homogeneous ensembles, varying the sources of stochasticity through parameter initialization, mini-batch sampling, and data-augmentation realizations, results in different fairness outcomes.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Fairness of Deep Ensembles: On the interplay between per-group task difficulty and under-representation

    cs.LG 2025-01 conditional novelty 6.0 of 10

    Homogeneous deep ensembles shrink accuracy gaps between demographic groups without lowering overall accuracy, and the optimal training-data balance shifts toward the harder group when per-group task difficulty differs.

Pith tools