Pith. sign in

REVIEW 2 cited by

Conformal Prediction Sets Can Cause Disparate Impact

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2410.01888 v2 pith:MOKKIHKE submitted 2024-10-02 cs.LG stat.ML

Conformal Prediction Sets Can Cause Disparate Impact

classification cs.LG stat.ML
keywords setspredictioncoveragedisparateimpactacrossconformalequalized
verification ladder T0 review T1 audit T2 compute T3 formal T4 reserved
0 comments
read the original abstract

Conformal prediction is a statistically rigorous method for quantifying uncertainty in models by having them output sets of predictions, with larger sets indicating more uncertainty. However, prediction sets are not inherently actionable; many applications require a single output to act on, not several. To overcome this limitation, prediction sets can be provided to a human who then makes an informed decision. In any such system it is crucial to ensure the fairness of outcomes across protected groups, and researchers have proposed that Equalized Coverage be used as the standard for fairness. By conducting experiments with human participants, we demonstrate that providing prediction sets can lead to disparate impact in decisions. Disquietingly, we find that providing sets that satisfy Equalized Coverage actually increases disparate impact compared to marginal coverage. Instead of equalizing coverage, we propose to equalize set sizes across groups which empirically leads to lower disparate impact.

discussion (0)

Sign in with ORCID, Apple, or X to comment. Anyone can read and Pith papers without signing in.

Forward citations

Cited by 2 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score.

  1. A Framework for Measuring Appropriate Reliance on Set-Valued AI Advice

    cs.AI 2026-06 unverdicted novelty 7.0

    Proposes new metrics (correct reliance rates for classification; quantity and quality of reliance for regression) to evaluate appropriate reliance on set-valued AI advice.

  2. Socio-Conformal Calibration in Complex Survey Data: Marginal Validity Is Not Enough for Subgroup Reliability

    stat.ME 2026-05 unverdicted novelty 5.0

    Standard conformal prediction gives nominal overall coverage on Pew survey data but leaves ~13-point weighted gaps across race-education subgroups, and group-specific Mondrian calibration does not reliably close them.