REVIEW 2 cited by
Conformal Prediction Sets Can Cause Disparate Impact
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
Conformal Prediction Sets Can Cause Disparate Impact
read the original abstract
Conformal prediction is a statistically rigorous method for quantifying uncertainty in models by having them output sets of predictions, with larger sets indicating more uncertainty. However, prediction sets are not inherently actionable; many applications require a single output to act on, not several. To overcome this limitation, prediction sets can be provided to a human who then makes an informed decision. In any such system it is crucial to ensure the fairness of outcomes across protected groups, and researchers have proposed that Equalized Coverage be used as the standard for fairness. By conducting experiments with human participants, we demonstrate that providing prediction sets can lead to disparate impact in decisions. Disquietingly, we find that providing sets that satisfy Equalized Coverage actually increases disparate impact compared to marginal coverage. Instead of equalizing coverage, we propose to equalize set sizes across groups which empirically leads to lower disparate impact.
Forward citations
Cited by 2 Pith papers
-
A Framework for Measuring Appropriate Reliance on Set-Valued AI Advice
Proposes new metrics (correct reliance rates for classification; quantity and quality of reliance for regression) to evaluate appropriate reliance on set-valued AI advice.
-
Socio-Conformal Calibration in Complex Survey Data: Marginal Validity Is Not Enough for Subgroup Reliability
Standard conformal prediction gives nominal overall coverage on Pew survey data but leaves ~13-point weighted gaps across race-education subgroups, and group-specific Mondrian calibration does not reliably close them.
discussion (0)
Sign in with ORCID, Apple, or X to comment. Anyone can read and Pith papers without signing in.