Pith. sign in

REVIEW 1 cited by

Can we Agree? On the Rash\=omon Effect and the Reliability of Post-Hoc Explainable AI

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2308.07247 v1 pith:Z444PM6V submitted 2023-08-14 cs.LG cs.AIstat.ML

classification cs.LGcs.AIstat.ML
keywords explanationsmodelsdataknowledgeomonrashagreementeffect
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

The Rash\=omon effect poses challenges for deriving reliable knowledge from machine learning models. This study examined the influence of sample size on explanations from models in a Rash\=omon set using SHAP. Experiments on 5 public datasets showed that explanations gradually converged as the sample size increased. Explanations from <128 samples exhibited high variability, limiting reliable knowledge extraction. However, agreement between models improved with more data, allowing for consensus. Bagging ensembles often had higher agreement. The results provide guidance on sufficient data to trust explanations. Variability at low samples suggests that conclusions may be unreliable without validation. Further work is needed with more model types, data domains, and explanation methods. Testing convergence in neural networks and with model-specific explanation methods would be impactful. The approaches explored here point towards principled techniques for eliciting knowledge from ambiguous models.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Systemizing Multiplicity: The Curious Case of Arbitrariness in Machine Learning

    cs.LG 2025-01 conditional novelty 6.0 of 10

    A systematic review of 80 papers on model multiplicity, with a new taxonomy of developer choices and a formal distinction between multiplicity, uncertainty, and variance.

Pith tools