Pith. sign in

REVIEW 1 cited by

Causal Shapley Values: Exploiting Causal Knowledge to Explain Individual Predictions of Complex Models

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2011.01625 v1 pith:PRYR6AYS submitted 2020-11-03 cs.AI cs.LG

classification cs.AIcs.LG
keywords valuesshapleycausalwhenassumptioncomplexcomputingdesirable
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Shapley values underlie one of the most popular model-agnostic methods within explainable artificial intelligence. These values are designed to attribute the difference between a model's prediction and an average baseline to the different features used as input to the model. Being based on solid game-theoretic principles, Shapley values uniquely satisfy several desirable properties, which is why they are increasingly used to explain the predictions of possibly complex and highly non-linear machine learning models. Shapley values are well calibrated to a user's intuition when features are independent, but may lead to undesirable, counterintuitive explanations when the independence assumption is violated. In this paper, we propose a novel framework for computing Shapley values that generalizes recent work that aims to circumvent the independence assumption. By employing Pearl's do-calculus, we show how these 'causal' Shapley values can be derived for general causal graphs without sacrificing any of their desirable properties. Moreover, causal Shapley values enable us to separate the contribution of direct and indirect effects. We provide a practical implementation for computing causal Shapley values based on causal chain graphs when only partial information is available and illustrate their utility on a real-world example.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Unifying Attribution-Based Explanations Using Functional Decomposition

    cs.LG 2024-12 reject novelty 6.0 of 10

    A unification framework for XAI attribution methods whose core canonical decomposition theorem fails because the components sum to the fully removed function rather than to the original function.

Pith tools