Pith. sign in

REVIEW 1 cited by

Ultra-marginal Feature Importance: Learning from Data with Causal Guarantees

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2204.09938 v5 pith:3LZ7QDNZ submitted 2022-04-21 stat.ML cs.ITcs.LGmath.ITstat.AP

classification stat.MLcs.ITcs.LGmath.ITstat.AP
keywords datafeatureimportancelearningcausalumfiaxiomsrelationships
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Scientists frequently prioritize learning from data rather than training the best possible model; however, research in machine learning often prioritizes the latter. Marginal contribution feature importance (MCI) was developed to break this trend by providing a useful framework for quantifying the relationships in data. In this work, we aim to improve upon the theoretical properties, performance, and runtime of MCI by introducing ultra-marginal feature importance (UMFI), which uses dependence removal techniques from the AI fairness literature as its foundation. We first propose axioms for feature importance methods that seek to explain the causal and associative relationships in data, and we prove that UMFI satisfies these axioms under basic assumptions. We then show on real and simulated data that UMFI performs better than MCI, especially in the presence of correlated interactions and unrelated features, while partially learning the structure of the causal graph and reducing the exponential runtime of MCI to super-linear.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Information-theoretic Quantification of High-order Feature Effects in Classification Problems

    cs.LG 2025-07 conditional novelty 5.0 of 10

    A kNN-based CMI estimator applied to classification decomposes feature importance into unique, redundant, and synergistic components, validated on synthetic and real data.

Pith tools