Pith. sign in

REVIEW 9 cited by

All Models are Wrong, but Many are Useful: Learning a Variable's Importance by Studying an Entire Class of Prediction Models Simultaneously

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 1801.01489 v5 pith:HGITWHKL submitted 2018-01-04 stat.ME

classification stat.ME
keywords modelimportancemodelspredictionclassvariablebetaconditional
verification ladder T0 review T1 audit T2 compute T3 formal

Signed reviews

No signed human review yet.

0 comments
abstract

Variable importance (VI) tools describe how much covariates contribute to a prediction model's accuracy. However, important variables for one well-performing model (for example, a linear model $f(\mathbf{x})=\mathbf{x}^{T}\beta$ with a fixed coefficient vector $\beta$) may be unimportant for another model. In this paper, we propose model class reliance (MCR) as the range of VI values across all well-performing model in a prespecified class. Thus, MCR gives a more comprehensive description of importance by accounting for the fact that many prediction models, possibly of different parametric forms, may fit the data well. In the process of deriving MCR, we show several informative results for permutation-based VI estimates, based on the VI measures used in Random Forests. Specifically, we derive connections between permutation importance estimates for a single prediction model, U-statistics, conditional variable importance, conditional causal effects, and linear model coefficients. We then give probabilistic bounds for MCR, using a novel, generalizable technique. We apply MCR to a public data set of Broward County criminal records to study the reliance of recidivism prediction models on sex and race. In this application, MCR can be used to help inform VI for unknown, proprietary models.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 9 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Scaling Inherently Interpretable Language Models

    cs.CL 2026-08 conditional novelty 7.0 of 10

    Training a language model with a built-in concept bottleneck preserves compute-optimal scaling and yields interpretability metrics that improve with scale, demonstrated on an 8B causal diffusion model.

  2. Automatic detection of Ellerman bombs using Deep Learning

    astro-ph.SR 2025-05 conditional novelty 6.0 of 10

    Neural networks can automatically detect Ellerman bombs in high-resolution Hα data, but SDO/AIA four-passband intensity maps alone are insufficient for reliable detection.

  3. How Your Location Relates to Health: Variable Importance and Interpretable Machine Learning for Environmental and Sociodemographic Data

    cs.LG 2025-01 conditional novelty 6.0 of 10

    Using the MEDSAT dataset, the authors find that NO2 is a robust global predictor for asthma, hypertension, and anxiety prescriptions, alongside outcome-specific factors such as occupation, marriage, and vegetation, wi...

  4. Shapley Decomposition of R-Squared in Machine Learning Models

    stat.ME 2019-08 conditional novelty 6.0 of 10

    A Shapley-value based normalization of residual variance increases yields feature-level R2 shares that sum to the model's overall R2.

  5. Efficient computation of counterfactual explanations of LVQ models

    cs.LG 2019-08 conditional novelty 6.0 of 10

    Counterfactual explanations for LVQ classifiers can be computed by solving closed-form linear, quadratic, or non-convex QCQP programs derived from the nearest-prototype rule, yielding faster and closer counterfactuals...

  6. AI-Spectra: A Visual Dashboard for Model Multiplicity to Enhance Informed and Transparent Decision-Making

    cs.HC 2024-11 conditional novelty 5.0 of 10

    The paper presents a robot-face dashboard for comparing multiple AI models, and demonstrates it on MNIST digit classifiers.

  7. Interpretable Event Diagnosis in Water Distribution Networks

    cs.AI 2025-05 conditional novelty 4.0 of 10

    Counterfactual fingerprints computed from sensor residuals can distinguish leakages from sensor faults in water distribution networks and provide operators with a contrastive explanation.

  8. On the global feature importance for interpretable and trustworthy heat demand forecasting

    cs.LG 2026-08 conditional novelty 3.0 of 10

    A case study comparing XGBoost intrinsic feature importances, partial dependence, ALE, and SHAP for a district heating heat demand forecasting model.

  9. Explainable AI the Latest Advancements and New Trends

    cs.AI 2025-05 conditional novelty 2.0 of 10

    A survey of explainable AI methods and a speculative proposal that meta-reasoning in reward space can explain AI decisions.

Pith tools