REVIEW 4 cited by
Sanity Checks Revisited: An Exploration to Repair the Model Parameter Randomisation Test
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
The Model Parameter Randomisation Test (MPRT) is widely acknowledged in the eXplainable Artificial Intelligence (XAI) community for its well-motivated evaluative principle: that the explanation function should be sensitive to changes in the parameters of the model function. However, recent works have identified several methodological caveats for the empirical interpretation of MPRT. To address these caveats, we introduce two adaptations to the original MPRT -- Smooth MPRT and Efficient MPRT, where the former minimises the impact that noise has on the evaluation results through sampling and the latter circumvents the need for biased similarity measurements by re-interpreting the test through the explanation's rise in complexity, after full parameter randomisation. Our experimental results demonstrate that these proposed variants lead to improved metric reliability, thus enabling a more trustworthy application of XAI methods.
Forward citations
Cited by 4 Pith papers
-
Value bounds and Convergence Analysis for Averages of LRP attributions
Averaged LRP-beta attributions have Hoeffding convergence bounds independent of weight norms, unlike gradient-based explanations.
-
Explainable embeddings with Distance Explainer
Distance Explainer adapts RISE-style random masking to produce attribution maps for pairwise distances in arbitrary embedding spaces, using rank-based mask selection and a two-sided mirror mode.
-
Complying with the EU AI Act: Innovations in Explainable and User-Centric Hand Gesture Recognition
A radar gesture recognition system with user-specific VAE thresholding and experience replay calibration reports 11.50% more flagged anomalies and 15.17% better user adaptation, plus a 97.5% characterization rate that...
-
xai_evals : A Framework for Evaluating Post-Hoc Local Explanation Methods
A technical report introducing xai_evals, a Python package that wraps existing explainability and metric libraries without adding new methods or validated results.
Discussion (0). Continue with ORCID to comment.