Pith. sign in

REVIEW

Interpretability Benchmark for Evaluating Spatial Misalignment of Prototypical Parts Explanations

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2308.08162 v1 pith:55JN2X4H submitted 2023-08-16 cs.CV cs.AIcs.LG

classification cs.CVcs.AIcs.LG
keywords benchmarkmisalignmentcompensationinterpretabilitypartsprototypicalregionspatial
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Prototypical parts-based networks are becoming increasingly popular due to their faithful self-explanations. However, their similarity maps are calculated in the penultimate network layer. Therefore, the receptive field of the prototype activation region often depends on parts of the image outside this region, which can lead to misleading interpretations. We name this undesired behavior a spatial explanation misalignment and introduce an interpretability benchmark with a set of dedicated metrics for quantifying this phenomenon. In addition, we propose a method for misalignment compensation and apply it to existing state-of-the-art models. We show the expressiveness of our benchmark and the effectiveness of the proposed compensation methodology through extensive empirical studies.

Discussion (0). Continue with ORCID to comment.

Pith tools