Pith. sign in

REVIEW 2 cited by

On the Impact of Knowledge Distillation for Model Interpretability

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2305.15734 v1 pith:YR67NIYU submitted 2023-05-25 cs.LG cs.AIcs.CV

classification cs.LGcs.AIcs.CV
keywords interpretabilitymodelinformationmodelsclass-similaritydifferentdistillationknowledge
verification ladder T0 review T1 audit T2 compute T3 formal

Signed reviews

No signed human review yet.

0 comments
read the original abstract

Several recent studies have elucidated why knowledge distillation (KD) improves model performance. However, few have researched the other advantages of KD in addition to its improving model performance. In this study, we have attempted to show that KD enhances the interpretability as well as the accuracy of models. We measured the number of concept detectors identified in network dissection for a quantitative comparison of model interpretability. We attributed the improvement in interpretability to the class-similarity information transferred from the teacher to student models. First, we confirmed the transfer of class-similarity information from the teacher to student model via logit distillation. Then, we analyzed how class-similarity information affects model interpretability in terms of its presence or absence and degree of similarity information. We conducted various quantitative and qualitative experiments and examined the results on different datasets, different KD methods, and according to different measures of interpretability. Our research showed that KD models by large models could be used more reliably in various fields.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 2 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. ReaLM: Reflection-Enhanced Autonomous Reasoning with Small Language Models

    cs.CL 2025-08 conditional novelty 6.0 of 10

    ReaLM trains small language models to learn from both right and wrong reasoning chains, then fades the chains out so the model reasons independently, improving benchmark accuracy.

  2. Taming Vision-Language Models for Medical Image Analysis: A Comprehensive Review

    eess.IV 2025-06 conditional novelty 3.0 of 10

    A survey that classifies vision-language model adaptation for medical imaging into five strategies across eleven tasks, with challenges and future directions.

Pith tools