Pith. sign in

REVIEW 2 cited by

A Survey of Privacy-Preserving Model Explanations: Privacy Risks, Attacks, and Countermeasures

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2404.00673 v2 pith:O6EKMV6F submitted 2024-03-31 cs.CR cs.AIcs.CYcs.LG

classification cs.CRcs.AIcs.CYcs.LG
keywords privacyresearchexplanationsattackscountermeasuresmodelsurveyanalysis
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

As the adoption of explainable AI (XAI) continues to expand, the urgency to address its privacy implications intensifies. Despite a growing corpus of research in AI privacy and explainability, there is little attention on privacy-preserving model explanations. This article presents the first thorough survey about privacy attacks on model explanations and their countermeasures. Our contribution to this field comprises a thorough analysis of research papers with a connected taxonomy that facilitates the categorisation of privacy attacks and countermeasures based on the targeted explanations. This work also includes an initial investigation into the causes of privacy leaks. Finally, we discuss unresolved issues and prospective research directions uncovered in our analysis. This survey aims to be a valuable resource for the research community and offers clear insights for those new to this domain. To support ongoing research, we have established an online resource repository, which will be continuously updated with new and relevant findings. Interested readers are encouraged to access our repository at https://github.com/tamlhp/awesome-privex.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 2 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Differentially Private Explanations for Clusters

    cs.CR 2025-06 conditional novelty 7.0 of 10

    DPClustX privately selects the most informative attributes for each cluster and releases noisy histograms only for those attributes, providing differentially private explanations of clustering results.

  2. Reconciling Privacy and Explainability in High-Stakes: A Systematic Inquiry

    cs.CR 2024-12 conditional novelty 6.0 of 10

    Gradient-based explainers yield almost uncorrelated attributions on DP-trained chest X-ray models, so the authors recommend privatizing explanations from a non-private model instead.

Pith tools