Pith. sign in

REVIEW 4 cited by

Comprehensive Privacy Analysis of Deep Learning: Passive and Active White-box Inference Attacks against Centralized and Federated Learning

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 1812.00910 v2 pith:633F5YUT submitted 2018-12-03 stat.ML cs.CRcs.LG

classification stat.MLcs.CRcs.LG
keywords inferenceattackslearningdeepmodelswhite-boxprivacytraining
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Deep neural networks are susceptible to various inference attacks as they remember information about their training data. We design white-box inference attacks to perform a comprehensive privacy analysis of deep learning models. We measure the privacy leakage through parameters of fully trained models as well as the parameter updates of models during training. We design inference algorithms for both centralized and federated learning, with respect to passive and active inference attackers, and assuming different adversary prior knowledge. We evaluate our novel white-box membership inference attacks against deep learning algorithms to trace their training data records. We show that a straightforward extension of the known black-box attacks to the white-box setting (through analyzing the outputs of activation functions) is ineffective. We therefore design new algorithms tailored to the white-box setting by exploiting the privacy vulnerabilities of the stochastic gradient descent algorithm, which is the algorithm used to train deep neural networks. We investigate the reasons why deep learning models may leak information about their training data. We then show that even well-generalized models are significantly susceptible to white-box membership inference attacks, by analyzing state-of-the-art pre-trained and publicly available models for the CIFAR dataset. We also show how adversarial participants, in the federated learning setting, can successfully run active membership inference attacks against other participants, even when the global model achieves high prediction accuracies.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 4 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Stealth by Conformity: Evading Robust Aggregation through Adaptive Poisoning

    cs.CR 2025-09 conditional novelty 6.0 of 10

    An adaptive federated-learning backdoor attack uses membership-inference feedback on the global model to keep malicious updates statistically similar to benign ones, evading nine robust aggregation defenses in two ima...

  2. Hyperparameters in Score-Based Membership Inference Attacks

    cs.LG 2025-02 accept novelty 6.0 of 10

    A new shadow-model hyperparameter selection method (KL-LiRA) makes membership inference attacks nearly as effective without knowing target hyperparameters, and training-data-based hyperparameter tuning shows no detect...

  3. Code Division Modulation Layers Against Forgetting and Inference in Continual Gait Identification

    cs.MM 2026-07 conditional novelty 5.0 of 10

    A fixed random sign-pattern layer per task preserves continual-learning accuracy and blocks confidence-based membership inference for attackers who lack the code.

  4. ZKP-FedEval: Verifiable and Privacy-Preserving Federated Evaluation using Zero-Knowledge Proofs

    cs.LG 2025-07 reject novelty 3.0 of 10

    A threshold-based ZKP protocol for federated evaluation is proposed, but the implemented circuit only checks the threshold, not the loss computation, leaving the central guarantee unsupported.

Pith tools