Pith. sign in

REVIEW 2 cited by

Get Rid Of Your Trail: Remotely Erasing Backdoors in Federated Learning

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2304.10638 v1 pith:PMYE6V6Z submitted 2023-04-20 cs.LG cs.CR

classification cs.LGcs.CR
keywords adversariesbackdoorscentralizedmodelattacksbackdoorlearningacross
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Federated Learning (FL) enables collaborative deep learning training across multiple participants without exposing sensitive personal data. However, the distributed nature of FL and the unvetted participants' data makes it vulnerable to backdoor attacks. In these attacks, adversaries inject malicious functionality into the centralized model during training, leading to intentional misclassifications for specific adversary-chosen inputs. While previous research has demonstrated successful injections of persistent backdoors in FL, the persistence also poses a challenge, as their existence in the centralized model can prompt the central aggregation server to take preventive measures to penalize the adversaries. Therefore, this paper proposes a methodology that enables adversaries to effectively remove backdoors from the centralized model upon achieving their objectives or upon suspicion of possible detection. The proposed approach extends the concept of machine unlearning and presents strategies to preserve the performance of the centralized model and simultaneously prevent over-unlearning of information unrelated to backdoor patterns, making the adversaries stealthy while removing backdoors. To the best of our knowledge, this is the first work that explores machine unlearning in FL to remove backdoors to the benefit of adversaries. Exhaustive evaluation considering image classification scenarios demonstrates the efficacy of the proposed method in efficient backdoor removal from the centralized model, injected by state-of-the-art attacks across multiple configurations.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 2 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Mind the Cost of Scaffold! Benign Clients May Even Become Accomplices of Backdoor Attack

    cs.LG 2024-11 conditional novelty 5.0 of 10

    BadSFL is a backdoor attack that exploits Scaffold's control variate to make benign clients amplify and preserve a planted backdoor in non-IID federated learning.

  2. Upcycling Noise for Federated Unlearning

    cs.LG 2024-12 reject novelty 4.0 of 10

    FUI claims to erase a client's data in differentially private federated learning by retracting the local model and adding calibrated Gaussian noise to mimic retraining.

Pith tools