Curse or Redemption? How Data Heterogeneity Affects the Robustness of Federated Learning-Reference-Cited by-同舟云学术

Curse or Redemption? How Data Heterogeneity Affects the Robustness of Federated Learning

Published:2021-05-18 Issue:12 Volume:35 Page:10807-10814
ISSN:2374-3468
Container-title:Proceedings of the AAAI Conference on Artificial Intelligence
language:
Short-container-title:AAAI

Author:

Zawad Syed,Ali Ahsan,Chen Pin-Yu,Anwar Ali,Zhou Yi,Baracaldo Nathalie,Tian Yuan,Yan Feng

Abstract

Data heterogeneity has been identified as one of the key features in federated learning but often overlooked in the lens of robustness to adversarial attacks. This paper focuses on characterizing and understanding its impact on backdooring attacks in federated learning through comprehensive experiments using synthetic and the LEAF benchmarks. The initial impression driven by our experimental results suggests that data heterogeneity is the dominant factor in the effectiveness of attacks and it may be a redemption for defending against backdooring as it makes the attack less efficient, more challenging to design effective attack strategies, and the attack result also becomes less predictable. However, with further investigations, we found data heterogeneity is more of a curse than a redemption as the attack effectiveness can be significantly boosted by simply adjusting the client-side backdooring timing. More importantly, data heterogeneity may result in overfitting at the local training of benign clients, which can be utilized by attackers to disguise themselves and fool skewed-feature based defenses. In addition, effective attack strategies can be made by adjusting attack data distribution. Finally, we discuss the potential directions of defending the curses brought by data heterogeneity. The results and lessons learned from our extensive experiments and analysis offer new insights for designing robust federated learning methods and systems.

Publisher

Association for the Advancement of Artificial Intelligence (AAAI)

Subject

General Medicine

Cited by 25 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. WBSP: Addressing stragglers in distributed machine learning with worker-busy synchronous parallel;Parallel Computing;2024-09

2. ODE: An Online Data Selection Framework for Federated Learning With Limited Storage;IEEE/ACM Transactions on Networking;2024-08

3. Improving Federated Learning Through Low-Entropy Client Sampling Based on Learned High-Level Features;2024 IEEE 17th International Conference on Cloud Computing (CLOUD);2024-07-07

4. Precision Guided Approach to Mitigate Data Poisoning Attacks in Federated Learning;Proceedings of the Fourteenth ACM Conference on Data and Application Security and Privacy;2024-06-19

5. Evaluation of Data Heterogeneity in FL Environment;2024 XXVII International Conference on Soft Computing and Measurements (SCM);2024-05-22