<scp>B</scp> <sup>3</sup> : Backdoor Attacks against Black-box Machine Learning Models-Reference-Cited by-同舟云学术

B ³ : Backdoor Attacks against Black-box Machine Learning Models

Published:2023-08-08 Issue:4 Volume:26 Page:1-24
ISSN:2471-2566
Container-title:ACM Transactions on Privacy and Security
language:en
Short-container-title:ACM Trans. Priv. Secur.

Author:

Gong Xueluan¹^ORCID,Chen Yanjiao²^ORCID,Yang Wenbin³^ORCID,Huang Huayang³^ORCID,Wang Qian³^ORCID

Affiliation:

1. School of Computer Science, Wuhan University, China

2. College of Electrical Engineering, Zhejiang University, China

3. School of Cyber Science and Engineering, Wuhan University, China

Abstract

Backdoor attacks aim to inject backdoors to victim machine learning models during training time, such that the backdoored model maintains the prediction power of the original model towards clean inputs and misbehaves towards backdoored inputs with the trigger. The reason for backdoor attacks is that resource-limited users usually download sophisticated models from model zoos or query the models from MLaaS rather than training a model from scratch, thus a malicious third party has a chance to provide a backdoored model. In general, the more precious the model provided (i.e., models trained on rare datasets), the more popular it is with users. In this article, from a malicious model provider perspective, we propose a black-box backdoor attack, named B 3 , where neither the rare victim model (including the model architecture, parameters, and hyperparameters) nor the training data is available to the adversary. To facilitate backdoor attacks in the black-box scenario, we design a cost-effective model extraction method that leverages a carefully constructed query dataset to steal the functionality of the victim model with a limited budget. As the trigger is key to successful backdoor attacks, we develop a novel trigger generation algorithm that intensifies the bond between the trigger and the targeted misclassification label through the neuron with the highest impact on the targeted label. Extensive experiments have been conducted on various simulated deep learning models and the commercial API of Alibaba Cloud Compute Service. We demonstrate that B 3 has a high attack success rate and maintains high prediction accuracy for benign inputs. It is also shown that B 3 is robust against state-of-the-art defense strategies against backdoor attacks, such as model pruning and NC.

Funder

National Key R&D Program of China

NSFC

Publisher

Association for Computing Machinery (ACM)

Subject

Safety, Risk, Reliability and Quality,General Computer Science

Link

https://dl.acm.org/doi/pdf/10.1145/3605212

Reference52 articles.

1. Eitan Borgnia, Valeriia Cherepanova, Liam Fowl, Amin Ghiasi, Jonas Geiping, Micah Goldblum, Tom Goldstein, and Arjun Gupta. 2021. Strong data augmentation sanitizes poisoning and backdoor attacks without an accuracy tradeoff. In IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP). IEEE, 3855–3859.

2. Detecting backdoor attacks on deep neural networks by activation clustering;Chen Bryant;arXiv preprint arXiv:1811.03728,2018

3. Targeted backdoor attacks on deep learning systems using data poisoning;Chen Xinyun;arXiv preprint arXiv:1712.05526,2017

4. Backdoor attacks and defenses for deep neural networks in outsourced cloud environments;Chen Yanjiao;IEEE Netw.,2020

5. MARNet: Backdoor attacks against cooperative multi-agent reinforcement learning;Chen Yanjiao;IEEE Trans. Depend. Secure Comput.,2022

Cited by 2 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. Machine learning security and privacy: a review of threats and countermeasures;EURASIP Journal on Information Security;2024-04-23

2. SGBA: A stealthy scapegoat backdoor attack against deep neural networks;Computers & Security;2024-01