article

Incremental Learning, Incremental Backdoor Threats

  • IEEE Transactions on Dependable and Secure Computing
  • IEEE Computer Society
Research footprint

At a glance

Citations
18
References
58
Comments
0
Paper overview

Abstract

Class incremental learning from a pre-trained DNN model is gaining lots of popularity. Unfortunately, the pre-trained model also introduces a new attack vector, which enables an adversary to inject a backdoor into it and further compromise the downstream models learned from it. Prior works proposed backdoor attacks against the pre-trained models in the transfer learning scenario. However, they become less effective when the adversary does not have the knowledge of the downstream tasks or new data, which is more practical and considered in this paper. To this end, we design the first latent backdoor attacks against incremental learning. We propose two novel techniques, which can effectively and stealthily embed a backdoor into the pre-trained model. Such backdoor can only be activated when the pre-trained model is extended to a downstream model with incremental learning. It has a very high attack success rate, and is able to bypass existing backdoor detection approaches. Extensive experiments confirm the effectiveness of our attacks over different datasets and incremental learning methods, as well as strong robustness against state-of-the-art backdoor defense mechanisms includingNeural Cleanse,Fine-PruningandSTRIP.

Record transparency

Publication details

DOI
10.1109/tdsc.2022.3201234
OpenAlex
W4293812146
Document type
article
Language
EN
Source
IEEE Transactions on Dependable and Secure Computing
Last metadata update
Community

Comments

Log in to join the discussion.

  1. No comments yet. Start the discussion.