A Tale of Evil Twins: Adversarial Inputs versus Poisoned Models

Pang, Ren; Shen, Hua; Zhang, Xinyang; Ji, Shouling; Vorobeychik, Yevgeniy; Luo, Xiapu; Liu, Alex; Wang, Ting

doi:10.1145/3372297.3417253

Computer Science > Machine Learning

arXiv:1911.01559 (cs)

[Submitted on 5 Nov 2019 (v1), last revised 21 Nov 2020 (this version, v3)]

Title:A Tale of Evil Twins: Adversarial Inputs versus Poisoned Models

Authors:Ren Pang, Hua Shen, Xinyang Zhang, Shouling Ji, Yevgeniy Vorobeychik, Xiapu Luo, Alex Liu, Ting Wang

View PDF

Abstract:Despite their tremendous success in a range of domains, deep learning systems are inherently susceptible to two types of manipulations: adversarial inputs -- maliciously crafted samples that deceive target deep neural network (DNN) models, and poisoned models -- adversely forged DNNs that misbehave on pre-defined inputs. While prior work has intensively studied the two attack vectors in parallel, there is still a lack of understanding about their fundamental connections: what are the dynamic interactions between the two attack vectors? what are the implications of such interactions for optimizing existing attacks? what are the potential countermeasures against the enhanced attacks? Answering these key questions is crucial for assessing and mitigating the holistic vulnerabilities of DNNs deployed in realistic settings.
Here we take a solid step towards this goal by conducting the first systematic study of the two attack vectors within a unified framework. Specifically, (i) we develop a new attack model that jointly optimizes adversarial inputs and poisoned models; (ii) with both analytical and empirical evidence, we reveal that there exist intriguing "mutual reinforcement" effects between the two attack vectors -- leveraging one vector significantly amplifies the effectiveness of the other; (iii) we demonstrate that such effects enable a large design spectrum for the adversary to enhance the existing attacks that exploit both vectors (e.g., backdoor attacks), such as maximizing the attack evasiveness with respect to various detection methods; (iv) finally, we discuss potential countermeasures against such optimized attacks and their technical challenges, pointing to several promising research directions.

Comments:	Accepted as a full paper at ACM CCS 2020
Subjects:	Machine Learning (cs.LG); Cryptography and Security (cs.CR); Machine Learning (stat.ML)
Cite as:	arXiv:1911.01559 [cs.LG]
	(or arXiv:1911.01559v3 [cs.LG] for this version)
	https://doi.org/10.48550/arXiv.1911.01559
Related DOI:	https://doi.org/10.1145/3372297.3417253

Submission history

From: Ren Pang [view email]
[v1] Tue, 5 Nov 2019 01:32:57 UTC (2,598 KB)
[v2] Fri, 15 May 2020 22:14:26 UTC (5,178 KB)
[v3] Sat, 21 Nov 2020 04:54:13 UTC (4,718 KB)

Computer Science > Machine Learning

Title:A Tale of Evil Twins: Adversarial Inputs versus Poisoned Models

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Machine Learning

Title:A Tale of Evil Twins: Adversarial Inputs versus Poisoned Models

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators