Computer Science > Computer Vision and Pattern Recognition

arXiv:2405.16226 (cs)

[Submitted on 25 May 2024 (v1), last revised 25 Sep 2024 (this version, v3)]

Title:Detecting Adversarial Data via Perturbation Forgery

Authors:Qian Wang, Chen Li, Yuchen Luo, Hefei Ling, Ping Li, Jiazhong Chen, Shijuan Huang, Ning Yu

Abstract:As a defense strategy against adversarial attacks, adversarial detection aims to identify and filter out adversarial data from the data flow based on discrepancies in distribution and noise patterns between natural and adversarial data. Although previous detection methods achieve high performance in detecting gradient-based adversarial attacks, new attacks based on generative models with imbalanced and anisotropic noise patterns evade detection. Even worse, existing techniques either necessitate access to attack data before deploying a defense or incur a significant time cost for inference, rendering them impractical for defending against newly emerging attacks that are unseen by defenders. In this paper, we explore the proximity relationship between adversarial noise distributions and demonstrate the existence of an open covering for them. By learning to distinguish this open covering from the distribution of natural data, we can develop a detector with strong generalization capabilities against all types of adversarial attacks. Based on this insight, we heuristically propose Perturbation Forgery, which includes noise distribution perturbation, sparse mask generation, and pseudo-adversarial data production, to train an adversarial detector capable of detecting unseen gradient-based, generative-model-based, and physical adversarial attacks, while remaining agnostic to any specific models. Comprehensive experiments conducted on multiple general and facial datasets, with a wide spectrum of attacks, validate the strong generalization of our method.

Subjects:	Computer Vision and Pattern Recognition (cs.CV); Machine Learning (cs.LG)
Cite as:	arXiv:2405.16226 [cs.CV]
	(or arXiv:2405.16226v3 [cs.CV] for this version)
	https://doi.org/10.48550/arXiv.2405.16226

Submission history

From: Qian Wang [view email]
[v1] Sat, 25 May 2024 13:34:16 UTC (1,689 KB)
[v2] Sat, 24 Aug 2024 15:00:36 UTC (4,443 KB)
[v3] Wed, 25 Sep 2024 00:09:58 UTC (4,443 KB)

Computer Science > Computer Vision and Pattern Recognition

Title:Detecting Adversarial Data via Perturbation Forgery

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computer Vision and Pattern Recognition

Title:Detecting Adversarial Data via Perturbation Forgery

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators