Universal Distributional Decision-based Black-box Adversarial Attack with Reinforcement Learning

Huang, Yiran; Zhou, Yexu; Hefenbrock, Michael; Riedel, Till; Fang, Likun; Beigl, Michael

doi:10.5445/IR/1000152792

Universal Distributional Decision-based Black-box Adversarial Attack with Reinforcement Learning

Huang, Yiran

¹; Zhou, Yexu ¹; Hefenbrock, Michael ¹; Riedel, Till

¹; Fang, Likun

¹; Beigl, Michael

¹
¹ Institut für Telematik (TM), Karlsruher Institut für Technologie (KIT)

Abstract:

The vulnerability of the high-performance machine learning models implies a security risk in applications with real-world consequences. Research on adversarial attacks is beneficial in guiding the development of machine learning models on the one hand and finding targeted defenses on the other. However, most of the adversarial attacks today leverage the gradient or logit information from the models to generate adversarial perturbation. Works in the more realistic domain: decision-based attacks, which generate adversarial perturbation solely based on observing the output label of the targeted model, are still relatively rare and mostly use gradient-estimation strategies. In this work, we propose a pixel-wise decision-based attack algorithm that finds a distribution of adversarial perturbation through a reinforcement learning algorithm. We call this method Decision-based Black-box Attack with Reinforcement learning (DBAR). Experiments show that the proposed approach outperforms state-of-the-art decision-based attacks with a higher attack success rate and greater transferability.

KITopen-Download

Volltext

DOI: 10.5445/IR/1000152792

Veröffentlicht am 22.11.2022

Export

Statistiken

Seitenaufrufe: 92
seit 22.11.2022

Downloads: 28
seit 23.11.2022

Zugehörige Institution(en) am KIT	Institut für Telematik (TM)
Publikationstyp	Forschungsbericht/Preprint
Publikationsdatum	15.11.2022
Sprache	Englisch
Identifikator	KITopen-ID: 1000152792
Verlag	arxiv
Schlagwörter	Adversarial attack and Decision attack and Reinforcement Learning

Repository KITopen

Universal Distributional Decision-based Black-box Adversarial Attack with Reinforcement Learning

Abstract: