Patch is enough: naturalistic adversarial patch against vision-language pre-training models

Dehong Kong; Siyuan Liang; Xiaopeng Zhu; Yuansheng Zhong; Wenqi Ren

doi:10.1007/s44267-024-00066-7

Visual Intelligence (Dec 2024)

Patch is enough: naturalistic adversarial patch against vision-language pre-training models

Dehong Kong,
Siyuan Liang,
Xiaopeng Zhu,
Yuansheng Zhong,
Wenqi Ren

Affiliations

Dehong Kong: School of Cyber Science and Technology
Siyuan Liang: School of Computing, National University of Singapore
Xiaopeng Zhu: Guangdong Testing Institute of Product Quality Supervision
Yuansheng Zhong: Guangdong Testing Institute of Product Quality Supervision
Wenqi Ren: School of Cyber Science and Technology

DOI: https://doi.org/10.1007/s44267-024-00066-7
Journal volume & issue: Vol. 2, no. 1
pp. 1 – 10

Abstract

Read online

Abstract Visual language pre-training (VLP) models have demonstrated significant success in various domains, but they remain vulnerable to adversarial attacks. Addressing these adversarial vulnerabilities is crucial for enhancing security in multi-modal learning. Traditionally, adversarial methods that target VLP models involve simultaneous perturbation of images and text. However, this approach faces significant challenges. First, adversarial perturbations often fail to translate effectively into real-world scenarios. Second, direct modifications to the text are conspicuously visible. To overcome these limitations, we propose a novel strategy that uses only image patches for attacks, thus preserving the integrity of the original text. Our method leverages prior knowledge from diffusion models to enhance the authenticity and naturalness of the perturbations. Moreover, to optimize patch placement and improve the effectiveness of our attacks, we utilize the cross-attention mechanism, which encapsulates inter-modal interactions by generating attention maps to guide strategic patch placement. Extensive experiments conducted in a white-box setting for image-to-text scenarios reveal that our proposed method significantly outperforms existing techniques, achieving a 100% attack success rate.

Published in Visual Intelligence

ISSN: 2097-3330 (Print); 2731-9008 (Online)
Publisher: Springer
Country of publisher: Singapore
LCC subjects: Science: Mathematics: Instruments and machines: Electronic computers. Computer science; Science: Physiology: Neurophysiology and neuropsychology
Website: https://link.springer.com/journal/44267

About the journal

Abstract

Keywords