paper-with-me

Papers

CaptionFool: Universal Image Captioning Model Attacks

2026-02-28 · Swapnil Parekh arxiv

Image captioning models are encoder-decoder architectures trained on large-scale image-text datasets, making them susceptible to adversarial attacks. We present CaptionFool, a novel universal (input-agnostic) adversarial attack against state-of-the-art transformer-based captioning models. By modifying only 7 out of 577 image patches (approximately 1.2% of the image), our attack achieves 94-96% success rate in generating arbitrary target captions, including offensive content. We further demonstrate that CaptionFool can generate "slang" terms specifically designed to evade existing content moderation filters. Our findings expose critical vulnerabilities in deployed vision-language models and underscore the urgent need for robust defenses against such attacks. Warning: This paper contains model outputs which are offensive in nature.

📄 PDF Abstract BibTeX arXiv:2603.00529

Code (0)

등록된 구현이 없습니다.

Tasks

Adversarial AttackImage Captioning

Similar Papers 제목 키워드 기반

Unveiling the Fragility of Vision-Language Models: Multi-Modal Adversarial Synergy via Texture-Constrained Perturbations and Cross-Modal Optimization

2026-05-26 · Xiang Fang, Wanlong Fang, Changshuo Wang arxiv

Large Vision-Language Models (LVLMs) have transformed multi-modal understanding, excelling in tasks like image captioning and visual question answering by integrating visual and textual inputs. However, their robustness …

Visual Question AnsweringAutonomous DrivingImage Captioning

Controlled Caption Generation for Images Through Adversarial Attacks

2021-07-07 · Nayyer Aafaq, Naveed Akhtar, Wei Liu, Mubarak Shah 외

Deep learning is found to be vulnerable to adversarial examples. However, its adversarial susceptibility in image caption generation is under-explored. We study adversarial examples for vision and language models, which …

Caption GenerationImage CaptioningLanguage Modelling

Protect, Show, Attend and Tell: Empowering Image Captioning Models with Ownership Protection

2020-08-25 · Jian Han Lim, Chee Seng Chan, Kam Woh Ng, Lixin Fan 외

By and large, existing Intellectual Property (IP) protection on deep neural networks typically i) focus on image classification task only, and ii) follow a standard digital watermarking framework that was conventionally …

Image Captioningimage-classificationImage Classification

Generalizing Universal Adversarial Attacks Beyond Additive Perturbations

2020-10-15 · Yanghao Zhang, Wenjie Ruan, Fu Wang, Xiaowei Huang

The previous study has shown that universal adversarial attacks can fool deep neural networks over a large set of input images with a single human-invisible perturbation. However, current methods for universal adversaria…

Adversarial Attack

Exact Adversarial Attack to Image Captioning via Structured Output Learning with Latent Variables

2019-05-10 · CVPR 2019 6 · Yan Xu, Baoyuan Wu, Fumin Shen, Yanbo Fan 외

In this work, we study the robustness of a CNN+RNN based image captioning system being subjected to adversarial noises. We propose to fool an image captioning system to generate some targeted partial captions for an imag…

Adversarial AttackImage Captioning