paper-with-me

홈 › Papers

Dual Attention Guided Defense Against Malicious Edits

2025-12-16 · Jie Zhang, Shuai Dong, Shiguang Shan, Xilin Chen arxiv

Recent progress in text-to-image diffusion models has transformed image editing via text prompts, yet this also introduces significant ethical challenges from potential misuse in creating deceptive or harmful content. While current defenses seek to mitigate this risk by embedding imperceptible perturbations, their effectiveness is limited against malicious tampering. To address this issue, we propose a Dual Attention-Guided Noise Perturbation (DANP) immunization method that adds imperceptible perturbations to disrupt the model's semantic understanding and generation process. DANP functions over multiple timesteps to manipulate both cross-attention maps and the noise prediction process, using a dynamic threshold to generate masks that identify text-relevant and irrelevant regions. It then reduces attention in relevant areas while increasing it in irrelevant ones, thereby misguides the edit towards incorrect regions and preserves the intended targets. Additionally, our method maximizes the discrepancy between the injected noise and the model's predicted noise to further interfere with the generation. By targeting both attention and noise prediction mechanisms, DANP exhibits impressive immunity against malicious edits, and extensive experiments confirm that our method achieves state-of-the-art performance.

📄 PDF Abstract BibTeX arXiv:2512.14333

Code (0)

등록된 구현이 없습니다.

Tasks

Image Editing

Similar Papers 제목 키워드 기반

Vid-Freeze: Protecting Images from Malicious Image-to-Video Generation via Temporal Freezing

2025-09-27 · Rohit Chowdhury, Aniruddha Bala, Rohan Jaiswal, Siddharth Roheda arxiv

The rapid progress of image-to-video (I2V) generation models has introduced significant risks by enabling deceptive or malicious video synthesis from a single image. Prior defenses such as I2VGuard attempt to immunize im…

Adversarial DefenseMotion SynthesisVideo Generation

Efficient and Robust Video Defense Framework against 3D-field Personalized Talking Face

2025-12-24 · Rui-qing Sun, Xingshan Yao, Tian Lan, Jia-Ling Shi 외 arxiv

State-of-the-art 3D-field video-referenced Talking Face Generation (TFG) methods synthesize high-fidelity personalized talking-face videos in real time by modeling 3D geometry and appearance from reference portrait video…

Computational EfficiencyTalking Face Generation

Anti-Inpainting: A Proactive Defense against Malicious Diffusion-based Inpainters under Unknown Conditions

2025-05-19 · Yimao Guo, Zuomin Qu, Wei Lu, Xiangyang Luo

As diffusion-based malicious image manipulation becomes increasingly prevalent, multiple proactive defense methods are developed to safeguard images against unauthorized tampering. However, most proactive defense methods…

Data AugmentationDenoisingImage Manipulation

AttentionDefense: Leveraging System Prompt Attention for Explainable Defense Against Novel Jailbreaks

2025-04-10 · Charlotte Siska, Anush Sankaran

In the past few years, Language Models (LMs) have shown par-human capabilities in several domains. Despite their practical applications and exceeding user consumption, they are susceptible to jailbreaks when malicious in…

The Attacker Moves Second: Stronger Adaptive Attacks Bypass Defenses Against Llm Jailbreaks and Prompt Injections

2025-10-10 · Milad Nasr, Nicholas Carlini, Chawin Sitawarin, Sander V. Schulhoff 외 arxiv

How should we evaluate the robustness of language model defenses? Current defenses against jailbreaks and prompt injections (which aim to prevent an attacker from eliciting harmful knowledge or remotely triggering malici…

Reinforcement Learning