paper-with-me

Papers

Groot: Adversarial Testing for Generative Text-to-Image Models with Tree-based Semantic Transformation

2024-02-19 · Yi Liu, Guowei Yang, Gelei Deng, Feiyue Chen, Yuqi Chen, Ling Shi, Tianwei Zhang, Yang Liu

With the prevalence of text-to-image generative models, their safety becomes a critical concern. adversarial testing techniques have been developed to probe whether such models can be prompted to produce Not-Safe-For-Work (NSFW) content. However, existing solutions face several challenges, including low success rate and inefficiency. We introduce Groot, the first automated framework leveraging tree-based semantic transformation for adversarial testing of text-to-image models. Groot employs semantic decomposition and sensitive element drowning strategies in conjunction with LLMs to systematically refine adversarial prompts. Our comprehensive evaluation confirms the efficacy of Groot, which not only exceeds the performance of current state-of-the-art approaches but also achieves a remarkable success rate (93.66%) on leading text-to-image models such as DALL-E 3 and Midjourney.

📄 PDF Abstract BibTeX arXiv:2402.12100

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

GROOT: Corrective Reward Optimization for Generative Sequential Labeling

2022-09-29 · Kazuma Hashimoto, Karthik Raman

Sequential labeling is a fundamental NLP task, forming the backbone of many applications. Supervised learning of Seq2Seq models has shown great success on these problems. However, the training objectives are still signif…

Decoder

Efficient Training of Robust Decision Trees Against Adversarial Examples

2020-12-18 · Daniël Vos, Sicco Verwer

In the present day we use machine learning for sensitive tasks that require models to be both understandable and robust. Although traditional models such as decision trees are understandable, they suffer from adversarial…

Adversarial Attack

GROOT: Generating Robust Watermark for Diffusion-Model-Based Audio Synthesis

2024-07-15 · Weizhi Liu, Yue Li, Dongdong Lin, Hui Tian 외

Amid the burgeoning development of generative models like diffusion models, the task of differentiating synthesized audio from its natural counterpart grows more daunting. Deepfake detection offers a viable solution to c…

Audio SynthesisDecoderDeepFake DetectionFace Swapping

GROOT: Learning to Follow Instructions by Watching Gameplay Videos

2023-10-12 · Shaofei Cai, Bowei Zhang, ZiHao Wang, Xiaojian Ma 외

We study the problem of building a controller that can follow open-ended instructions in open-world environments. We propose to follow reference videos as instructions, which offer expressive goal specifications while el…

DecoderInstruction FollowingMinecraft

GrootVL: Tree Topology is All You Need in State Space Model

2024-06-04 · Yicheng Xiao, Lin Song, Shaoli Huang, Jiangshan Wang 외

The state space models, employing recursively propagated features, demonstrate strong representation capabilities comparable to Transformer models and superior efficiency. However, constrained by the inherent geometric c…

Allimage-classificationImage Classificationobject-detection+2