paper-with-me

홈 › Papers

Disrupting Model Training with Adversarial Shortcuts

2021-06-12 · ICML Workshop AML 2021 7 · Ivan Evtimov, Ian Covert, Aditya Kusupati, Tadayoshi Kohno

When data is publicly released for human consumption, it is unclear how to prevent its unauthorized usage for machine learning purposes. Successful model training may be preventable with carefully designed dataset modifications, and we present a proof-of-concept approach for the image classification setting. We propose methods based on the notion of adversarial shortcuts, which encourage models to rely on non-robust signals rather than semantic features, and our experiments demonstrate that these measures successfully prevent deep learning models from achieving high accuracy on real, unmodified data examples.

📄 PDF Abstract BibTeX arXiv:2106.06654

Code (0)

등록된 구현이 없습니다.

Tasks

BIG-bench Machine Learningimage-classificationImage Classificationmodel

Similar Papers 제목 키워드 기반

Hiding Faces in Plain Sight: Disrupting AI Face Synthesis with Adversarial Perturbations

2019-06-21 · Yuezun Li, Xin Yang, Baoyuan Wu, Siwei Lyu

Recent years have seen fast development in synthesizing realistic human faces using AI technologies. Such fake faces can be weaponized to cause negative personal and social impact. In this work, we develop technologies t…

Face DetectionFace Generation

Disrupting Deepfakes: Adversarial Attacks Against Conditional Image Translation Networks and Facial Manipulation Systems

2020-03-03 · Nataniel Ruiz, Sarah Adel Bargal, Stan Sclaroff

Face modification systems using deep learning have become increasingly powerful and accessible. Given images of a person's face, such systems can generate new images of that same person under different expressions and po…

Adversarial AttackAttributeTranslation

Avoiding Reasoning Shortcuts: Adversarial Evaluation, Training, and Model Development for Multi-Hop QA

2019-06-17 · ACL 2019 7 · Yichen Jiang, Mohit Bansal

Multi-hop question answering requires a model to connect multiple pieces of evidence scattered in a long context to answer the question. In this paper, we show that in the multi-hop HotpotQA (Yang et al., 2018) dataset, …

Multi-hop Question AnsweringQuestion AnsweringSentence

Layer-Aware Analysis of Catastrophic Overfitting: Revealing the Pseudo-Robust Shortcut Dependency

2024-05-25 · Runqi Lin, Chaojian Yu, Bo Han, Hang Su 외

Catastrophic overfitting (CO) presents a significant challenge in single-step adversarial training (AT), manifesting as highly distorted deep neural networks (DNNs) that are vulnerable to multi-step adversarial attacks. …

Hiding Faces in Plain Sight: Defending DeepFakes by Disrupting Face Detection

2024-12-02 · Delong Zhu, Yuezun Li, Baoyuan Wu, Jiaran Zhou 외

This paper investigates the feasibility of a proactive DeepFake defense framework, {\em FacePosion}, to prevent individuals from becoming victims of DeepFake videos by sabotaging face detection. The motivation stems from…

Adversarial AttackFace Detection