paper-with-me

홈 › Papers

On the Robustness of the CVPR 2018 White-Box Adversarial Example Defenses

2018-04-10 · Anish Athalye, Nicholas Carlini

Neural networks are known to be vulnerable to adversarial examples. In this note, we evaluate the two white-box defenses that appeared at CVPR 2018 and find they are ineffective: when applying existing techniques, we can reduce the accuracy of the defended models to 0%.

📄 PDF Abstract BibTeX arXiv:1804.03286

Code (2)

anishathalye/Guided-Denoise 공식 구현 tf
anishathalye/pixel-deflection tf

Similar Papers 제목 키워드 기반

Beware the Black-Box: on the Robustness of Recent Defenses to Adversarial Examples

2020-06-18 · Kaleel Mahmood, Deniz Gurevin, Marten van Dijk, Phuong Ha Nguyen

Many defenses have recently been proposed at venues like NIPS, ICML, ICLR and CVPR. These defenses are mainly focused on mitigating white-box attacks. They do not properly examine black-box attacks. In this paper, we exp…

Diversity

Buffer Zone based Defense against Adversarial Examples in Image Classification

2021-01-01 · Kaleel Mahmood, Phuong Ha Nguyen, Lam M. Nguyen, Thanh V Nguyen 외

Recent defenses published at venues like NIPS, ICML, ICLR and CVPR are mainly focused on mitigating white-box attacks. These defenses do not properly consider adaptive adversaries. In this paper, we expand the scope of t…

Adversarial RobustnessClassificationGeneral Classificationimage-classification+1

Adversarial Attacks on ML Defense Models Competition

2021-10-15 · Yinpeng Dong, Qi-An Fu, Xiao Yang, Wenzhao Xiang 외

Due to the vulnerability of deep neural networks (DNNs) to adversarial examples, a large number of defense techniques have been proposed to alleviate this problem in recent years. However, the progress of building more r…

Adversarial AttackAdversarial Robustnessimage-classificationImage Classification

Towards the Worst-case Robustness of Large Language Models

2025-01-31 · Huanran Chen, Yinpeng Dong, Zeming Wei, Hang Su 외

Recent studies have revealed the vulnerability of large language models to adversarial attacks, where adversaries craft specific input sequences to induce harmful, violent, private, or incorrect outputs. In this work, we…

Language ModelingLanguage Modelling

Improving White-box Robustness of Pre-processing Defenses via Joint Adversarial Training

2021-06-10 · Dawei Zhou, Nannan Wang, Xinbo Gao, Bo Han 외

Deep neural networks (DNNs) are vulnerable to adversarial noise. A range of adversarial defense techniques have been proposed to mitigate the interference of adversarial noise, among which the input pre-processing method…

Adversarial DefenseAdversarial Robustness