paper-with-me

홈 › Papers

Distinguishability of Adversarial Examples

2019-05-01 · ICLR 2019 5 · Yi Qin, Ryan Hunt, Chuan Yue

Machine learning models including traditional models and neural networks can be easily fooled by adversarial examples which are generated from the natural examples with small perturbations. This poses a critical challenge to machine learning security, and impedes the wide application of machine learning in many important domains such as computer vision and malware detection. Unfortunately, even state-of-the-art defense approaches such as adversarial training and defensive distillation still suffer from major limitations and can be circumvented. From a unique angle, we propose to investigate two important research questions in this paper: Are adversarial examples distinguishable from natural examples? Are adversarial examples generated by different methods distinguishable from each other? These two questions concern the distinguishability of adversarial examples. Answering them will potentially lead to a simple yet effective approach, termed as defensive distinction in this paper under the formulation of multi-label classification, for protecting against adversarial examples. We design and perform experiments using the MNIST dataset to investigate these two questions, and obtain highly positive results demonstrating the strong distinguishability of adversarial examples. We recommend that this unique defensive distinction approach should be seriously considered to complement other defense approaches.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

BIG-bench Machine LearningMalware DetectionMulti-Label ClassificationMUlTI-LABEL-ClASSIFICATION

Similar Papers 제목 키워드 기반

Making Adversarial Examples More Transferable and Indistinguishable

2020-07-08 · Junhua Zou, Yexin Duan, Boyu Li, Wu Zhang 외

Fast gradient sign attack series are popular methods that are used to generate adversarial examples. However, most of the approaches based on fast gradient sign attack series cannot balance the indistinguishability and t…

Wasserstein Adversarial Examples on Univariant Time Series Data

2023-03-22 · Wenjie Wang, Li Xiong, Jian Lou

Adversarial examples are crafted by adding indistinguishable perturbations to normal examples in order to fool a well-trained deep learning model to misclassify. In the context of computer vision, this notion of indistin…

Adversarial AttackTime Series

Why GANs are overkill for NLP

2022-05-19 · David Alvarez-Melis, Vikas Garg, Adam Tauman Kalai

This work offers a novel theoretical perspective on why, despite numerous attempts, adversarial approaches to generative modeling (e.g., GANs) have not been as popular for certain generation tasks, particularly sequentia…

Text Generation

Improved Natural Language Generation via Loss Truncation

2020-04-30 · ACL 2020 6 · Daniel Kang, Tatsunori Hashimoto

Neural language models are usually trained to match the distributional properties of a large-scale corpus by minimizing the log loss. While straightforward to optimize, this approach forces the model to reproduce all var…

Text Generation

Analysis of classifiers' robustness to adversarial perturbations

2015-02-09 · Alhussein Fawzi, Omar Fawzi, Pascal Frossard

The goal of this paper is to analyze an intriguing phenomenon recently discovered in deep networks, namely their instability to adversarial perturbations (Szegedy et. al., 2014). We provide a theoretical framework for an…

General Classification