paper-with-me

홈 › Papers

Learning to Separate Clusters of Adversarial Representations for Robust Adversarial Detection

2020-12-07 · Byunggill Joe, Jihun Hamm, Sung Ju Hwang, Sooel Son, Insik Shin

Although deep neural networks have shown promising performances on various tasks, they are susceptible to incorrect predictions induced by imperceptibly small perturbations in inputs. A large number of previous works proposed to detect adversarial attacks. Yet, most of them cannot effectively detect them against adaptive whitebox attacks where an adversary has the knowledge of the model and the defense method. In this paper, we propose a new probabilistic adversarial detector motivated by a recently introduced non-robust feature. We consider the non-robust features as a common property of adversarial examples, and we deduce it is possible to find a cluster in representation space corresponding to the property. This idea leads us to probability estimate distribution of adversarial representations in a separate cluster, and leverage the distribution for a likelihood based adversarial detector.

📄 PDF Abstract BibTeX arXiv:2012.03483

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Radial Basis Feature Transformation to Arm CNNs Against Adversarial Attacks

2019-05-01 · ICLR 2019 5 · Saeid Asgari Taghanaki, Shekoofeh Azizi, Ghassan Hamarneh

The linear and non-flexible nature of deep convolutional models makes them vulnerable to carefully crafted adversarial perturbations. To tackle this problem, in this paper, we propose a nonlinear radial basis convolution…

image-classificationImage Classification

Improved Detection of Adversarial Attacks via Penetration Distortion Maximization

2019-09-25 · Shai Rozenberg, Gal Elidan, Ran El-Yaniv

This paper is concerned with the defense of deep models against adversarial at- tacks. We develop an adversarial detection method, which is inspired by the cer- tificate defense approach, and captures the idea of separat…

Hamming Similarity and Graph Laplacians for Class Partitioning and Adversarial Image Detection

2023-05-02 · Huma Jamil, Yajing Liu, Turgay Caglar, Christina M. Cole 외

Researchers typically investigate neural network representations by examining activation outputs for one or more layers of a network. Here, we investigate the potential for ReLU activation patterns (encoded as bit vector…

Adversarial Domain Adaptation for Metal Cutting Sound Detection: Leveraging Abundant Lab Data for Scarce Industry Data

2024-10-23 · Mir Imtiaz Mostafiz, Eunseob Kim, Adrian Shuai Li, Elisa Bertino 외

Cutting state monitoring in the milling process is crucial for improving manufacturing efficiency and tool life. Cutting sound detection using machine learning (ML) models, inspired by experienced machinists, can be empl…

Domain Adaptation

A Unified Perspective on Adversarial Membership Manipulation in Vision Models

2026-04-03 · Ruize Gao, Kaiwen Zhou, Yongqiang Chen, Feng Liu arxiv

Membership inference attacks (MIAs) aim to determine whether a specific data point was part of a model's training set, serving as effective tools for evaluating privacy leakage of vision models. However, existing MIAs im…

Adversarial Robustness