paper-with-me

홈 › Papers

Interpreting Adversarially Trained Convolutional Neural Networks

2019-05-23 · Tianyuan Zhang, Zhanxing Zhu

We attempt to interpret how adversarially trained convolutional neural networks (AT-CNNs) recognize objects. We design systematic approaches to interpret AT-CNNs in both qualitative and quantitative ways and compare them with normally trained models. Surprisingly, we find that adversarial training alleviates the texture bias of standard CNNs when trained on object recognition tasks, and helps CNNs learn a more shape-biased representation. We validate our hypothesis from two aspects. First, we compare the salience maps of AT-CNNs and standard CNNs on clean images and images under different transformations. The comparison could visually show that the prediction of the two types of CNNs is sensitive to dramatically different types of features. Second, to achieve quantitative verification, we construct additional test datasets that destroy either textures or shapes, such as style-transferred version of clean data, saturated images and patch-shuffled ones, and then evaluate the classification accuracy of AT-CNNs and normal CNNs on these datasets. Our findings shed some light on why AT-CNNs are more robust than those normally trained ones and contribute to a better understanding of adversarial training over CNNs from an interpretation perspective.

📄 PDF Abstract BibTeX arXiv:1905.09797

Code (1)

PKUAI26/AT-CNN 공식 구현 pytorch

Tasks

Object Recognition

Similar Papers 제목 키워드 기반

Interpreting Attributions and Interactions of Adversarial Attacks

2021-08-16 · ICCV 2021 10 · Xin Wang, Shuyun Lin, Hao Zhang, Yufei Zhu 외

This paper aims to explain adversarial attacks in terms of how adversarial perturbations contribute to the attacking task. We estimate attributions of different image regions to the decrease of the attacking cost based o…

Large Norms of CNN Layers Do Not Hurt Adversarial Robustness

2020-09-17 · Youwei Liang, Dong Huang

Since the Lipschitz properties of convolutional neural networks (CNNs) are widely considered to be related to adversarial robustness, we theoretically characterize the $\ell_1$ norm and $\ell_\infty$ norm of 2D multi-cha…

Adversarial Robustness

Improving Interpretability in Medical Imaging Diagnosis using Adversarial Training

2020-12-02 · Andrei Margeloiu, Nikola Simidjievski, Mateja Jamnik, Adrian Weller

We investigate the influence of adversarial training on the interpretability of convolutional neural networks (CNNs), specifically applied to diagnosing skin cancer. We show that gradient-based saliency maps of adversari…

Lung Segmentation and Nodule Detection in Computed Tomography Scan using a Convolutional Neural Network Trained Adversarially using Turing Test Loss

2020-06-16 · Rakshith Sathish, Rachana Sathish, Ramanathan Sethuraman, Debdoot Sheet

Lung cancer is the most common form of cancer found worldwide with a high mortality rate. Early detection of pulmonary nodules by screening with a low-dose computed tomography (CT) scan is crucial for its effective clini…

Computed Tomography (CT)Management

GAT: Generative Adversarial Training for Adversarial Example Detection and Robust Classification

2019-05-27 · Xuwang Yin, Soheil Kolouri, Gustavo K. Rohde

The vulnerabilities of deep neural networks against adversarial examples have become a significant concern for deploying these models in sensitive domains. Devising a definitive defense against such attacks is proven to …

ClassificationGeneral ClassificationRobust classificationvalid