paper-with-me

Papers

Beyond Explainability: Leveraging Interpretability for Improved Adversarial Learning

2019-04-21 · Devinder Kumar, Ibrahim Ben-Daya, Kanav Vats, Jeffery Feng, Graham Taylor and, Alexander Wong

In this study, we propose the leveraging of interpretability for tasks beyond purely the purpose of explainability. In particular, this study puts forward a novel strategy for leveraging gradient-based interpretability in the realm of adversarial examples, where we use insights gained to aid adversarial learning. More specifically, we introduce the concept of spatially constrained one-pixel adversarial perturbations, where we guide the learning of such adversarial perturbations towards more susceptible areas identified via gradient-based interpretability. Experimental results using different benchmark datasets show that such a spatially constrained one-pixel adversarial perturbation strategy can noticeably improve the speed of convergence as well as produce successful attacks that were also visually difficult to perceive, thus illustrating an effective use of interpretability methods for tasks outside of the purpose of purely explainability.

📄 PDF Abstract BibTeX arXiv:1904.09633

Code (0)

등록된 구현이 없습니다.

Methods 이 논문이 사용한 방법론

SPEED The monocular depth estimation (MDE) is the task of estimating depth from a single frame. This information is an essential knowledge in many computer vision tasks such as scene…
Interpretability 설명 없음

Similar Papers 제목 키워드 기반

Counterfactual Image Generation for adversarially robust and interpretable Classifiers

2023-10-01 · Rafael Bischof, Florian Scheidegger, Michael A. Kraus, A. Cristiano I. Malossi

Neural Image Classifiers are effective but inherently hard to interpret and susceptible to adversarial attacks. Solutions to both problems exist, among others, in the form of counterfactual examples generation to enhance…

counterfactualDescriptiveImage GenerationImage-to-Image Translation+1

Exploiting Explainability to Design Adversarial Attacks and Evaluate Attack Resilience in Hate-Speech Detection Models

2023-05-29 · Pranath Reddy Kumbam, Sohaib Uddin Syed, Prashanth Thamminedi, Suhas Harish 외

The advent of social media has given rise to numerous ethical challenges, with hate speech among the most significant concerns. Researchers are attempting to tackle this problem by leveraging hate-speech detection and em…

Adversarial RobustnessDecision MakingHate Speech Detection

On the Relationship Between Interpretability and Explainability in Machine Learning

2023-11-20 · Benjamin Leblanc, Pascal Germain

Interpretability and explainability have gained more and more attention in the field of machine learning as they are crucial when it comes to high-stakes decisions and troubleshooting. Since both provide information abou…

Position

A Novel XAI-Enhanced Quantum Adversarial Networks for Velocity Dispersion Modeling in MaNGA Galaxies

2025-10-28 · Sathwik Narkedimilli, N V Saran Kumar, Aswath Babu H, Manjunath K Vanahalli 외 arxiv

Current quantum machine learning approaches often face challenges balancing predictive accuracy, robustness, and interpretability. To address this, we propose a novel quantum adversarial framework that integrates a hybri…

Quantum Machine Learning

Kolmogorov-Arnold Graph Neural Networks

2024-06-26 · Gianluca De Carlo, Andrea Mastropietro, Aris Anagnostopoulos

Graph neural networks (GNNs) excel in learning from network-like data but often lack interpretability, making their application challenging in domains requiring transparent decision-making. We propose the Graph Kolmogoro…

Decision MakingGraph ClassificationLink PredictionNode Classification