paper-with-me

Papers

FeatureLens: A Highly Generalizable and Interpretable Framework for Detecting Adversarial Examples Based on Image Features

2025-12-03 · Zhigang Yang, Yuan Liu, Jiawei Zhang, Puning Zhang, Xinqiang Ma arxiv

Although the remarkable performance of deep neural networks (DNNs) in image classification, their vulnerability to adversarial attacks remains a critical challenge. Most existing detection methods rely on complex and poorly interpretable architectures, which compromise interpretability and generalization. To address this, we propose FeatureLens, a lightweight framework that acts as a lens to scrutinize anomalies in image features. Comprising an Image Feature Extractor (IFE) and shallow classifiers (e.g., SVM, MLP, or XGBoost) with model sizes ranging from 1,000 to 30,000 parameters, FeatureLens achieves high detection accuracy ranging from 97.8% to 99.75% in closed-set evaluation and 86.17% to 99.6% in generalization evaluation across FGSM, PGD, CW, and DAmageNet attacks, using only 51 dimensional features. By combining strong detection performance with excellent generalization, interpretability, and computational efficiency, FeatureLens offers a practical pathway toward transparent and effective adversarial defense.

📄 PDF Abstract BibTeX arXiv:2512.03625

Code (0)

등록된 구현이 없습니다.

Tasks

Computational EfficiencyImage ClassificationAdversarial Defense

Similar Papers 제목 키워드 기반

MOMENTA: A Multimodal Framework for Detecting Harmful Memes and Their Targets

2021-09-11 · Findings (EMNLP) 2021 11 · Shraman Pramanick, Shivam Sharma, Dimitar Dimitrov, Md Shad Akhtar 외

Internet memes have become powerful means to transmit political, psychological, and socio-cultural ideas. Although memes are typically humorous, recent days have witnessed an escalation of harmful memes used for trolling…

Vulnerability-Aware Spatio-Temporal Learning for Generalizable and Interpretable Deepfake Video Detection

2025-01-02 · Dat Nguyen, Marcella Astrid, Anis Kacem, Enjie Ghorbel 외

Detecting deepfake videos is highly challenging due to the complex intertwined spatial and temporal artifacts in forged sequences. Most recent approaches rely on binary classifiers trained on both real and fake data. How…

Face SwappingMulti-Task Learning

Detecting Pipeline Failures through Fine-Grained Analysis of Web Agents

2025-09-17 · Daniel Röder, Akhil Juneja, Roland Roller, Sven Schmeier arxiv

Web agents powered by large language models (LLMs) can autonomously perform complex, multistep tasks in dynamic web environments. However, current evaluations mostly focus on the overall success while overlooking interme…

A generalizable large-scale foundation model for musculoskeletal radiographs

2026-02-03 · Shinn Kim, Soobin Lee, Kyoungseob Shin, Han-Soo Kim 외 arxiv

Artificial intelligence (AI) has shown promise in detecting and characterizing musculoskeletal diseases from radiographs. However, most existing models remain task-specific, annotation-dependent, and limited in generaliz…

Self-Supervised LearningFracture detection

ProtoEEGNet: An Interpretable Approach for Detecting Interictal Epileptiform Discharges

2023-12-03 · Dennis Tang, Frank Willard, Ronan Tegerdine, Luke Triplett 외

In electroencephalogram (EEG) recordings, the presence of interictal epileptiform discharges (IEDs) serves as a critical biomarker for seizures or seizure-like events.Detecting IEDs can be difficult; even highly trained …

Decision MakingEEGElectroencephalogram (EEG)