paper-with-me

홈 › Papers

Adversarial Filters of Dataset Biases

2020-02-10 · ICML 2020 1 · Ronan Le Bras, Swabha Swayamdipta, Chandra Bhagavatula, Rowan Zellers, Matthew E. Peters, Ashish Sabharwal, Yejin Choi

Large neural models have demonstrated human-level performance on language and vision benchmarks, while their performance degrades considerably on adversarial or out-of-distribution samples. This raises the question of whether these models have learned to solve a dataset rather than the underlying task by overfitting to spurious dataset biases. We investigate one recently proposed approach, AFLite, which adversarially filters such dataset biases, as a means to mitigate the prevalent overestimation of machine performance. We provide a theoretical understanding for AFLite, by situating it in the generalized framework for optimum bias reduction. We present extensive supporting evidence that AFLite is broadly applicable for reduction of measurable dataset biases, and that models trained on the filtered datasets yield better generalization to out-of-distribution tasks. Finally, filtering results in a large drop in model performance (e.g., from 92% to 62% for SNLI), while human performance still remains high. Our work thus shows that such filtered datasets can pose new research challenges for robust generalization by serving as upgraded benchmarks.

📄 PDF Abstract BibTeX arXiv:2002.04108

Code (1)

swabhs/notebooks_for_aflite

Tasks

Natural Language Inference

Similar Papers 제목 키워드 기반

Gabor filter incorporated CNN for compression

2021-10-29 · Akihiro Imamura, Nana Arizumi

Convolutional neural networks (CNNs) are remarkably successful in many computer vision tasks. However, the high cost of inference is problematic for embedded and real-time systems, so there are many studies on compressin…

Adversarial Robustness via Random Projection Filters

2023-01-01 · CVPR 2023 1 · Minjing Dong, Chang Xu

Deep Neural Networks show superior performance in various tasks but are vulnerable to adversarial attacks. Most defense techniques are devoted to the adversarial training strategies, however, it is difficult to achie…

Adversarial RobustnessAttributeLEMMA

DVS-Attacks: Adversarial Attacks on Dynamic Vision Sensors for Spiking Neural Networks

2021-07-01 · Alberto Marchisio, Giacomo Pira, Maurizio Martina, Guido Masera 외

Spiking Neural Networks (SNNs), despite being energy-efficient when implemented on neuromorphic hardware and coupled with event-based Dynamic Vision Sensors (DVS), are vulnerable to security threats, such as adversarial …

Adversarial Attack

A Comprehensive Analysis of Adversarial Attacks against Spam Filters

2025-05-04 · Esra Hotoğlu, Sevil Sen, Burcu Can

Deep learning has revolutionized email filtering, which is critical to protect users from cyber threats such as spam, malware, and phishing. However, the increasing sophistication of adversarial attacks poses a significa…

Deep LearningSentenceSpam detection

A Neurosymbolic Framework for Bias Correction in Convolutional Neural Networks

2024-05-24 · Parth Padalkar, Natalia Ślusarz, Ekaterina Komendantskaya, Gopal Gupta

Recent efforts in interpreting Convolutional Neural Networks (CNNs) focus on translating the activation of CNN filters into a stratified Answer Set Program (ASP) rule-sets. The CNN filters are known to capture high-level…

Decision Makingimage-classificationImage ClassificationSemantic Similarity+1