paper-with-me

홈 › Papers

Using Random Perturbations to Mitigate Adversarial Attacks on Sentiment Analysis Models

2022-02-11 · ICON 2021 12 · Abigail Swenor, Jugal Kalita

Attacks on deep learning models are often difficult to identify and therefore are difficult to protect against. This problem is exacerbated by the use of public datasets that typically are not manually inspected before use. In this paper, we offer a solution to this vulnerability by using, during testing, random perturbations such as spelling correction if necessary, substitution by random synonym, or simply dropping the word. These perturbations are applied to random words in random sentences to defend NLP models against adversarial attacks. Our Random Perturbations Defense and Increased Randomness Defense methods are successful in returning attacked models to similar accuracy of models before attacks. The original accuracy of the model used in this work is 80% for sentiment classification. After undergoing attacks, the accuracy drops to accuracy between 0% and 44%. After applying our defense methods, the accuracy of the model is returned to the original accuracy within statistical significance.

📄 PDF Abstract BibTeX arXiv:2202.05758

Code (0)

등록된 구현이 없습니다.

Tasks

Sentiment AnalysisSentiment ClassificationSpelling Correction

Similar Papers 제목 키워드 기반

Learning to Discriminate Perturbations for Blocking Adversarial Attacks in Text Classification

2019-09-06 · IJCNLP 2019 11 · Yichao Zhou, Jyun-Yu Jiang, Kai-Wei Chang, Wei Wang

Adversarial attacks against machine learning models have threatened various real-world applications such as spam filtering and sentiment analysis. In this paper, we propose a novel framework, learning to DIScriminate Per…

BlockingGeneral ClassificationSentiment Analysistext-classification+1

Adversarial Evasion Attack Efficiency against Large Language Models

2024-06-12 · João Vitorino, Eva Maia, Isabel Praça

Large Language Models (LLMs) are valuable for text classification, but their vulnerabilities must not be disregarded. They lack robustness against adversarial examples, so it is pertinent to understand the impacts of dif…

Adversarial DefenseClassificationSentiment AnalysisSentiment Classification+2

A Mask-Based Adversarial Defense Scheme

2022-04-21 · Weizhen Xu, Chenyi Zhang, Fangzhen Zhao, Liangda Fang

Adversarial attacks hamper the functionality and accuracy of Deep Neural Networks (DNNs) by meddling with subtle perturbations to their inputs.In this work, we propose a new Mask-based Adversarial Defense scheme (MAD) fo…

Adversarial AttackAdversarial DefenseDenoising

End-to-End Adversarial White Box Attacks on Music Instrument Classification

2020-07-29 · Katharina Prinz, Arthur Flexer

Small adversarial perturbations of input data are able to drastically change performance of machine learning systems, thereby challenging the validity of such systems. We present the very first end-to-end adversarial att…

BIG-bench Machine LearningGeneral Classification

How adversarial attacks can disrupt seemingly stable accurate classifiers

2023-09-07 · Oliver J. Sutton, Qinghua Zhou, Ivan Y. Tyukin, Alexander N. Gorban 외

Adversarial attacks dramatically change the output of an otherwise accurate learning system using a seemingly inconsequential modification to a piece of input data. Paradoxically, empirical evidence indicates that even s…

image-classificationImage Classification