paper-with-me

홈 › Papers

Defending Pre-trained Language Models from Adversarial Word Substitution Without Performance Sacrifice

2021-08-01 · Findings (ACL) 2021 8 · Rongzhou Bao, Jiayi Wang, Hai Zhao
📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Self-Supervised Contrastive Learning with Adversarial Perturbations for Defending Word Substitution-based Attacks

2021-07-15 · Findings (NAACL) 2022 7 · Zhao Meng, Yihan Dong, Mrinmaya Sachan, Roger Wattenhofer

In this paper, we present an approach to improve the robustness of BERT language models against word substitution-based adversarial attacks by leveraging adversarial perturbations for self-supervised contrastive learning…

Adversarial AttackContrastive LearningLanguage Modelling

Defending Pre-trained Language Models from Adversarial Word Substitutions Without Performance Sacrifice

2021-05-30 · Rongzhou Bao, Jiayi Wang, Hai Zhao

Pre-trained contextualized language models (PrLMs) have led to strong performance gains in downstream natural language understanding tasks. However, PrLMs can still be easily fooled by adversarial word substitution, whic…

Adversarial AttackAnomaly DetectionMulti-Task LearningNatural Language Understanding

Text Adversarial Purification as Defense against Adversarial Attacks

2022-03-27 · Linyang Li, Demin Song, Xipeng Qiu

Adversarial purification is a successful defense mechanism against adversarial attacks without requiring knowledge of the form of the incoming attack. Generally, adversarial purification aims to remove the adversarial pe…

Adversarial AttackAdversarial DefenseAdversarial Purification

Generalization to Mitigate Synonym Substitution Attacks

2020-11-01 · EMNLP (DeeLIO) 2020 11 · Basemah Alshemali, Jugal Kalita

Studies have shown that deep neural networks (DNNs) are vulnerable to adversarial examples – perturbed inputs that cause DNN-based models to produce incorrect results. One robust adversarial attack in the NLP domain is t…

Adversarial Attack

Certified Robustness to Adversarial Word Substitutions

2019-09-03 · IJCNLP 2019 11 · Robin Jia, aditi raghunathan, Kerem Göksel, Percy Liang

State-of-the-art NLP models can often be fooled by adversaries that apply seemingly innocuous label-preserving transformations (e.g., paraphrasing) to input text. The number of possible transformations scales exponential…

Data AugmentationNatural Language InferenceSentiment Analysis