paper-with-me

홈 › Papers

Gender Bias and Universal Substitution Adversarial Attacks on Grammatical Error Correction Systems for Automated Assessment

2022-08-19 · Vyas Raina, Mark Gales

Grammatical Error Correction (GEC) systems perform a sequence-to-sequence task, where an input word sequence containing grammatical errors, is corrected for these errors by the GEC system to output a grammatically correct word sequence. With the advent of deep learning methods, automated GEC systems have become increasingly popular. For example, GEC systems are often used on speech transcriptions of English learners as a form of assessment and feedback - these powerful GEC systems can be used to automatically measure an aspect of a candidate's fluency. The count of \textit{edits} from a candidate's input sentence (or essay) to a GEC system's grammatically corrected output sentence is indicative of a candidate's language ability, where fewer edits suggest better fluency. The count of edits can thus be viewed as a \textit{fluency score} with zero implying perfect fluency. However, although deep learning based GEC systems are extremely powerful and accurate, they are susceptible to adversarial attacks: an adversary can introduce a small, specific change at the input of a system that causes a large, undesired change at the output. When considering the application of GEC systems to automated language assessment, the aim of an adversary could be to cheat by making a small change to a grammatically incorrect input sentence that conceals the errors from a GEC system, such that no edits are found and the candidate is unjustly awarded a perfect fluency score. This work examines a simple universal substitution adversarial attack that non-native speakers of English could realistically employ to deceive GEC systems used for assessment.

📄 PDF Abstract BibTeX arXiv:2208.09466

Code (0)

등록된 구현이 없습니다.

Tasks

Adversarial AttackGrammatical Error CorrectionSentence

Similar Papers 제목 키워드 기반

Language Models That Walk the Talk: A Framework for Formal Fairness Certificates

2025-05-19 · Danqing Chen, Tobias Ladner, Ahmed Rayen Mhadhbi, Matthias Althoff

As large language models become integral to high-stakes applications, ensuring their robustness and fairness is critical. Despite their success, large language models remain vulnerable to adversarial attacks, where small…

Fairness

Universal Acoustic Adversarial Attacks for Flexible Control of Speech-LLMs

2025-05-20 · Rao Ma, Mengjie Qian, Vyas Raina, Mark Gales 외

The combination of pre-trained speech encoders with large language models has enabled the development of speech LLMs that can handle a wide range of spoken language processing tasks. While these models are powerful and f…

Attribute

Are Synonym Substitution Attacks Really Synonym Substitution Attacks?

2022-10-06 · Cheng-Han Chiang, Hung-Yi Lee

In this paper, we explore the following question: Are synonym substitution attacks really synonym substitution attacks (SSAs)? We approach this question by examining how SSAs replace words in the original sentence and sh…

Sentence

It's All in the Name: Mitigating Gender Bias with Name-Based Counterfactual Data Substitution

2019-09-02 · IJCNLP 2019 11 · Rowan Hall Maudslay, Hila Gonen, Ryan Cotterell, Simone Teufel

This paper treats gender bias latent in word embeddings. Previous mitigation attempts rely on the operationalisation of gender bias as a projection over a linear subspace. An alternative approach is Counterfactual Data A…

AllcounterfactualData AugmentationWord Embeddings

Self-Supervised Contrastive Learning with Adversarial Perturbations for Defending Word Substitution-based Attacks

2021-07-15 · Findings (NAACL) 2022 7 · Zhao Meng, Yihan Dong, Mrinmaya Sachan, Roger Wattenhofer

In this paper, we present an approach to improve the robustness of BERT language models against word substitution-based adversarial attacks by leveraging adversarial perturbations for self-supervised contrastive learning…

Adversarial AttackContrastive LearningLanguage Modelling