paper-with-me

홈 › Papers

Adversarial attacks against Fact Extraction and VERification

2019-03-13 · James Thorne, Andreas Vlachos

This paper describes a baseline for the second iteration of the Fact Extraction and VERification shared task (FEVER2.0) which explores the resilience of systems through adversarial evaluation. We present a collection of simple adversarial attacks against systems that participated in the first FEVER shared task. FEVER modeled the assessment of truthfulness of written claims as a joint information retrieval and natural language inference task using evidence from Wikipedia. A large number of participants made use of deep neural networks in their submissions to the shared task. The extent as to whether such models understand language has been the subject of a number of recent investigations and discussion in literature. In this paper, we present a simple method of generating entailment-preserving and entailment-altering perturbations of instances by common patterns within the training data. We find that a number of systems are greatly affected with absolute losses in classification accuracy of up to $29\%$ on the newly perturbed instances. Using these newly generated instances, we construct a sample submission for the FEVER2.0 shared task. Addressing these types of attacks will aid in building more robust fact-checking models, as well as suggest directions to expand the datasets.

📄 PDF Abstract BibTeX arXiv:1903.05543

Code (0)

등록된 구현이 없습니다.

Tasks

Fact CheckingInformation RetrievalNatural Language InferenceRetrieval

Similar Papers 제목 키워드 기반

Evaluating adversarial attacks against multiple fact verification systems

2019-11-01 · IJCNLP 2019 11 · James Thorne, Andreas Vlachos, Christos Christodoulopoulos, Arpit Mittal

Automated fact verification has been progressing owing to advancements in modeling and availability of large datasets. Due to the nature of the task, it is critical to understand the vulnerabilities of these systems agai…

Fact Verification

The FEVER2.0 Shared Task

2019-11-01 · WS 2019 11 · James Thorne, Andreas Vlachos, Oana Cocarascu, Christos Christodoulopoulos 외

We present the results of the second Fact Extraction and VERification (FEVER2.0) Shared Task. The task challenged participants to both build systems to verify factoid claims using evidence retrieved from Wikipedia and to…

Adversarial Attack

DECEIVE-AFC: Adversarial Claim Attacks against Search-Enabled LLM-based Fact-Checking Systems

2026-01-31 · Haoran Ou, Kangjie Chen, Gelei Deng, Hangcheng Liu 외 arxiv

Fact-checking systems with search-enabled large language models (LLMs) have shown strong potential for verifying claims by dynamically retrieving external evidence. However, the robustness of such systems against adversa…

Adversarial Attack

Adversarial Attacks Against Automated Fact-Checking: A Survey

2025-09-10 · Fanzhen Liu, Alsharif Abuadbba, Kristen Moore, Surya Nepal 외 arxiv

In an era where misinformation spreads freely, fact-checking (FC) plays a crucial role in verifying claims and promoting reliable information. While automated fact-checking (AFC) has advanced significantly, existing syst…

The defender's perspective on automatic speaker verification: An overview

2023-05-22 · Haibin Wu, Jiawen Kang, Lingwei Meng, Helen Meng 외

Automatic speaker verification (ASV) plays a critical role in security-sensitive environments. Regrettably, the reliability of ASV has been undermined by the emergence of spoofing attacks, such as replay and synthetic sp…

Speaker Verification