paper-with-me

홈 › Papers

RAID: A Shared Benchmark for Robust Evaluation of Machine-Generated Text Detectors

2024-05-13 · Liam Dugan, Alyssa Hwang, Filip Trhlik, Josh Magnus Ludan, Andrew Zhu, Hainiu Xu, Daphne Ippolito, Chris Callison-Burch

Many commercial and open-source models claim to detect machine-generated text with extremely high accuracy (99% or more). However, very few of these detectors are evaluated on shared benchmark datasets and even when they are, the datasets used for evaluation are insufficiently challenging-lacking variations in sampling strategy, adversarial attacks, and open-source generative models. In this work we present RAID: the largest and most challenging benchmark dataset for machine-generated text detection. RAID includes over 6 million generations spanning 11 models, 8 domains, 11 adversarial attacks and 4 decoding strategies. Using RAID, we evaluate the out-of-domain and adversarial robustness of 8 open- and 4 closed-source detectors and find that current detectors are easily fooled by adversarial attacks, variations in sampling strategies, repetition penalties, and unseen generative models. We release our data along with a leaderboard to encourage future research.

📄 PDF Abstract BibTeX arXiv:2405.07940

Code (1)

liamdugan/raid 공식 구현 pytorch

Tasks

Adversarial RobustnessText Detection

Similar Papers 제목 키워드 기반

GenAI Content Detection Task 3: Cross-Domain Machine-Generated Text Detection Challenge

2025-01-15 · Liam Dugan, Andrew Zhu, Firoj Alam, Preslav Nakov 외

Recently there have been many shared tasks targeting the detection of generated text from Large Language Models (LLMs). However, these shared tasks tend to focus either on cases where text is limited to one particular do…

Text Detection

Generative Responsible AI Data Evaluation Schema (GRAIDES) for AI Assurance in Local Government

2026-06-18 · Ethan Knights, Christopher Conlan, Temilorun Gbolahan, Stephen Waterman 외 arxiv

Trust in the application of generative Artificial Intelligence (AI) relies on well-governed measurable evidence of performance and safety. In practice, however, evaluation data is often fragmented across systems, inconsi…

Raidar: geneRative AI Detection viA Rewriting

2024-01-23 · Chengzhi Mao, Carl Vondrick, Hao Wang, Junfeng Yang

We find that large language models (LLMs) are more likely to modify human-written text than AI-generated text when tasked with rewriting. This tendency arises because LLMs often perceive AI-generated text as high-quality…

RAID: A Dataset for Testing the Adversarial Robustness of AI-Generated Image Detectors

2025-06-04 · Hicham Eddoubi, Jonas Ricker, Federico Cocchi, Angelo Sotgiu 외

AI-generated images have reached a quality level at which humans are incapable of reliably distinguishing them from real images. To counteract the inherent risk of fraud and disinformation, the detection of AI-generated …

Adversarial Robustness

Machine learning discovers invariants of braids and flat braids

2023-07-22 · Alexei Lisitsa, Mateo Salles, Alexei Vernitski

We use machine learning to classify examples of braids (or flat braids) as trivial or non-trivial. Our ML takes form of supervised learning using neural networks (multilayer perceptrons). When they achieve good results i…