paper-with-me

홈 › Papers

Explaining Generalization of AI-Generated Text Detectors Through Linguistic Analysis

2026-01-12 · Yuxi Xia, Kinga Stańczak, Benjamin Roth arxiv

AI-text detectors achieve high accuracy on in-domain benchmarks, but often struggle to generalize across different generation conditions such as unseen prompts, model families, or domains. While prior work has reported these generalization gaps, there are limited insights about the underlying causes. In this work, we present a systematic study aimed at explaining generalization behavior through linguistic analysis. We construct a comprehensive benchmark that spans 6 prompting strategies, 7 large language models (LLMs), and 4 domain datasets, resulting in a diverse set of human- and AI-generated texts. Using this dataset, we fine-tune classification-based detectors on various generation settings and evaluate their cross-prompt, cross-model, and cross-dataset generalization. To explain the performance variance, we compute correlations between generalization accuracies and feature shifts of 80 linguistic features between training and test conditions. Our analysis reveals that generalization performance for specific detectors and evaluation conditions is significantly associated with linguistic features such as tense usage and pronoun frequency.

📄 PDF Abstract BibTeX arXiv:2601.07974

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

MAGA-Bench: Machine-Augment-Generated Text via Alignment Detection Benchmark

2026-01-08 · Anyang Song, Ying Cheng, Yiqian Xu, Rui Feng arxiv

Machine-Generated Text (MGT) is becoming increasingly difficult to distinguish from Human-Written Text (HWT). This trend has exacerbated malicious activities such as fake news and online fraud. The generalization ability…

EAGLE: A Domain Generalization Framework for AI-generated Text Detection

2024-03-23 · Amrita Bhattacharjee, Raha Moraffah, Joshua Garland, Huan Liu

With the advancement in capabilities of Large Language Models (LLMs), one major step in the responsible and safe use of such LLMs is to be able to detect text generated by these models. While supervised AI-generated text…

Contrastive LearningDomain GeneralizationText Detection

On the Zero-Shot Generalization of Machine-Generated Text Detectors

2023-10-08 · Xiao Pu, Jingyu Zhang, Xiaochuang Han, Yulia Tsvetkov 외

The rampant proliferation of large language models, fluent enough to generate text indistinguishable from human-written language, gives unprecedented importance to the detection of machine-generated text. This work is mo…

Zero-shot Generalization

Threat Scenarios and Best Practices to Detect Neural Fake News

2022-10-01 · COLING 2022 10 · Artidoro Pagnoni, Martin Graciarena, Yulia Tsvetkov

In this work, we discuss different threat scenarios from neural fake news generated by state-of-the-art language models. Through our experiments, we assess the performance of generated text detection systems under these …

Out-of-Distribution GeneralizationText Detection

On the Generalization Ability of Machine-Generated Text Detectors

2024-12-23 · Yule Liu, Zhiyuan Zhong, Yifan Liao, Zhen Sun 외

The rise of large language models (LLMs) has raised concerns about machine-generated text (MGT), including ethical and practical issues like plagiarism and misinformation. Building a robust and highly generalizable MGT d…

BenchmarkingMisinformation