paper-with-me

홈 › Papers

DetectRL-X: Towards Reliable Multilingual and Real-World LLM-Generated Text Detection

2026-05-15 · Junchao Wu, Yefeng Liu, Chenyu Zhu, Hao Zhang, Zeyu Wu, Tianqi Shi, Yichao Du, Longyue Wang, Weihua Luo, Jinsong Su, Derek F. Wong arxiv

The effective detection and governance of Large Language Model (LLM) generated content has become increasingly critical due to the growing risk of misuse. Despite the impressive performance of existing detectors, their reliability and potential in multilingual, real-world scenarios remain largely underexplored. In this study, we introduce DetectRL-X, a comprehensive multilingual benchmark designed to evaluate advanced detectors across 8 dimensions. The benchmark encompasses 8 languages commonly used in commercial contexts and collects human-written texts from 6 domains highly susceptible to LLM misuse. To better aligned with real-world applications, We create LLM-generated texts using 4 popular commercial LLMs, and include typical AI-assisted writing operations such as polishing, expanding, and condensing to capture authentic usage patterns. Furthermore, we develop a multilingual framework for paraphrasing and perturbation attacks to simulate diverse human modifications and writing noise, enabling stress testing of detectors across languages. Experimental results on DetectRL-X reveal the strengths and limitations of current state-of-the-art detectors when applied to diverse linguistic resources. We further analyze how domains, generators, attack strategies, text length, and refinement operations influence performance in different languages, underscoring DetectRL-X as an effective benchmark for strengthening multilingual and language-specific detectors.

📄 PDF Abstract BibTeX arXiv:2605.15518

Code (0)

등록된 구현이 없습니다.

Tasks

Text Detection

Similar Papers 제목 키워드 기반

DetectRL: Benchmarking LLM-Generated Text Detection in Real-World Scenarios

2024-10-31 · Junchao Wu, Runzhe Zhan, Derek F. Wong, Shu Yang 외

Detecting text generated by large language models (LLMs) is of great recent interest. With zero-shot methods like DetectGPT, detection capabilities have reached impressive levels. However, the reliability of existing det…

BenchmarkingLLM-generated Text DetectionText Detection

Zero-Shot Detection of LLM-Generated Text via Implicit Reward Model

2026-04-23 · Runheng Liu, Heyan Huang, Xingchen Xiao, Zhijing Wu arxiv

Large language models (LLMs) have demonstrated remarkable capabilities across various tasks. However, their ability to generate human-like text has raised concerns about potential misuse. This underscores the need for re…

Text Detection

Sure! Here's a short and concise title for your paper: "Contamination in Generated Text Detection Benchmarks"

2025-11-12 · Philipp Dingfelder, Christian Riess arxiv

Large language models are increasingly used for many applications. To prevent illicit use, it is desirable to be able to detect AI-generated text. Training and evaluation of such detectors critically depend on suitable b…

Text Detection

Exons-Detect: Identifying and Amplifying Exonic Tokens via Hidden-State Discrepancy for Robust AI-Generated Text Detection

2026-03-26 · Xiaowei Zhu, Yubing Ren, Fang Fang, Shi Wang 외 arxiv

The rapid advancement of large language models has increasingly blurred the boundary between human-written and AI-generated text, raising societal risks such as misinformation dissemination, authorship ambiguity, and thr…

Text Detection

"Be My Cheese?": Assessing Cultural Nuance in Multilingual LLM Translations

2025-09-25 · Madison Van Doren, Cory Holland arxiv

This pilot study explores the localisation capabilities of state-of-the-art multilingual AI models when translating figurative language, such as idioms and puns, from English into a diverse range of global languages. It …

Machine Translation