paper-with-me

Papers

Who Wrote This Line? Evaluating the Detection of LLM-Generated Classical Chinese Poetry

2026-04-11 · Jiang Li, Tian Lan, Shanshan Wang, Dongxing Zhang, Dianqing Lin, Guanglai Gao, Derek F. Wong, Xiangdong Su arxiv

The rapid development of large language models (LLMs) has extended text generation tasks into the literary domain. However, AI-generated literary creations has raised increasingly prominent issues of creative authenticity and ethics in literary world, making the detection of LLM-generated literary texts essential and urgent. While previous works have made significant progress in detecting AI-generated text, it has yet to address classical Chinese poetry. Due to the unique linguistic features of classical Chinese poetry, such as strict metrical regularity, a shared system of poetic imagery, and flexible syntax, distinguishing whether a poem is authored by AI presents a substantial challenge. To address these issues, we introduce ChangAn, a benchmark for detecting LLM-generated classical Chinese poetry that containing total 30,664 poems, 10,276 are human-written poems and 20,388 poems are generated by four popular LLMs. Based on ChangAn, we conducted a systematic evaluation of 12 AI detectors, investigating their performance variations across different text granularities and generation strategies. Our findings highlight the limitations of current Chinese text detectors, which fail to serve as reliable tools for detecting LLM-generated classical Chinese poetry. These results validate the effectiveness and necessity of our proposed ChangAn benchmark. Our dataset and code are available at https://github.com/VelikayaScarlet/ChangAn.

📄 PDF Abstract BibTeX arXiv:2604.10101

Code (0)

등록된 구현이 없습니다.

Tasks

Text Generation

Similar Papers 제목 키워드 기반

Who Wrote this Code? Watermarking for Code Generation

2023-05-24 · Taehyun Lee, Seokhee Hong, Jaewoo Ahn, Ilgee Hong 외

Since the remarkable generation performance of large language models raised ethical and legal concerns, approaches to detect machine-generated text by embedding watermarks are being developed. However, we discover that t…

Code GenerationText Detection

Who Wrote This? The Key to Zero-Shot LLM-Generated Text Detection Is GECScore

2024-05-07 · Junchao Wu, Runzhe Zhan, Derek F. Wong, Shu Yang 외

The efficacy of an large language model (LLM) generated text detector depends substantially on the availability of sizable training data. White-box zero-shot detectors, which require no such data, are nonetheless limited…

Language ModelingLanguage ModellingLarge Language ModelLLM-generated Text Detection+1

Human Centered NLP with User-Factor Adaptation

2017-09-01 · EMNLP 2017 9 · Veronica Lynn, Youngseo Son, Vivek Kulkarni, Niranjan Balasubramanian 외

We pose the general task of user-factor adaptation {--} adapting supervised learning models to real-valued user factors inferred from a background of their language, reflecting the idea that a piece of text should be und…

Document ClassificationDomain AdaptationPOSPOS Tagging+3

Who Wrote This? Identifying Machine vs Human-Generated Text in Hausa

2025-03-17 · Babangida Sani, Aakansha Soy, Sukairaj Hafiz Imam, Ahmad Mustapha 외

The advancement of large language models (LLMs) has allowed them to be proficient in various tasks, including content generation. However, their unregulated usage can lead to malicious activities such as plagiarism and g…

ArticlesText Detection

L3DAS22 Challenge: Learning 3D Audio Sources in a Real Office Environment

2022-02-21 · Eric Guizzo, Christian Marinoni, Marco Pennese, Xinlei Ren 외

The L3DAS22 Challenge is aimed at encouraging the development of machine learning strategies for 3D speech enhancement and 3D sound localization and detection in office-like environments. This challenge improves and exte…

Sound Event Localization and DetectionSpeech Enhancement