paper-with-me

Papers

LLM-DetectAIve: a Tool for Fine-Grained Machine-Generated Text Detection

2024-08-08 · Mervat Abassy, Kareem Elozeiri, Alexander Aziz, Minh Ngoc Ta, Raj Vardhan Tomar, Bimarsha Adhikari, Saad El Dine Ahmed, Yuxia Wang, Osama Mohammed Afzal, Zhuohan Xie, Jonibek Mansurov, Ekaterina Artemova, Vladislav Mikhailov, Rui Xing, Jiahui Geng, Hasan Iqbal, Zain Muhammad Mujahid, Tarek Mahmoud, Akim Tsvigun, Alham Fikri Aji, Artem Shelmanov, Nizar Habash, Iryna Gurevych, Preslav Nakov

The ease of access to large language models (LLMs) has enabled a widespread of machine-generated texts, and now it is often hard to tell whether a piece of text was human-written or machine-generated. This raises concerns about potential misuse, particularly within educational and academic domains. Thus, it is important to develop practical systems that can automate the process. Here, we present one such system, LLM-DetectAIve, designed for fine-grained detection. Unlike most previous work on machine-generated text detection, which focused on binary classification, LLM-DetectAIve supports four categories: (i) human-written, (ii) machine-generated, (iii) machine-written, then machine-humanized, and (iv) human-written, then machine-polished. Category (iii) aims to detect attempts to obfuscate the fact that a text was machine-generated, while category (iv) looks for cases where the LLM was used to polish a human-written text, which is typically acceptable in academic writing, but not in education. Our experiments show that LLM-DetectAIve can effectively identify the above four categories, which makes it a potentially useful tool in education, academia, and other domains. LLM-DetectAIve is publicly accessible at https://github.com/mbzuai-nlp/LLM-DetectAIve. The video describing our system is available at https://youtu.be/E8eT_bE7k8c.

📄 PDF Abstract BibTeX arXiv:2408.04284

Code (1)

mbzuai-nlp/llm-detectaive 공식 구현

Tasks

Binary ClassificationText Detection

Similar Papers 제목 키워드 기반

Stress-testing Machine Generated Text Detection: Shifting Language Models Writing Style to Fool Detectors

2025-05-30 · Andrea Pedrotti, Michele Papucci, Cristiano Ciaccio, Alessio Miaschi 외

Recent advancements in Generative AI and Large Language Models (LLMs) have enabled the creation of highly realistic synthetic content, raising concerns about the potential for malicious use, such as misinformation and ma…

MisinformationText Detection

LLMDet: A Third Party Large Language Models Generated Text Detection Tool

2023-05-24 · Kangxi Wu, Liang Pang, HuaWei Shen, Xueqi Cheng 외

Generated texts from large language models (LLMs) are remarkably close to high-quality human-authored text, raising concerns about their potential misuse in spreading false information and academic misconduct. Consequent…

Language ModellingLarge Language ModelText Detection

PUCP-Metrix: An Open-source and Comprehensive Toolkit for Linguistic Analysis of Spanish Texts

2025-11-21 · Javier Alonso Villegas Luis, Marco Antonio Sobrevilla Cabezudo arxiv

Linguistic features remain essential for interpretability and tasks that involve style, structure, and readability, but existing Spanish tools offer limited coverage. We present PUCP-Metrix, an open-source and comprehens…

Text Detection

MISMATCH: Fine-grained Evaluation of Machine-generated Text with Mismatch Error Types

2023-06-18 · Keerthiram Murugesan, Sarathkrishna Swaminathan, Soham Dan, Subhajit Chaudhury 외

With the growing interest in large language models, the need for evaluating the quality of machine text compared to reference (typically human-generated) text has become focal attention. Most recent works focus either on…

Sentence

Eval3D: Interpretable and Fine-grained Evaluation for 3D Generation

2025-01-01 · CVPR 2025 1 · Shivam Duggal, Yushi Hu, Oscar Michel, Aniruddha Kembhavi 외

Despite the unprecedented progress in the field of 3D generation, current systems still often fail to produce high-quality 3D assets that are visually appealing and geometrically and semantically consistent across mu…

3D Generation