paper-with-me

홈 › Papers

GenAI Content Detection Task 3: Cross-Domain Machine-Generated Text Detection Challenge

2025-01-15 · Liam Dugan, Andrew Zhu, Firoj Alam, Preslav Nakov, Marianna Apidianaki, Chris Callison-Burch

Recently there have been many shared tasks targeting the detection of generated text from Large Language Models (LLMs). However, these shared tasks tend to focus either on cases where text is limited to one particular domain or cases where text can be from many domains, some of which may not be seen during test time. In this shared task, using the newly released RAID benchmark, we aim to answer whether or not models can detect generated text from a large, yet fixed, number of domains and LLMs, all of which are seen during training. Over the course of three months, our task was attempted by 9 teams with 23 detector submissions. We find that multiple participants were able to obtain accuracies of over 99% on machine-generated text from RAID while maintaining a 5% False Positive Rate -- suggesting that detectors are able to robustly detect text from many domains and models simultaneously. We discuss potential interpretations of this result and provide directions for future research.

📄 PDF Abstract BibTeX arXiv:2501.08913

Code (1)

liamdugan/raid 공식 구현 pytorch

Tasks

Text Detection

Methods 이 논문이 사용한 방법론

Focus 설명 없음

Similar Papers 제목 키워드 기반

LuxVeri at GenAI Detection Task 3: Cross-Domain Detection of AI-Generated Text Using Inverse Perplexity-Weighted Ensemble of Fine-Tuned Transformer Models

2025-01-21 · Md Kamrujjaman Mobin, Md Saiful Islam

This paper presents our approach for Task 3 of the GenAI content detection workshop at COLING-2025, focusing on Cross-Domain Machine-Generated Text (MGT) Detection. We propose an ensemble of fine-tuned transformer models…

GenAI Content Detection Task 1: English and Multilingual Machine-Generated Text Detection: AI vs. Human

2025-01-19 · Yuxia Wang, Artem Shelmanov, Jonibek Mansurov, Akim Tsvigun 외

We present the GenAI Content Detection Task~1 -- a shared task on binary machine generated text detection, conducted as a part of the GenAI workshop at COLING 2025. The task consists of two subtasks: Monolingual (English…

Text Detection

GenAI vs. Human Fact-Checkers: Accurate Ratings, Flawed Rationales

2025-02-20 · Yuehong Cassandra Tai, Khushi Navin Patni, Nicholas Daniel Hemauer, Bruce Desmarais 외

Despite recent advances in understanding the capabilities and limits of generative artificial intelligence (GenAI) models, we are just beginning to understand their capacity to assess and reason about the veracity of con…

Misinformation

A Survey on Responsible Generative AI: What to Generate and What Not

2024-04-08 · Jindong Gu

In recent years, generative AI (GenAI), like large language models and text-to-image models, has received significant attention across various domains. However, ensuring the responsible generation of content by these mod…

How to Strategize Human Content Creation in the Era of GenAI?

2024-06-07 · Seyed A. Esmaeili, Kevin Lim, Kshipra Bhawalkar, Zhe Feng 외

Generative AI (GenAI) will have significant impact on content creation platforms. In this paper, we study the dynamic competition between a GenAI and a human contributor. Unlike the human, the GenAI's content only improv…