paper-with-me

홈 › Papers

AINL-Eval 2025 Shared Task: Detection of AI-Generated Scientific Abstracts in Russian

2025-08-13 · Tatiana Batura, Elena Bruches, Milana Shvenk, Valentin Malykh arxiv

The rapid advancement of large language models (LLMs) has revolutionized text generation, making it increasingly difficult to distinguish between human- and AI-generated content. This poses a significant challenge to academic integrity, particularly in scientific publishing and multilingual contexts where detection resources are often limited. To address this critical gap, we introduce the AINL-Eval 2025 Shared Task, specifically focused on the detection of AI-generated scientific abstracts in Russian. We present a novel, large-scale dataset comprising 52,305 samples, including human-written abstracts across 12 diverse scientific domains and AI-generated counterparts from five state-of-the-art LLMs (GPT-4-Turbo, Gemma2-27B, Llama3.3-70B, Deepseek-V3, and GigaChat-Lite). A core objective of the task is to challenge participants to develop robust solutions capable of generalizing to both (i) previously unseen scientific domains and (ii) models not included in the training data. The task was organized in two phases, attracting 10 teams and 159 submissions, with top systems demonstrating strong performance in identifying AI-generated content. We also establish a continuous shared task platform to foster ongoing research and long-term progress in this important area. The dataset and platform are publicly available at https://github.com/iis-research-team/AINL-Eval-2025.

📄 PDF Abstract BibTeX arXiv:2508.09622

Code (0)

등록된 구현이 없습니다.

Tasks

Text Generation

Similar Papers 제목 키워드 기반

Overview of SHROOM-Visions 2026: A Shared Task on Hallucination Detection in Large Vision-Language Models

2026-08-26 · Raúl Vázquez, Aman Sinha, Chuyuan Li, Claudio Savelli 외 arxiv

In 2026, we held the fourth iteration of the SHROOM Shared Task series: SHROOM-Visions (\textbf{S}hared-task on \textbf{H}allucinations and \textbf{R}elated \textbf{O}bservable \textbf{O}vergeneration \textbf{M}istakes i…

Image CaptioningText Generation

BUSTED at AraGenEval Shared Task: A Comparative Study of Transformer-Based Models for Arabic AI-Generated Text Detection

2025-10-23 · Ali Zain, Sareem Farooqui, Muhammad Rafi arxiv

This paper details our submission to the AraGenEval Shared Task on Arabic AI-generated text detection, where our team, BUSTED, secured 5th place. We investigated the effectiveness of three pre-trained transformer models:…

Binary ClassificationText Detection

MasonTigers at SemEval-2024 Task 8: Performance Analysis of Transformer-based Models on Machine-Generated Text Detection

2024-03-22 · Sadiya Sayara Chowdhury Puspo, Md Nishat Raihan, Dhiman Goswami, Al Nahian Bin Emran 외

This paper presents the MasonTigers entry to the SemEval-2024 Task 8 - Multigenerator, Multidomain, and Multilingual Black-Box Machine-Generated Text Detection. The task encompasses Binary Human-Written vs. Machine-Gener…

Sentencetext-classificationText ClassificationText Detection

Overview of the DAGPap22 Shared Task on Detecting Automatically Generated Scientific Papers

2022-10-01 · sdp (COLING) 2022 10 · Yury Kashnitsky, Drahomira Herrmannova, Anita de Waard, George Tsatsaronis 외

This paper provides an overview of the DAGPap22 shared task on the detection of automatically generated scientific papers at the Scholarly Document Process workshop colocated with COLING. We frame the detection problem a…

Binary Classification

M-DAIGT: A Shared Task on Multi-Domain Detection of AI-Generated Text

2025-11-14 · Salima Lamsiyah, Saad Ezzini, Abdelkader El Mahdaouy, Hamza Alami 외 arxiv

The generation of highly fluent text by Large Language Models (LLMs) poses a significant challenge to information integrity and academic research. In this paper, we introduce the Multi-Domain Detection of AI-Generated Te…

Binary Classification