paper-with-me

홈 › Papers

BUSTED at AraGenEval Shared Task: A Comparative Study of Transformer-Based Models for Arabic AI-Generated Text Detection

2025-10-23 · Ali Zain, Sareem Farooqui, Muhammad Rafi arxiv

This paper details our submission to the AraGenEval Shared Task on Arabic AI-generated text detection, where our team, BUSTED, secured 5th place. We investigated the effectiveness of three pre-trained transformer models: AraELECTRA, CAMeLBERT, and XLM-RoBERTa. Our approach involved fine-tuning each model on the provided dataset for a binary classification task. Our findings revealed a surprising result: the multilingual XLM-RoBERTa model achieved the highest performance with an F1 score of 0.7701, outperforming the specialized Arabic models. This work underscores the complexities of AI-generated text detection and highlights the strong generalization capabilities of multilingual models.

📄 PDF Abstract BibTeX arXiv:2510.20610

Code (0)

등록된 구현이 없습니다.

Tasks

Binary ClassificationText Detection

Similar Papers 제목 키워드 기반

RobustEdge: Low Power Adversarial Detection for Cloud-Edge Systems

2023-09-05 · Abhishek Moitra, Abhiroop Bhattacharjee, Youngeun Kim, Priyadarshini Panda

In practical cloud-edge scenarios, where a resource constrained edge performs data acquisition and a cloud system (having sufficient resources) performs inference tasks with a deep neural network (DNN), adversarial robus…

Adversarial RobustnessQuantization

Overview of the VLSP 2023 -- ComOM Shared Task: A Data Challenge for Comparative Opinion Mining from Vietnamese Product Reviews

2024-02-21 · Hoang-Quynh Le, Duy-Cat Can, Khanh-Vinh Nguyen, Mai-Vu Tran

This paper presents a comprehensive overview of the Comparative Opinion Mining from Vietnamese Product Reviews shared task (ComOM), held as part of the 10$^{th}$ International Workshop on Vietnamese Language and Speech P…

Opinion MiningSentence

Are Princelings Truly Busted? Evaluating Transaction Discounts in China's Land Market

2025-02-11 · Julia Manso

This paper narrowly replicates Chen and Kung's 2019 paper ($The$ $Quarterly$ $Journal$ $of$ $Economics$ 134(1): 185-226). Inspecting the data reveals that nearly one-third of the transactions (388,903 out of 1,208,621) a…

ComOM at VLSP 2023: A Dual-Stage Framework with BERTology and Unified Multi-Task Instruction Tuning Model for Vietnamese Comparative Opinion Mining

2023-12-14 · Dang Van Thin, Duong Ngoc Hao, Ngan Luu-Thuy Nguyen

The ComOM shared task aims to extract comparative opinions from product reviews in Vietnamese language. There are two sub-tasks, including (1) Comparative Sentence Identification (CSI) and (2) Comparative Element Extract…

Data AugmentationOpinion MiningSentence

A Comparative Study of Synthetic Data Generation Methods for Grammatical Error Correction

2020-07-01 · WS 2020 7 · Max White, Alla Rozovskaya

Grammatical Error Correction (GEC) is concerned with correcting grammatical errors in written text. Current GEC systems, namely those leveraging statistical and neural machine translation, require large quantities of ann…

Grammatical Error CorrectionMachine TranslationSynthetic Data GenerationTranslation