paper-with-me

홈 › Papers

DACTYL: Diverse Adversarial Corpus of Texts Yielded from Large Language Models

2025-08-01 · Shantanu Thorat, Andrew Caines arxiv

Existing AIG (AI-generated) text detectors struggle in real-world settings despite succeeding in internal testing, suggesting that they may not be robust enough. We rigorously examine the machine-learning procedure to build these detectors to address this. Most current AIG text detection datasets focus on zero-shot generations, but little work has been done on few-shot or one-shot generations, where LLMs are given human texts as an example. In response, we introduce the Diverse Adversarial Corpus of Texts Yielded from Language models (DACTYL), a challenging AIG text detection dataset focusing on one-shot/few-shot generations. We also include texts from domain-specific continued-pre-trained (CPT) language models, where we fully train all parameters using a memory-efficient optimization approach. Many existing AIG text detectors struggle significantly on our dataset, indicating a potential vulnerability to one-shot/few-shot and CPT-generated texts. We also train our own classifiers using two approaches: standard binary cross-entropy (BCE) optimization and a more recent approach, deep X-risk optimization (DXO). While BCE-trained classifiers marginally outperform DXO classifiers on the DACTYL test set, the latter excels on out-of-distribution (OOD) texts. In our mock deployment scenario in student essay detection with an OOD student essay dataset, the best DXO classifier outscored the best BCE-trained classifier by 50.56 macro-F1 score points at the lowest false positive rates for both. Our results indicate that DXO classifiers generalize better without overfitting to the test set. Our experiments highlight several areas of improvement for AIG text detectors.

📄 PDF Abstract BibTeX arXiv:2508.00619

Code (0)

등록된 구현이 없습니다.

Tasks

Text Detection

Similar Papers 제목 키워드 기반

Language Level Classification on German Texts using a Neural Approach

2021-11-16 · ACL ARR November 2021 11 · Anonymous

Studies on language level classification (LLC) for German are scarce. Of the few existing, most use a feature-engineered approach. To the best of our knowledge, there is no deep learning approach on German texts yet. Thi…

ArticlesClassificationSentence

Bukva: Russian Sign Language Alphabet

2024-10-11 · Karina Kvanchiani, Petr Surovtsev, Alexander Nagaev, Elizaveta Petrova 외

This paper investigates the recognition of the Russian fingerspelling alphabet, also known as the Russian Sign Language (RSL) dactyl. Dactyl is a component of sign languages where distinct hand movements represent indivi…

CPUSign Language Recognition

Hengam: An Adversarially Trained Transformer for Persian Temporal Tagging

2022-11-20 · Proceedings of the 2nd Conference of the Asia-Pacific Chapter of the Association for Computational Linguistics and the 12th International Joint Conference on Natural Language Processing 2022 11 · Sajad Mirzababaei, Amir Hossein Kargaran, Hinrich Schütze, Ehsaneddin Asgari

Many NLP main tasks benefit from an accurate understanding of temporal expressions, e.g., text summarization, question answering, and information retrieval. This paper introduces Hengam, an adversarially trained transfor…

Information RetrievalNamed Entity Recognition (NER)Question AnsweringRetrieval+3

Corpus Wide Argument Mining -- a Working Solution

2019-11-25 · Liat Ein-Dor, Eyal Shnarch, Lena Dankin, Alon Halfon 외

One of the main tasks in argument mining is the retrieval of argumentative content pertaining to a given topic. Most previous work addressed this task by retrieving a relatively small number of relevant documents as the …

Argument MiningArticlesRetrievalSentence

A Data-Validated Host-Parasite Model for Infectious Disease Outbreaks

2019-08-29

The use of model experimental systems and mathematical models is important to further understanding of infectious disease dynamics and strategize disease mitigation. Gyrodactylids are helminth ectoparasites of teleost fi…