paper-with-me

홈 › Papers

Decoding AI and Human Authorship: Nuances Revealed Through NLP and Statistical Analysis

2024-07-15 · Mayowa Akinwande, Oluwaseyi Adeliyi, Toyyibat Yussuph

This research explores the nuanced differences in texts produced by AI and those written by humans, aiming to elucidate how language is expressed differently by AI and humans. Through comprehensive statistical data analysis, the study investigates various linguistic traits, patterns of creativity, and potential biases inherent in human-written and AI- generated texts. The significance of this research lies in its contribution to understanding AI's creative capabilities and its impact on literature, communication, and societal frameworks. By examining a meticulously curated dataset comprising 500K essays spanning diverse topics and genres, generated by LLMs, or written by humans, the study uncovers the deeper layers of linguistic expression and provides insights into the cognitive processes underlying both AI and human-driven textual compositions. The analysis revealed that human-authored essays tend to have a higher total word count on average than AI-generated essays but have a shorter average word length compared to AI- generated essays, and while both groups exhibit high levels of fluency, the vocabulary diversity of Human authored content is higher than AI generated content. However, AI- generated essays show a slightly higher level of novelty, suggesting the potential for generating more original content through AI systems. The paper addresses challenges in assessing the language generation capabilities of AI models and emphasizes the importance of datasets that reflect the complexities of human-AI collaborative writing. Through systematic preprocessing and rigorous statistical analysis, this study offers valuable insights into the evolving landscape of AI-generated content and informs future developments in natural language processing (NLP).

📄 PDF Abstract BibTeX arXiv:2408.00769

Code (0)

등록된 구현이 없습니다.

Tasks

Text Generation

Similar Papers 제목 키워드 기반

JAMDEC: Unsupervised Authorship Obfuscation using Constrained Decoding over Small Language Models

2024-02-13 · Jillian Fisher, Ximing Lu, JaeHun Jung, Liwei Jiang 외

The permanence of online content combined with the enhanced authorship identification techniques calls for stronger computational methods to protect the identity and privacy of online authorship when needed, e.g., blind …

The author is dead, but what if they never lived? A reception experiment on Czech AI- and human-authored poetry

2025-11-26 · Anna Marklová, Ondřej Vinš, Martina Vokáčová, Jiří Milička arxiv

Large language models are increasingly capable of producing creative texts, yet most studies on AI-generated poetry focus on English -- a language that dominates training data. In this paper, we examine the perception of…

READER: Robust Evidence-based Authorship Decoding via Extracted Representations

2026-06-09 · Jiaxu Liu, Sunnan Mu, Dong Huang, Liuyin Wang 외 arxiv

As agentic applications increasingly route user tasks through official and third-party LLM APIs, provenance becomes an operational question: which model generated a given black-box response? We study Dynamic Black-Box LL…

Neural Authorship Attribution: Stylometric Analysis on Large Language Models

2023-08-14 · Tharindu Kumarage, Huan Liu

Large language models (LLMs) such as GPT-4, PaLM, and Llama have significantly propelled the generation of AI-crafted text. With rising concerns about their potential misuse, there is a pressing need for AI-generated-tex…

Authorship AttributionLanguage ModelingLanguage ModellingMisinformation

Figuratively Speaking: Authorship Attribution via Multi-Task Figurative Language Modeling

2024-06-12 · Gregorios A Katsios, Ning Sa, Tomek Strzalkowski

The identification of Figurative Language (FL) features in text is crucial for various Natural Language Processing (NLP) tasks, where understanding of the author's intended meaning and its nuances is key for successful c…

Authorship AttributionLanguage ModelingLanguage Modelling