paper-with-me

홈 › Papers

AI, write an essay for me: A large-scale comparison of human-written versus ChatGPT-generated essays

2023-04-24 · Steffen Herbold, Annette Hautli-Janisz, Ute Heuer, Zlata Kikteva, Alexander Trautsch

Background: Recently, ChatGPT and similar generative AI models have attracted hundreds of millions of users and become part of the public discourse. Many believe that such models will disrupt society and will result in a significant change in the education system and information generation in the future. So far, this belief is based on either colloquial evidence or benchmarks from the owners of the models -- both lack scientific rigour. Objective: Through a large-scale study comparing human-written versus ChatGPT-generated argumentative student essays, we systematically assess the quality of the AI-generated content. Methods: A large corpus of essays was rated using standard criteria by a large number of human experts (teachers). We augment the analysis with a consideration of the linguistic characteristics of the generated essays. Results: Our results demonstrate that ChatGPT generates essays that are rated higher for quality than human-written essays. The writing style of the AI models exhibits linguistic characteristics that are different from those of the human-written essays, e.g., it is characterized by fewer discourse and epistemic markers, but more nominalizations and greater lexical diversity. Conclusions: Our results clearly demonstrate that models like ChatGPT outperform humans in generating argumentative essays. Since the technology is readily available for anyone to use, educators must act immediately. We must re-invent homework and develop teaching concepts that utilize these AI models in the same way as math utilized the calculator: teach the general concepts first and then use AI tools to free up time for other learning objectives.

📄 PDF Abstract BibTeX arXiv:2304.14276

Code (0)

등록된 구현이 없습니다.

Tasks

Math

Similar Papers 제목 키워드 기반

Poor Alignment and Steerability of Large Language Models: Evidence from College Admission Essays

2025-03-25 · Jinsook Lee, AJ Alvero, Thorsten Joachims, René Kizilcec

People are increasingly using technologies equipped with large language models (LLM) to write texts for formal communication, which raises two important questions at the intersection of technology and society: Who do LLM…

FOXGLOVE: Understanding Goal-Oriented and Anchored Writing Feedback from Experts and LLMs on Argumentative Essays

2026-06-04 · Yijun Liu, Yifan Song, John Gallagher, Sarah Sterman 외 arxiv

While large language models (LLMs) are increasingly used to generate writing feedback, there remains no systematic comparison of LLM and expert feedback on the dimensions that writing research identifies as central to re…

Automated Essay Scoring Using Grammatical Variety and Errors with Multi-Task Learning and Item Response Theory

2024-06-13 · Kosuke Doi, Katsuhito Sudoh, Satoshi Nakamura

This study examines the effect of grammatical features in automatic essay scoring (AES). We use two kinds of grammatical features as input to an AES model: (1) grammatical items that writers used correctly in essays, and…

Automated Essay ScoringMulti-Task Learning

LLM-Based Persuasion Enables Guardrail Override in Frontier LLMs

2026-05-13 · Rodrigo Nogueira, Thales Sales Almeida, Giovana Kerche Bonás, Andrea Roque 외 arxiv

Frontier assistant LLMs ship with strong guardrails: asked directly to write a persuasive essay denying the Holocaust, denying vaccine safety, defending flat-earth cosmology, arguing for racial hierarchies, denying anthr…

A Corpus of Annotated Revisions for Studying Argumentative Writing

2017-07-01 · ACL 2017 7 · Fan Zhang, Homa B. Hashemi, Rebecca Hwa, Diane Litman

This paper presents ArgRewrite, a corpus of between-draft revisions of argumentative essays. Drafts are manually aligned at the sentence level, and the writer{'}s purpose for each revision is annotated with categories an…

Argument MiningSentence