paper-with-me

홈 › Papers

Everyone prefers human writers, including AI

2025-10-09 · Wouter Haverals, Meredith Martin arxiv

As AI writing tools become widespread, we need to understand how both humans and machines evaluate literary style, a domain where objective standards are elusive and judgments are inherently subjective. We conducted controlled experiments using Raymond Queneau's Exercises in Style (1947) to measure attribution bias across evaluators. Study 1 compared human participants (N=556) and AI models (N=13) evaluating literary passages from Queneau versus GPT-4-generated versions under three conditions: blind, accurately labeled, and counterfactually labeled. Study 2 tested bias generalization across a 14$\times$14 matrix of AI evaluators and creators. Both studies revealed systematic pro-human attribution bias. Humans showed +13.7 percentage point (pp) bias (Cohen's h = 0.28, 95% CI: 0.21-0.34), while AI models showed +34.3 percentage point bias (h = 0.70, 95% CI: 0.65-0.76), a 2.5-fold stronger effect (P$<$0.001). Study 2 confirmed this bias operates across AI architectures (+25.8pp, 95% CI: 24.1-27.6%), demonstrating that AI systems systematically devalue creative content when labeled as "AI-generated" regardless of which AI created it. We also find that attribution labels cause evaluators to invert assessment criteria, with identical features receiving opposing evaluations based solely on perceived authorship. This suggests AI models have absorbed human cultural biases against artificial creativity during training. Our study represents the first controlled comparison of attribution bias between human and artificial evaluators in aesthetic judgment, revealing that AI systems not only replicate but amplify this human tendency.

📄 PDF Abstract BibTeX arXiv:2510.08831

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Human diversity fuels collective creativity that large language models cannot simulate or sustain

2026-07-29 · Mengchen Dong, Hiromu Yakura arxiv

Diverse human groups produce diverse ideas, the raw material of innovation. Generative AI challenges this engine twice over: everyday AI assistance may homogenize what diverse people create, and AI-simulated diversity ma…

The Unlikely Duel: Evaluating Creative Writing in LLMs through a Unique Scenario

2024-06-22 · Carlos Gómez-Rodríguez, Paul Williams

This is a summary of the paper "A Confederacy of Models: a Comprehensive Evaluation of LLMs on Creative Writing", which was published in Findings of EMNLP 2023. We evaluate a range of recent state-of-the-art, instruction…

The AI Ghostwriter Effect: When Users Do Not Perceive Ownership of AI-Generated Text But Self-Declare as Authors

2023-03-06 · Fiona Draxler, Anna Werner, Florian Lehmann, Matthias Hoppe 외

Human-AI interaction in text production increases complexity in authorship. In two empirical studies (n1 = 30 & n2 = 96), we investigate authorship and ownership in human-AI collaboration for personalized language genera…

AttributeText Generation

Co-Writing with AI, on Human Terms: Aligning Research with User Demands Across the Writing Process

2025-04-16 · Mohi Reza, Jeb Thomas-Mitchell, Peter Dushniku, Nathan Laundry 외

As generative AI tools like ChatGPT become integral to everyday writing, critical questions arise about how to preserve writers' sense of agency and ownership when using these tools. Yet, a systematic understanding of ho…

"It was 80% me, 20% AI": Seeking Authenticity in Co-Writing with Large Language Models

2024-11-20 · Angel Hsing-Chi Hwang, Q. Vera Liao, Su Lin Blodgett, Alexandra Olteanu 외

Given the rising proliferation and diversity of AI writing assistance tools, especially those powered by large language models (LLMs), both writers and readers may have concerns about the impact of these tools on the aut…