paper-with-me

홈 › Papers

StoryScope: Investigating idiosyncrasies in AI fiction

2026-04-03 · Jenna Russell, Rishanth Rajendhran, Chau Minh Pham, Mohit Iyyer, John Wieting arxiv

As AI-generated fiction becomes increasingly prevalent, questions of authorship and originality are becoming central to how written work is evaluated. While most existing work in this space focuses on identifying surface-level signatures of AI writing, we ask instead whether AI-generated stories can be distinguished from human ones without relying on stylistic signals, focusing on discourse-level narrative choices such as character agency and chronological discontinuity. We propose StoryScope, a pipeline that automatically induces a fine-grained, interpretable feature space of discourse-level narrative features across 10 dimensions. We apply StoryScope to a parallel corpus of 10,272 writing prompts, each written by a human author and five LLMs, yielding 61,608 stories, each ~5,000 words, and 304 extracted features per story. Narrative features alone achieve 93.2% macro-F1 for human vs. AI detection and 68.4% macro-F1 for six-way authorship attribution, retaining over 97% of the performance of models that include stylistic cues. A compact set of 30 core narrative features captures much of this signal: AI stories over-explain themes and favor tidy, single-track plots while human stories frame protagonist' choices as more morally ambiguous and have increased temporal complexity. Per-model fingerprint features enable six-way attribution: for example, Claude produces notably flat event escalation, GPT over-indexes on dream sequences, and Gemini defaults to external character description. We find that AI-generated stories cluster in a shared region of narrative space, while human-authored stories exhibit greater diversity. More broadly, these results suggest that differences in underlying narrative construction, not just writing style, can be used to separate human-written original works from AI-generated fiction.

📄 PDF Abstract BibTeX arXiv:2604.03136

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Shirtless and Dangerous: Quantifying Linguistic Signals of Gender Bias in an Online Fiction Writing Community

2016-03-29 · Ethan Fast, Tina Vachovsky, Michael S. Bernstein

Imagine a princess asleep in a castle, waiting for her prince to slay the dragon and rescue her. Tales like the famous Sleeping Beauty clearly divide up gender roles. But what about more modern stories, borne of a genera…

Idiosyncrasies in Large Language Models

2025-02-17 · MingJie Sun, Yida Yin, Zhiqiu Xu, J. Zico Kolter 외

In this work, we unveil and study idiosyncrasies in Large Language Models (LLMs) -- unique patterns in their outputs that can be used to distinguish the models. To do so, we consider a simple classification task: given a…

Idiosyncrasies and challenges of data driven learning in electronic trading

2018-11-30

We outline the idiosyncrasies of neural information processing and machine learning in quantitative finance. We also present some of the approaches we take towards solving the fundamental challenges we face.

BIG-bench Machine Learning

Investigating the Scalability of Approximate Sparse Retrieval Algorithms to Massive Datasets

2025-01-20 · Sebastian Bruch, Franco Maria Nardini, Cosimo Rulli, Rossano Venturini 외

Learned sparse text embeddings have gained popularity due to their effectiveness in top-k retrieval and inherent interpretability. Their distributional idiosyncrasies, however, have long hindered their use in real-world …

Retrieval

Levels of Non-Fictionality in Fictional Texts

2022-06-01 · ISA (LREC) 2022 6 · Florian Barth, Hanna Varachkina, Tillmann Dönicke, Luisa Gödeke

The annotation and automatic recognition of non-fictional discourse within a text is an important, yet unresolved task in literary research. While non-fictional passages can consist of several clauses or sentences, we ar…