paper-with-me

홈 › Papers

Help Me Write a Story: Evaluating LLMs' Ability to Generate Writing Feedback

2025-07-21 · Hannah Rashkin, Elizabeth Clark, Fantine Huot, Mirella Lapata arxiv

Can LLMs provide support to creative writers by giving meaningful writing feedback? In this paper, we explore the challenges and limitations of model-generated writing feedback by defining a new task, dataset, and evaluation frameworks. To study model performance in a controlled manner, we present a novel test set of 1,300 stories that we corrupted to intentionally introduce writing issues. We study the performance of commonly used LLMs in this task with both automatic and human evaluation metrics. Our analysis shows that current models have strong out-of-the-box behavior in many respects -- providing specific and mostly accurate writing feedback. However, models often fail to identify the biggest writing issue in the story and to correctly decide when to offer critical vs. positive feedback.

📄 PDF Abstract BibTeX arXiv:2507.16007

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Evaluating Creative Short Story Generation in Humans and Large Language Models

2024-11-04 · Mete Ismayilzada, Claire Stevenson, Lonneke van der Plas

Story-writing is a fundamental aspect of human imagination, relying heavily on creativity to produce narratives that are novel, effective, and surprising. While large language models (LLMs) have demonstrated the ability …

DiversitySentenceStory Generation

Pretraining Language Models on Historical Text

2026-06-02 · Xiaoxi Luo, Zachary Shinnick, Niclas Griesshaber, Yixuan Wang 외 arxiv

We introduce TypewriterLM, a 7.24B History language model (LM) trained exclusively on English text predating 1913. Developing History LMs requires addressing challenges in data quality and availability, preventing tempor…

Choose Your Own Adventure: Paired Suggestions in Collaborative Writing for Evaluating Story Generation Models

2021-06-01 · NAACL 2021 4 · Elizabeth Clark, Noah A. Smith

Story generation is an open-ended and subjective task, which poses a challenge for evaluating story generation models. We present Choose Your Own Adventure, a collaborative writing setup for pairwise model evaluation. Tw…

Story Generation

Generating Constructive Feedback on Stories via Reinforcement Learning

2026-09-04 · Maja Stahl, Timon Ziegenbein, Henning Wachsmuth arxiv

Constructive feedback is crucial for creative writers to refine their storytelling abilities. Since receiving feedback from human experts is often costly and time-intensive, large language models (LLMs) offer a scalable …

Reinforcement Learning

StoryWriter: A Multi-Agent Framework for Long Story Generation

2025-06-19 · Haotian Xia, Hao Peng, Yunjia Qi, Xiaozhi Wang 외

Long story generation remains a challenge for existing large language models (LLMs), primarily due to two main factors: (1) discourse coherence, which requires plot consistency, logical coherence, and completeness in the…

Story Generation