paper-with-me

홈 › Papers

StoryAlign: Evaluating and Training Reward Models for Story Generation

2026-05-06 · Haotian Xia, Hao Peng, Yunjia Qi, Xiaozhi Wang, Bin Xu, Lei Hou, Juanzi Li arxiv

Story generation aims to automatically produce coherent, structured, and engaging narratives. Although large language models (LLMs) have significantly advanced text generation, stories generated by LLMs still diverge from human-authored works regarding complex narrative structure and human-aligned preferences. A key reason is the absence of effective modeling of human story preferences, which are inherently subjective and under-explored. In this work, we systematically evaluate the modeling of human story preferences and introduce StoryRMB, the first benchmark for assessing reward models on story preferences. StoryRMB contains $1,133$ high-quality, human-verified instances, each consisting of a prompt, one chosen story, and three rejected stories. We find existing reward models struggle to select human-preferred stories, with the best model achieving only $66.3\%$ accuracy. To address this limitation, we construct roughly $100,000$ high-quality story preference pairs across diverse domains and develop StoryReward, an advanced reward model for story preference trained on this dataset. StoryReward achieves state-of-the-art (SoTA) performance on StoryRMB, outperforming much larger models. We also adopt StoryReward in downstream test-time scaling applications for best-of-n (BoN) story selection and find that it generally chooses stories better aligned with human preferences. We will release our dataset, model, and code to facilitate future research. Related code and data are available at https://github.com/THU-KEG/StoryReward.

📄 PDF Abstract BibTeX arXiv:2605.04831

Code (0)

등록된 구현이 없습니다.

Tasks

Story GenerationText Generation

Similar Papers 제목 키워드 기반

Capturing Classic Authorial Style in Long-Form Story Generation with GRPO Fine-Tuning

2025-12-05 · Jinlong Liu, Mohammed Bahja, Venelin Kovatchev, Mark Lee arxiv

Evaluating and optimising authorial style in long-form story generation remains challenging because style is often assessed with ad hoc prompting and is frequently conflated with overall writing quality. We propose a two…

Story GenerationStyle Transfer

Controllable Neural Story Plot Generation via Reward Shaping

2018-09-27 · Pradyumna Tambwekar, Murtaza Dhuliawala, Lara J. Martin, Animesh Mehta 외

Language-modeling--based approaches to story plot generation attempt to construct a plot by sampling from a language model (LM) to predict the next character, word, or sentence to add to the story. LM techniques lack the…

Language ModelingLanguage Modellingreinforcement-learningReinforcement Learning+4

From Plots to Endings: A Reinforced Pointer Generator for Story Ending Generation

2019-01-11 · Yan Zhao, Lu Liu, Chunhua Liu, Ruoyao Yang 외

We introduce a new task named Story Ending Generation (SEG), whic-h aims at generating a coherent story ending from a sequence of story plot. Wepropose a framework consisting of a Generator and a Reward Manager for thist…

Learning to Reason for Long-Form Story Generation

2025-03-28 · Alexander Gurung, Mirella Lapata

Generating high-quality stories spanning thousands of tokens requires competency across a variety of skills, from tracking plot and character arcs to keeping a consistent and engaging style. Due to the difficulty of sour…

FormMathStory Generation

Choose Your Own Adventure: Paired Suggestions in Collaborative Writing for Evaluating Story Generation Models

2021-06-01 · NAACL 2021 4 · Elizabeth Clark, Noah A. Smith

Story generation is an open-ended and subjective task, which poses a challenge for evaluating story generation models. We present Choose Your Own Adventure, a collaborative writing setup for pairwise model evaluation. Tw…

Story Generation