paper-with-me

Papers

Retell, Reward, Repeat: Reinforcement Learning for Narrative Theory-Informed Story Retelling

2026-01-23 · David Y. Liu, Xanthe Muston, Dipankar Srirag, Aditya Joshi, Sebastian Sequoiah-Grayson arxiv

Counterfactual story retelling exposes LLM shortcomings in constrained narrative solution spaces where they can no longer rely on recalling memorised training data. Ground-truth-based post-training, such as SFT, fails to teach LLMs how to generate logical and rational narrative events. In this paper, we introduce Retell, Reward, Repeat (RRR), an RL-based pipeline synthesising Structuralist Narratology with scalar narrativity to teach storytelling structure. We extend the TimeTravel dataset with human-annotated stages of narrative equilibrium to evaluate reward models. By using d-RLAIF, RRR derives training signals from the narrativity of textual features without the need for reference outputs. Evaluations demonstrate that RRR-trained LLMs outperform few-shot and SFT baselines in logic, rationality, and completeness, with output quality additionally validated by blind human preference. Relying on a small, query-only dataset, RRR provides a linguistically grounded, cost-effective post-training mechanism for storytelling--a domain currently lacking effective post-training methods. RRR highlights the continued relevance of integrating established linguistic theories into contemporary NLP.

📄 PDF Abstract BibTeX arXiv:2601.17226

Code (0)

등록된 구현이 없습니다.

Tasks

Reinforcement Learning

Similar Papers 제목 키워드 기반

Discriminative Joint Modeling of Lexical Variation and Acoustic Confusion for Automated Narrative Retelling Assessment

2013-06-01 · NAACL 2013 6 · Maider Lehr, Izhak Shafran, Emily Prud{'}hommeaux, Brian Roark
Machine TranslationReading ComprehensionSpeech RecognitionSpoken Language Understanding+1

Tell, Don't Show: Leveraging Language Models' Abstractive Retellings to Model Literary Themes

2025-05-29 · Li Lucy, Camilla Griffiths, Sarah Levine, Jennifer L. Eberhardt 외

Conventional bag-of-words approaches for topic modeling, like latent Dirichlet allocation (LDA), struggle with literary text. Literature challenges lexical methods because narrative language focuses on immersive sensory …

PersonaBank: A Corpus of Personal Narratives and Their Story Intention Graphs

2017-08-30 · LREC 2016 5 · Stephanie M. Lukin, Kevin Bowden, Casey Barackman, Marilyn A. Walker

We present a new corpus, PersonaBank, consisting of 108 personal stories from weblogs that have been annotated with their Story Intention Graphs, a deep representation of the fabula of a story. We describe the topics of …

Can Base ChatGPT be Used for Forecasting without Additional Optimization?

2024-04-11 · Van Pham, Scott Cunningham

This study investigates whether OpenAI's ChatGPT-3.5 and ChatGPT-4 can forecast future events. To evaluate the accuracy of the predictions, we take advantage of the fact that the training data at the time of our experime…

GNAT: A General Narrative Alignment Tool

2023-11-07 · Tanzir Pial, Steven Skiena

Algorithmic sequence alignment identifies similar segments shared between pairs of documents, and is fundamental to many NLP tasks. But it is difficult to recognize similarities between distant versions of narratives suc…

text similarity