paper-with-me

Papers

RoleRMBench & RoleRM: Towards Reward Modeling for Profile-Based Role Play in Dialogue Systems

2025-12-11 · Hang Ding, Qiming Feng, Dongqi Liu, Qi Zhao, Tao Yao, Shuo Wang, Dongsheng Chen, Jian Li, Zhenye Gan, Jiangning Zhang, Chengjie Wang, Yabiao Wang arxiv

Reward modeling has become a cornerstone of aligning large language models (LLMs) with human preferences. Yet, when extended to subjective and open-ended domains such as role play, existing reward models exhibit severe degradation, struggling to capture nuanced and persona-grounded human judgments. To address this gap, we introduce RoleRMBench, the first systematic benchmark for reward modeling in role-playing dialogue, covering seven fine-grained capabilities from narrative management to role consistency and engagement. Evaluation on RoleRMBench reveals large and consistent gaps between general-purpose reward models and human judgment, particularly in narrative and stylistic dimensions. We further propose RoleRM, a reward model trained with Continuous Implicit Preferences (CIP), which reformulates subjective evaluation as continuous consistent pairwise supervision under multiple structuring strategies. Comprehensive experiments show that RoleRM surpasses strong open- and closed-source reward models by over 24% on average, demonstrating substantial gains in narrative coherence and stylistic fidelity. Our findings highlight the importance of continuous preference representation and annotation consistency, establishing a foundation for subjective alignment in human-centered dialogue systems.

📄 PDF Abstract BibTeX arXiv:2512.10575

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Improving General Role-Playing Agents via Psychology-Grounded Reasoning and Role-Aware Policy Optimization

2026-06-25 · Zhenhua Xu, Dongsheng Chen, Jian Li, Yitong Lin 외 arxiv

Building general-purpose role-playing agents that faithfully portray any character from a natural-language profile remains challenging. The dominant paradigm -- supervised fine-tuning -- encourages behavioral mimicry wit…

Reinforcement Learning

UserLM-R1: Modeling Human Reasoning in User Language Models with Multi-Reward Reinforcement Learning

2026-01-14 · Feng Zhang, Shijia Li, Chunmao Zhang, Zhanyu Ma 외 arxiv

User simulators serve as the critical interactive environment for agent post-training, and an ideal user simulator generalizes across domains and proactively engages in negotiation by challenging or bargaining. However, …

Reinforcement Learning

User Profile with Large Language Models: Construction, Updating, and Benchmarking

2025-02-15 · Nusrat Jahan Prottasha, Md Kowsher, Hafijur Raman, Israt Jahan Anny 외

User profile modeling plays a key role in personalized systems, as it requires building accurate profiles and updating them with new information. In this paper, we present two high-quality open-source user profile datase…

BenchmarkingProfile Generation

Explaining heterogeneity in medial entorhinal cortex with task-driven neural networks

2021-12-01 · NeurIPS 2021 12 · Aran Nayebi, Alexander Attinger, Malcolm Campbell, Kiah Hardcastle 외

Medial entorhinal cortex (MEC) supports a wide range of navigational and memory related behaviors.Well-known experimental results have revealed specialized cell types in MEC --- e.g. grid, border, and head-direction cell…

Self-supervised User Profile Generation for Personalization

2026-06-03 · Clark Mingxuan Ju, Yuwei Qiu, Tong Zhao, Neil Shah arxiv

Personalizing large language models (LLMs) has become a central challenge as LLMs are deployed across recommendation, search, dialogue, and content generation -- settings where the same query should yield different answe…

Profile Generation