paper-with-me

홈 › Papers

MUSE: Multi-Domain Chinese User Simulation via Self-Evolving Profiles and Rubric-Guided Alignment

2026-04-15 · Zihao Liu, Hantao Zhou, Jiguo Li, Jun Xu, Jiuchong Gao, Jinghua Hao, Renqing He, Peng Wang arxiv

User simulators are essential for the scalable training and evaluation of interactive AI systems. However, existing approaches often rely on shallow user profiling, struggle to maintain persona consistency over long interactions, and are largely limited to English or single-domain settings. We present MUSE, a multi-domain Chinese user simulation framework designed to generate human-like, controllable, and behaviorally consistent responses. First, we propose Iterative Profile Self-Evolution (IPSE), which gradually optimizes user profiles by comparing and reasoning discrepancies between simulated trajectories and real dialogue behaviors. We then apply Role-Reversal Supervised Fine-Tuning to improve local response realism and human-like expression. To enable fine-grained behavioral alignment, we further train a specialized rubric-based reward model and incorporate it into rubric-guided multi-turn reinforcement learning, which optimizes the simulator at the dialogue level and enhances long-horizon behavioral consistency. Experiments show that MUSE consistently outperforms strong baselines in both utterance-level and session-level evaluations, generating responses that are more realistic, coherent, and persona-consistent over extended interactions.

📄 PDF Abstract BibTeX arXiv:2604.13828

Code (0)

등록된 구현이 없습니다.

Tasks

Reinforcement Learning

Similar Papers 제목 키워드 기반

MUSER: A Multi-View Similar Case Retrieval Dataset

2023-10-24 · Qingquan Li, Yiran Hu, Feng Yao, Chaojun Xiao 외

Similar case retrieval (SCR) is a representative legal AI application that plays a pivotal role in promoting judicial fairness. However, existing SCR datasets only focus on the fact description section when judging the s…

FairnessRetrievalSentencetext-classification+1

CrossWOZ: A Large-Scale Chinese Cross-Domain Task-Oriented Dialogue Dataset

2020-02-27 · TACL 2020 1 · Qi Zhu, Kaili Huang, Zheng Zhang, Xiaoyan Zhu 외

To advance multi-domain (cross-domain) dialogue modeling as well as alleviate the shortage of Chinese task-oriented datasets, we propose CrossWOZ, the first large-scale Chinese Cross-Domain Wizard-of-Oz task-oriented dat…

Dialogue State TrackingTask-Oriented Dialogue SystemsUser Simulation

MuSeCLIR: A Multiple Senses and Cross-lingual Information Retrieval Dataset

2022-10-01 · COLING 2022 10 · Wing Yan Li, Julie Weeds, David Weir

This paper addresses a deficiency in existing cross-lingual information retrieval (CLIR) datasets and provides a robust evaluation of CLIR systems’ disambiguation ability. CLIR is commonly tackled by combining translatio…

Cross-Lingual Information RetrievalInformation RetrievalRetrievalTranslation

VaseMuseum: Digital Intelligent Museum for Ancient Greek Pottery

2026-07-07 · Jiazi Wang, Nonghai Zhang, Qiushi Xie, Zeyu Zhang 외 arxiv

Vision-language models (VLMs) have made interactive digital museums increasingly feasible by connecting 3D digitization with natural-language artifact exploration. However, in cultural heritage domains such as ancient Gr…

MOSAIC: Multimodal Multistakeholder-aware Visual Art Recommendation

2024-07-31 · Bereket A. Yilma, Luis A. Leiva

Visual art (VA) recommendation is complex, as it has to consider the interests of users (e.g. museum visitors) and other stakeholders (e.g. museum curators). We study how to effectively account for key stakeholders in VA…

Diversity