paper-with-me

홈 › Papers

Know You First and Be You Better: Modeling Human-Like User Simulators via Implicit Profiles

2025-02-26 · Kuang Wang, Xianfei Li, Shenghao Yang, Li Zhou, Feng Jiang, Haizhou Li

User simulators are crucial for replicating human interactions with dialogue systems, supporting both collaborative training and automatic evaluation, especially for large language models (LLMs). However, existing simulators often rely solely on text utterances, missing implicit user traits such as personality, speaking style, and goals. In contrast, persona-based methods lack generalizability, as they depend on predefined profiles of famous individuals or archetypes. To address these challenges, we propose User Simulator with implicit Profiles (USP), a framework that infers implicit user profiles from human-machine conversations and uses them to generate more personalized and realistic dialogues. We first develop an LLM-driven extractor with a comprehensive profile schema. Then, we refine the simulation through conditional supervised fine-tuning and reinforcement learning with cycle consistency, optimizing it at both the utterance and conversation levels. Finally, we adopt a diverse profile sampler to capture the distribution of real-world user profiles. Experimental results demonstrate that USP outperforms strong baselines in terms of authenticity and diversity while achieving comparable performance in consistency. Furthermore, dynamic multi-turn evaluations based on USP strongly align with mainstream benchmarks, demonstrating its effectiveness in real-world applications.

📄 PDF Abstract BibTeX arXiv:2502.18968

Code (1)

wangkevin02/USP 공식 구현 pytorch

Methods 이 논문이 사용한 방법론

ALIGN In the ALIGN method, visual and language representations are jointly trained from noisy image alt-text data. The image and text encoders are learned via contrastive loss…
ADOPT Please enter a description about the method here

Similar Papers 제목 키워드 기반

SocialGen: Modeling Multi-Human Social Interaction with Language Models

2025-03-28 · Heng Yu, Juze Zhang, Changan Chen, Tiange Xiang 외

Human interactions in everyday life are inherently social, involving engagements with diverse individuals across various contexts. Modeling these social interactions is fundamental to a wide range of real-world applicati…

Beyond Binary Preferences: A Principled Framework for Reward Modeling with Ordinal Feedback

2026-02-13 · Amirhossein Afsharrad, Ruida Zhou, Luca Viano, Sanjay Lall 외 arxiv

Reward modeling is crucial for aligning large language models with human preferences, yet current approaches lack a principled mathematical framework for leveraging ordinal preference data. When human annotators provide …

On Faithfulness and Factuality in Abstractive Summarization

2020-05-02 · ACL 2020 6 · Joshua Maynez, Shashi Narayan, Bernd Bohnet, Ryan Mcdonald

It is well known that the standard likelihood training and approximate decoding objectives in neural text generation models lead to less human-like responses for open-ended tasks such as language modeling and story gener…

Abstractive Text SummarizationDocument SummarizationLanguage ModelingLanguage Modelling+3

Contrast Sensitivity Functions in Autoencoders

2021-02-28 · Qiang Li, Alex Gomez-Villa, Marcelo Bertalmio, Jesus Malo

Three decades ago, Atick et al. suggested that human frequency sensitivity may emerge from the enhancement required for a more efficient analysis of retinal images. Here we reassess the relevance of low-level vision task…

Sensitivity

PREF: Predictability Regularized Neural Motion Fields

2022-09-21 · Liangchen Song, Xuan Gong, Benjamin Planche, Meng Zheng 외

Knowing the 3D motions in a dynamic scene is essential to many vision applications. Recent progress is mainly focused on estimating the activity of some specific elements like humans. In this paper, we leverage a neural …