paper-with-me

Papers

Simulating Before Planning: Constructing Intrinsic User World Model for User-Tailored Dialogue Policy Planning

2025-04-18 · Tao He, Lizi Liao, Ming Liu, Bing Qin

Recent advancements in dialogue policy planning have emphasized optimizing system agent policies to achieve predefined goals, focusing on strategy design, trajectory acquisition, and efficient training paradigms. However, these approaches often overlook the critical role of user characteristics, which are essential in real-world scenarios like conversational search and recommendation, where interactions must adapt to individual user traits such as personality, preferences, and goals. To address this gap, we first conduct a comprehensive study utilizing task-specific user personas to systematically assess dialogue policy planning under diverse user behaviors. By leveraging realistic user profiles for different tasks, our study reveals significant limitations in existing approaches, highlighting the need for user-tailored dialogue policy planning. Building on this foundation, we present the User-Tailored Dialogue Policy Planning (UDP) framework, which incorporates an Intrinsic User World Model to model user traits and feedback. UDP operates in three stages: (1) User Persona Portraying, using a diffusion model to dynamically infer user profiles; (2) User Feedback Anticipating, leveraging a Brownian Bridge-inspired anticipator to predict user reactions; and (3) User-Tailored Policy Planning, integrating these insights to optimize response strategies. To ensure robust performance, we further propose an active learning approach that prioritizes challenging user personas during training. Comprehensive experiments on benchmarks, including collaborative and non-collaborative settings, demonstrate the effectiveness of UDP in learning user-specific dialogue strategies. Results validate the protocol's utility and highlight UDP's robustness, adaptability, and potential to advance user-centric dialogue systems.

📄 PDF Abstract BibTeX arXiv:2504.13643

Code (0)

등록된 구현이 없습니다.

Tasks

Active LearningConversational Search

Methods 이 논문이 사용한 방법론

Diffusion Diffusion models generate samples by gradually removing noise from a signal, and their training objective can be expressed as a reweighted variational lower-bound…

Similar Papers 제목 키워드 기반

"Think Before You Speak": Improving Multi-Action Dialog Policy by Planning Single-Action Dialogs

2022-04-25 · Shuo Zhang, Junzhou Zhao, Pinghui Wang, Yu Li 외

Multi-action dialog policy (MADP), which generates multiple atomic dialog actions per turn, has been widely applied in task-oriented dialog systems to provide expressive and efficient system responses. Existing MADP mode…

Multi-Task Learning

USimAgent: Large Language Models for Simulating Search Users

2024-03-14 · Erhan Zhang, Xingzhu Wang, Peiyuan Gong, Yankai Lin 외

Due to the advantages in the cost-efficiency and reproducibility, user simulation has become a promising solution to the user-centric evaluation of information retrieval systems. Nonetheless, accurately simulating user s…

Information RetrievalUser Simulation

Simulating Makeup Through Physics-Based Manipulation of Intrinsic Image Layers

2015-06-01 · CVPR 2015 6 · Chen Li, Kun Zhou, Stephen Lin

We present a method for simulating makeup in a face image. To generate realistic results without detailed geometric and reflectance measurements of the user, we propose to separate the image into intrinsic image layers a…

Capturing, Reconstructing, and Simulating: the UrbanScene3D Dataset

2021-07-09 · Liqiang Lin, Yilin Liu, Yue Hu, Xingguang Yan 외

We present UrbanScene3D, a large-scale data platform for research of urban scene perception and reconstruction. UrbanScene3D contains over 128k high-resolution images covering 16 scenes including large-scale real urban r…

3D ReconstructionInstance SegmentationSemantic Segmentation

Ask-before-Plan: Proactive Language Agents for Real-World Planning

2024-06-18 · Xuan Zhang, Yang Deng, Zifeng Ren, See-Kiong Ng 외

The evolution of large language models (LLMs) has enhanced the planning capabilities of language agents in diverse real-world scenarios. Despite these advancements, the potential of LLM-powered agents to comprehend ambig…

Decision Makingvalid