paper-with-me

홈 › Papers

Search-Based Interaction For Conversation Recommendation via Generative Reward Model Based Simulated User

2025-04-29 · Xiaolei Wang, Chunxuan Xia, Junyi Li, Fanzhe Meng, Lei Huang, Jinpeng Wang, Wayne Xin Zhao, Ji-Rong Wen

Conversational recommendation systems (CRSs) use multi-turn interaction to capture user preferences and provide personalized recommendations. A fundamental challenge in CRSs lies in effectively understanding user preferences from conversations. User preferences can be multifaceted and complex, posing significant challenges for accurate recommendations even with access to abundant external knowledge. While interaction with users can clarify their true preferences, frequent user involvement can lead to a degraded user experience. To address this problem, we propose a generative reward model based simulated user, named GRSU, for automatic interaction with CRSs. The simulated user provides feedback to the items recommended by CRSs, enabling them to better capture intricate user preferences through multi-turn interaction. Inspired by generative reward models, we design two types of feedback actions for the simulated user: i.e., generative item scoring, which offers coarse-grained feedback, and attribute-based item critique, which provides fine-grained feedback. To ensure seamless integration, these feedback actions are unified into an instruction-based format, allowing the development of a unified simulated user via instruction tuning on synthesized data. With this simulated user, automatic multi-turn interaction with CRSs can be effectively conducted. Furthermore, to strike a balance between effectiveness and efficiency, we draw inspiration from the paradigm of reward-guided search in complex reasoning tasks and employ beam search for the interaction process. On top of this, we propose an efficient candidate ranking method to improve the recommendation results derived from interaction. Extensive experiments on public datasets demonstrate the effectiveness, efficiency, and transferability of our approach.

📄 PDF Abstract BibTeX arXiv:2504.20458

Code (0)

등록된 구현이 없습니다.

Tasks

AttributeConversational RecommendationRecommendation Systems

Similar Papers 제목 키워드 기반

MobileConvRec: A Conversational Dataset for Mobile Apps Recommendations

2024-05-28 · Srijata Maji, Moghis Fereidouni, Vinaik Chhetri, Umar Farooq 외

Existing recommendation systems have focused on two paradigms: 1- historical user-item interaction-based recommendations and 2- conversational recommendations. Conversational recommendation systems facilitate natural lan…

Conversational RecommendationRecommendation Systems

Multi-Objective Intrinsic Reward Learning for Conversational Recommender Systems

2023-10-31 · NeurIPS 2023 11

Conversational Recommender Systems (CRS) actively elicit user preferences to generate adaptive recommendations. Mainstream reinforcement learning-based CRS solutions heavily rely on handcrafted reward functions, which ma…

Recommendation Systems

Hierarchical Conversational Preference Elicitation with Bandit Feedback

2022-09-06 · Jinhang Zuo, Songwen Hu, Tong Yu, Shuai Li 외

The recent advances of conversational recommendations provide a promising way to efficiently elicit users' preferences via conversational interactions. To achieve this, the recommender system conducts conversations with …

Recommendation Systems

Ask to Be Sure: Informative Interactions for Confident Multi-Turn LLM Recommendation

2026-08-16 · Cedar Site Bai, Zhenyu Liao, Duanshun Li, Sheikh Sarwar 외 arxiv

Recent advances in large language models (LLMs) have enabled their use as conversational recommender systems (CRS), demonstrating strong recommendation accuracy and natural dialogue. However, guiding multi-turn interacti…

Reinforcement Learning

Generative Inverse Deep Reinforcement Learning for Online Recommendation

2020-11-04 · Xiaocong Chen, Lina Yao, Aixin Sun, Xianzhi Wang 외

Deep reinforcement learning enables an agent to capture user's interest through interactions with the environment dynamically. It has attracted great interest in the recommendation research. Deep reinforcement learning u…

Deep Reinforcement Learningreinforcement-learningReinforcement LearningReinforcement Learning (RL)