paper-with-me

Papers

Behavior Alignment: A New Perspective of Evaluating LLM-based Conversational Recommender Systems

2024-04-17 · Dayu Yang, Fumian Chen, Hui Fang

Large Language Models (LLMs) have demonstrated great potential in Conversational Recommender Systems (CRS). However, the application of LLMs to CRS has exposed a notable discrepancy in behavior between LLM-based CRS and human recommenders: LLMs often appear inflexible and passive, frequently rushing to complete the recommendation task without sufficient inquiry.This behavior discrepancy can lead to decreased accuracy in recommendations and lower user satisfaction. Despite its importance, existing studies in CRS lack a study about how to measure such behavior discrepancy. To fill this gap, we propose Behavior Alignment, a new evaluation metric to measure how well the recommendation strategies made by a LLM-based CRS are consistent with human recommenders'. Our experiment results show that the new metric is better aligned with human preferences and can better differentiate how systems perform than existing evaluation metrics. As Behavior Alignment requires explicit and costly human annotations on the recommendation strategies, we also propose a classification-based method to implicitly measure the Behavior Alignment based on the responses. The evaluation results confirm the robustness of the method.

📄 PDF Abstract BibTeX arXiv:2404.11773

Code (1)

dayuyang1999/behavior-alignment 공식 구현 pytorch

Tasks

Conversational RecommendationRecommendation Systems

Similar Papers 제목 키워드 기반

Evaluating Conversational Recommender Systems with Large Language Models: A User-Centric Evaluation Framework

2025-01-16 · Nuo Chen, Quanyu Dai, Xiaoyu Dong, Xiao-Ming Wu 외

Conversational recommender systems (CRS) involve both recommendation and dialogue tasks, which makes their evaluation a unique challenge. Although past research has analyzed various factors that may affect user satisfact…

Recommendation Systems

Prompt Optimization for User Simulation in Conversational Recommender Systems: A Multi-Objective Framework

2026-05-08 · Nipun B Nair, Tongtong Wu, Weiqing Wang arxiv

Conversational recommender systems (CRSs) are a core component of next-generation intelligent recommender systems because they enable users to actively elicit preferences, clarify intentions, and adapt recommendations in…

Prompt Engineering

RecToM: A Benchmark for Evaluating Machine Theory of Mind in LLM-based Conversational Recommender Systems

2025-11-27 · Mengfan Li, Xuanhua Shi, Yang Deng arxiv

Large Language models are revolutionizing the conversational recommender systems through their impressive capabilities in instruction comprehension, reasoning, and human interaction. A core factor underlying effective re…

Reformulating Conversational Recommender Systems as Tri-Phase Offline Policy Learning

2024-08-13 · Gangyi Zhang, Chongming Gao, Hang Pan, Runzhe Teng 외

Existing Conversational Recommender Systems (CRS) predominantly utilize user simulators for training and evaluating recommendation policies. These simulators often oversimplify the complexity of user interactions by focu…

Recommendation SystemsUser Simulation

Analyzing and Simulating User Utterance Reformulation in Conversational Recommender Systems

2022-05-03 · Shuo Zhang, Mu-Chun Wang, Krisztian Balog

User simulation has been a cost-effective technique for evaluating conversational recommender systems. However, building a human-like simulator is still an open challenge. In this work, we focus on how users reformulate …

Recommendation SystemsUser Simulation