Evaluating Conversational Recommender Systems via User Simulation
Conversational information access is an emerging research area. Currently, human evaluation is used for end-to-end system evaluation, which is both very time and resource intensive at scale, and thus becomes a bottleneck of progress. As an alternative, we propose automated evaluation by means of simulating users. Our user simulator aims to generate responses that a real human would give by considering both individual preferences and the general flow of interaction with the system. We evaluate our simulation approach on an item recommendation task by comparing three existing conversational recommender systems. We show that preference modeling and task-specific interaction models both contribute to more realistic simulations, and can help achieve high correlation between automatic evaluation measures and manual human assessments.
Code (1)
Tasks
Conversational Information AccessRecommendation SystemsUser SimulationSimilar Papers 제목 키워드 기반
UserSimCRS: A User Simulation Toolkit for Evaluating Conversational Recommender Systems
We present an extensible user simulation toolkit to facilitate automatic evaluation of conversational recommender systems. It builds on an established agenda-based approach and extends it with several novel elements, inc…
Recommendation SystemsText GenerationUser SimulationUser Simulation for Evaluating Information Access Systems
Information access systems, such as search engines, recommender systems, and conversational assistants, have become integral to our daily lives as they help us satisfy our information needs. However, evaluating the effec…
Recommendation SystemsUser SimulationAnalyzing and Simulating User Utterance Reformulation in Conversational Recommender Systems
User simulation has been a cost-effective technique for evaluating conversational recommender systems. However, building a human-like simulator is still an open challenge. In this work, we focus on how users reformulate …
Recommendation SystemsUser SimulationIdentifying Breakdowns in Conversational Recommender Systems using User Simulation
We present a methodology to systematically test conversational recommender systems with regards to conversational breakdowns. It involves examining conversations generated between the system and simulated users for a set…
Conversational RecommendationDiagnosticUser SimulationEvaluating Conversational Recommender Systems: A Landscape of Research
Conversational recommender systems aim to interactively support online users in their information search and decision-making processes in an intuitive way. With the latest advances in voice-controlled devices, natural la…
Decision MakingRecommendation Systems