Offline Evaluation for Reinforcement Learning-based Recommendation: A Critical Issue and Some Alternatives
In this paper, we argue that the paradigm commonly adopted for offline evaluation of sequential recommender systems is unsuitable for evaluating reinforcement learning-based recommenders. We find that most of the existing offline evaluation practices for reinforcement learning-based recommendation are based on a next-item prediction protocol, and detail three shortcomings of such an evaluation protocol. Notably, it cannot reflect the potential benefits that reinforcement learning (RL) is expected to bring while it hides critical deficiencies of certain offline RL agents. Our suggestions for alternative ways to evaluate RL-based recommender systems aim to shed light on the existing possibilities and inspire future research on reliable evaluation protocols.
Code (0)
등록된 구현이 없습니다.
Tasks
Offline RLRecommendation Systemsreinforcement-learningReinforcement LearningReinforcement Learning (RL)Similar Papers 제목 키워드 기반
Empowering Clinicians with Medical Decision Transformers: A Framework for Sepsis Treatment
Offline reinforcement learning has shown promise for solving tasks in safety-critical settings, such as clinical decision support. Its application, however, has been limited by the lack of interpretability and interactiv…
Off-policy evaluationreinforcement-learningReinforcement LearningAssessing Fashion Recommendations: A Multifaceted Offline Evaluation Approach
Fashion is a unique domain for developing recommender systems (RS). Personalization is critical to fashion users. As a result, highly accurate recommendations are not sufficient unless they are also specific to users. Mo…
Collaborative FilteringRecommendation SystemsA Critical Study on Data Leakage in Recommender System Offline Evaluation
Recommender models are hard to evaluate, particularly under offline setting. In this paper, we provide a comprehensive and critical analysis of the data leakage issue in recommender system offline evaluation. Data leakag…
Collaborative FilteringRecommendation SystemsAccelerating Offline Reinforcement Learning Application in Real-Time Bidding and Recommendation: Potential Use of Simulation
In recommender systems (RecSys) and real-time bidding (RTB) for online advertisements, we often try to optimize sequential decision making using bandit and reinforcement learning (RL) techniques. In these applications, o…
Decision MakingOffline RLOff-policy evaluationPosition+4An Efficient Continuous Control Perspective for Reinforcement-Learning-based Sequential Recommendation
Sequential recommendation, where user preference is dynamically inferred from sequential historical behaviors, is a critical task in recommender systems (RSs). To further optimize long-term user engagement, offline reinf…
continuous-controlContinuous ControlRecommendation SystemsSequential Recommendation