paper-with-me

홈 › Papers

Offline Evaluation for Reinforcement Learning-based Recommendation: A Critical Issue and Some Alternatives

2023-01-03 · Romain Deffayet, Thibaut Thonet, Jean-Michel Renders, Maarten de Rijke

In this paper, we argue that the paradigm commonly adopted for offline evaluation of sequential recommender systems is unsuitable for evaluating reinforcement learning-based recommenders. We find that most of the existing offline evaluation practices for reinforcement learning-based recommendation are based on a next-item prediction protocol, and detail three shortcomings of such an evaluation protocol. Notably, it cannot reflect the potential benefits that reinforcement learning (RL) is expected to bring while it hides critical deficiencies of certain offline RL agents. Our suggestions for alternative ways to evaluate RL-based recommender systems aim to shed light on the existing possibilities and inspire future research on reliable evaluation protocols.

📄 PDF Abstract BibTeX arXiv:2301.00993

Code (0)

등록된 구현이 없습니다.

Tasks

Offline RLRecommendation Systemsreinforcement-learningReinforcement LearningReinforcement Learning (RL)

Similar Papers 제목 키워드 기반

Empowering Clinicians with Medical Decision Transformers: A Framework for Sepsis Treatment

2024-07-28 · Aamer Abdul Rahman, Pranav Agarwal, Rita Noumeir, Philippe Jouvet 외

Offline reinforcement learning has shown promise for solving tasks in safety-critical settings, such as clinical decision support. Its application, however, has been limited by the lack of interpretability and interactiv…

Off-policy evaluationreinforcement-learningReinforcement Learning

Assessing Fashion Recommendations: A Multifaceted Offline Evaluation Approach

2019-09-05 · Jake Sherman, Chinmay Shukla, Rhonda Textor, Su Zhang 외

Fashion is a unique domain for developing recommender systems (RS). Personalization is critical to fashion users. As a result, highly accurate recommendations are not sufficient unless they are also specific to users. Mo…

Collaborative FilteringRecommendation Systems

A Critical Study on Data Leakage in Recommender System Offline Evaluation

2020-10-21 · Yitong Ji, Aixin Sun, Jie Zhang, Chenliang Li

Recommender models are hard to evaluate, particularly under offline setting. In this paper, we provide a comprehensive and critical analysis of the data leakage issue in recommender system offline evaluation. Data leakag…

Collaborative FilteringRecommendation Systems

Accelerating Offline Reinforcement Learning Application in Real-Time Bidding and Recommendation: Potential Use of Simulation

2021-09-17 · Haruka Kiyohara, Kosuke Kawakami, Yuta Saito

In recommender systems (RecSys) and real-time bidding (RTB) for online advertisements, we often try to optimize sequential decision making using bandit and reinforcement learning (RL) techniques. In these applications, o…

Decision MakingOffline RLOff-policy evaluationPosition+4

An Efficient Continuous Control Perspective for Reinforcement-Learning-based Sequential Recommendation

2024-08-15 · Jun Wang, Likang Wu, Qi Liu, Yu Yang

Sequential recommendation, where user preference is dynamically inferred from sequential historical behaviors, is a critical task in recommender systems (RSs). To further optimize long-term user engagement, offline reinf…

continuous-controlContinuous ControlRecommendation SystemsSequential Recommendation