paper-with-me

홈 › Papers

Offline A/B testing for Recommender Systems

2018-01-22 · Alexandre Gilotte, Clément Calauzènes, Thomas Nedelec, Alexandre Abraham, Simon Dollé

Before A/B testing online a new version of a recommender system, it is usual to perform some offline evaluations on historical data. We focus on evaluation methods that compute an estimator of the potential uplift in revenue that could generate this new technology. It helps to iterate faster and to avoid losing money by detecting poor policies. These estimators are known as counterfactual or off-policy estimators. We show that traditional counterfactual estimators such as capped importance sampling and normalised importance sampling are experimentally not having satisfying bias-variance compromises in the context of personalised product recommendation for online advertising. We propose two variants of counterfactual estimates with different modelling of the bias that prove to be accurate in real-world conditions. We provide a benchmark of these estimators by showing their correlation with business metrics observed by running online A/B tests on a commercial recommender system.

📄 PDF Abstract BibTeX arXiv:1801.07030

Code (0)

등록된 구현이 없습니다.

Tasks

counterfactualProduct RecommendationRecommendation Systems

Similar Papers 제목 키워드 기반

Accelerated learning from recommender systems using multi-armed bandit

2019-08-16 · Meisam Hejazinia, Kyler Eastman, Shuqin Ye, Abbas Amirabadi 외

Recommendation systems are a vital component of many online marketplaces, where there are often millions of items to potentially present to users who have a wide variety of wants or needs. Evaluating recommender system a…

Recommendation Systems

ZoRRO: A Zero-Weight Personalized Recommender System for Scalable News Recommendation

2026-07-12 · Johannes Kruse, Ryotaro Shimizu, Kasper Lindskow, Jon Tofteskov 외 arxiv

We present ZoRRO (Zero-Weight Personalized Recommender System), a zero-weight, training-free framework for personalized news recommendation designed for scalable real-world deployment. ZoRRO outperforms strong neural bas…

On the Opportunities and Challenges of Offline Reinforcement Learning for Recommender Systems

2023-08-22 · Xiaocong Chen, Siyu Wang, Julian McAuley, Dietmar Jannach 외

Reinforcement learning serves as a potent tool for modeling dynamic user interests within recommender systems, garnering increasing research attention of late. However, a significant drawback persists: its poor data effi…

Recommendation Systemsreinforcement-learningReinforcement Learning

Beyond Offline A/B Testing: Context-Aware Agent Simulation for Recommender System Evaluation

2026-01-26 · Nicolas Bougie, Gian Maria Marconi, Xiaotong Ye, Narimasa Watanabe arxiv

Recommender systems are central to online services, enabling users to navigate through massive amounts of content across various domains. However, their evaluation remains challenging due to the disconnect between offlin…

Where Do We Go From Here? Guidelines For Offline Recommender Evaluation

2022-11-02 · Tobias Schnabel

Various studies in recent years have pointed out large issues in the offline evaluation of recommender systems, making it difficult to assess whether true progress has been made. However, there has been little research i…

Hyperparameter OptimizationRecommendation SystemsUncertainty Quantification