paper-with-me

홈 › Papers

Genie: An Open Box Counterfactual Policy Estimator for Optimizing Sponsored Search Marketplace

2018-08-22 · Murat Ali Bayir, Mingsen Xu, Yaojia Zhu, Yifan Shi

In this paper, we propose an offline counterfactual policy estimation framework called Genie to optimize Sponsored Search Marketplace. Genie employs an open box simulation engine with click calibration model to compute the KPI impact of any modification to the system. From the experimental results on Bing traffic, we showed that Genie performs better than existing observational approaches that employs randomized experiments for traffic slices that have frequent policy updates. We also show that Genie can be used to tune completely new policies efficiently without creating risky randomized experiments due to cold start problem. As time of today, Genie hosts more than 10000 optimization jobs yearly which runs more than 30 Million processing node hours of big data jobs for Bing Ads. For the last 3 years, Genie has been proven to be the one of the major platforms to optimize Bing Ads Marketplace due to its reliability under frequent policy changes and its efficiency to minimize risks in real experiments.

📄 PDF Abstract BibTeX arXiv:1808.07251

Code (0)

등록된 구현이 없습니다.

Tasks

counterfactual

Similar Papers 제목 키워드 기반

Unifying Online and Counterfactual Learning to Rank

2020-12-08 · Harrie Oosterhuis, Maarten de Rijke

Optimizing ranking systems based on user interactions is a well-studied problem. State-of-the-art methods for optimizing ranking systems based on user interactions are divided into online approaches - that learn by direc…

counterfactualLearning-To-RankSelection bias

CAB: Continuous Adaptive Blending Estimator for Policy Evaluation and Learning

2018-11-06 · Yi Su, Lequn Wang, Michele Santacatterina, Thorsten Joachims

The ability to perform offline A/B-testing and off-policy learning using logged contextual bandit feedback is highly desirable in a broad range of applications, including recommender systems, search engines, ad placement…

counterfactualRecommendation Systems

The Self-Normalized Estimator for Counterfactual Learning

2015-12-01 · NeurIPS 2015 12 · Adith Swaminathan, Thorsten Joachims

This paper identifies a severe problem of the counterfactual risk estimator typically used in batch learning from logged bandit feedback (BLBF), and proposes the use of an alternative estimator that avoids this problem.I…

counterfactual

Cost-Effective Incentive Allocation via Structured Counterfactual Inference

2019-02-07 · Romain Lopez, Chenchen Li, Xiang Yan, Junwu Xiong 외

We address a practical problem ubiquitous in modern marketing campaigns, in which a central agent tries to learn a policy for allocating strategic financial incentives to customers and observes only bandit feedback. In c…

counterfactualCounterfactual InferenceDomain AdaptationMarketing

Estimation and Inference for Policy Relevant Treatment Effects

2020-07-16

The policy relevant treatment effect (PRTE) measures the average effect of switching from a status-quo policy to a counterfactual policy. Estimation of the PRTE involves estimation of multiple preliminary parameters, inc…

counterfactual