paper-with-me

홈 › Papers

Offline Comparison of Ranking Functions using Randomized Data

2018-10-11 · Agarwal Aman, Wang Xuanhui, Li Cheng, Bendersky Michael, Najork Marc

Ranking functions return ranked lists of items, and users often interact with these items. How to evaluate ranking functions using historical interaction logs, also known as off-policy evaluation, is an important but challenging problem. The commonly used Inverse Propensity Scores (IPS) approaches work better for the single item case, but suffer from extremely low data efficiency for the ranked list case. In this paper, we study how to improve the data efficiency of IPS approaches in the offline comparison setting. We propose two approaches Trunc-match and Rand-interleaving for offline comparison using uniformly randomized data. We show that these methods can improve the data efficiency and also the comparison sensitivity based on one of the largest email search engines.

📄 PDF Abstract BibTeX arXiv:1810.05252

Code (0)

등록된 구현이 없습니다.

Tasks

Off-policy evaluation

Similar Papers 제목 키워드 기반

Diverse Randomized Value Functions: A Provably Pessimistic Approach for Offline Reinforcement Learning

2024-04-09 · Xudong Yu, Chenjia Bai, Hongyi Guo, Changhong Wang 외

Offline Reinforcement Learning (RL) faces distributional shift and unreliable value estimation, especially for out-of-distribution (OOD) actions. To address this, existing uncertainty-based methods penalize the value fun…

DiversityReinforcement Learning (RL)Uncertainty Quantification

The Partial Testimony of Logs: Evaluation of Language Model Generation under Confounded Model Choice

2026-05-02 · Jikai Jin, Vasilis Syrgkanis arxiv

Offline evaluation of language models from usage logs is biased when model choice is confounded: the same user-side factors that influence which model is used can also influence how its output is judged, so raw compariso…

Randomized Value Functions via Multiplicative Normalizing Flows

2018-06-06 · Ahmed Touati, Harsh Satija, Joshua Romoff, Joelle Pineau 외

Randomized value functions offer a promising approach towards the challenge of efficient exploration in complex environments with high dimensional state and action spaces. Unlike traditional point estimate methods, rando…

Efficient ExplorationThompson Sampling

As you like it: Localization via paired comparisons

2018-02-19 · Andrew K. Massimino, Mark A. Davenport

Suppose that we wish to estimate a vector $\mathbf{x}$ from a set of binary paired comparisons of the form "$\mathbf{x}$ is closer to $\mathbf{p}$ than to $\mathbf{q}$" for various choices of vectors $\mathbf{p}$ and $\m…

Rate-Optimal Rank Aggregation with Private Pairwise Rankings

2024-02-26 · SHIRONG XU, Will Wei Sun, Guang Cheng

In various real-world scenarios, such as recommender systems and political surveys, pairwise rankings are commonly collected and utilized for rank aggregation to derive an overall ranking of items. However, preference ra…

Recommendation Systems