paper-with-me

Papers

Uncertainty quantification in the Bradley-Terry-Luce model

2021-10-08 · Chao GAO, Yandi Shen, Anderson Y. Zhang

The Bradley-Terry-Luce (BTL) model is a benchmark model for pairwise comparisons between individuals. Despite recent progress on the first-order asymptotics of several popular procedures, the understanding of uncertainty quantification in the BTL model remains largely incomplete, especially when the underlying comparison graph is sparse. In this paper, we fill this gap by focusing on two estimators that have received much recent attention: the maximum likelihood estimator (MLE) and the spectral estimator. Using a unified proof strategy, we derive sharp and uniform non-asymptotic expansions for both estimators in the sparsest possible regime (up to some poly-logarithmic factors) of the underlying comparison graph. These expansions allow us to obtain: (i) finite-dimensional central limit theorems for both estimators; (ii) construction of confidence intervals for individual ranks; (iii) optimal constant of $\ell_2$ estimation, which is achieved by the MLE but not by the spectral estimator. Our proof is based on a self-consistent equation of the second-order remainder vector and a novel leave-two-out analysis.

📄 PDF Abstract BibTeX arXiv:2110.03874

Code (0)

등록된 구현이 없습니다.

Tasks

modelUncertainty Quantification

Similar Papers 제목 키워드 기반

Lagrangian Inference for Ranking Problems

2021-10-01 · Yue Liu, Ethan X. Fang, Junwei Lu

We propose a novel combinatorial inference framework to conduct general uncertainty quantification in ranking problems. We consider the widely adopted Bradley-Terry-Luce (BTL) model, where each item is assigned a positiv…

Uncertainty Quantification

A Judge-Aware Ranking Framework for Evaluating Large Language Models without Ground Truth

2026-01-29 · Mingyuan Xu, Xinzi Tan, Jiawei Wu, Doudou Zhou arxiv

Evaluating large language models (LLMs) on open-ended tasks without ground-truth labels is increasingly done via the LLM-as-a-judge paradigm. A critical but under-modeled issue is that judge LLMs differ substantially in …

LLM Evaluation as Tensor Completion: Low Rank Structure and Semiparametric Efficiency

2026-04-07 · Jiachun Li, David Simchi-Levi, Will Wei Sun arxiv

Large language model (LLM) evaluation platforms increasingly rely on pairwise human judgments. These data are noisy, sparse, and non-uniform, yet leaderboards are reported with limited uncertainty quantification. We stud…

Spectral Ranking Inferences based on General Multiway Comparisons

2023-08-05 · Jianqing Fan, Zhipeng Lou, Weichen Wang, Mengxin Yu

This paper studies the performance of the spectral method in the estimation and uncertainty quantification of the unobserved preference scores of compared entities in a general and more realistic setup. Specifically, the…

Uncertainty Quantification

An Analysis of Elo Rating Systems via Markov Chains

2024-06-09 · Sam Olesker-Taylor, Luca Zanetti

We present a theoretical analysis of the Elo rating system, a popular method for ranking skills of players in an online setting. In particular, we study Elo under the Bradley--Terry--Luce model and, using techniques from…