paper-with-me

홈 › Papers

When Less Is More: Binary Feedback Can Outperform Ordinal Comparisons in Ranking Recovery

2025-07-02 · Shirong Xu, Jingnan Zhang, Junhui Wang arxiv

Paired comparison data, where users evaluate items in pairs, play a central role in ranking and preference learning tasks. While ordinal comparison data intuitively offer richer information than binary comparisons, this paper challenges that conventional wisdom. We propose a general parametric framework for modeling ordinal paired comparisons without ties. The model adopts a generalized additive structure, featuring a link function that quantifies the preference difference between two items and a pattern function that governs the distribution over ordinal response levels. This framework encompasses classical binary comparison models as special cases, by treating binary responses as binarized versions of ordinal data. Within this framework, we show that binarizing ordinal data can significantly improve the accuracy of ranking recovery. Specifically, we prove that under the counting algorithm, the ranking error associated with binary comparisons exhibits a faster exponential convergence rate than that of ordinal data. Furthermore, we characterize a substantial performance gap between binary and ordinal data in terms of a signal-to-noise ratio (SNR) determined by the pattern function. We identify the pattern function that minimizes the SNR and maximizes the benefit of binarization. Extensive simulations and a real application on the MovieLens dataset further corroborate our theoretical findings.

📄 PDF Abstract BibTeX arXiv:2507.01613

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Restless bandits with imperfect binary feedback: PCL-indexability analysis and computation

2026-03-27 · José Niño-Mora arxiv

We study restless bandits with binary latent states and imperfect binary feedback, motivated by opportunistic spectrum access with sensing errors. For the associated belief-state model, we develop a partial conservation …

Some things are more CRINGE than others: Iterative Preference Optimization with the Pairwise Cringe Loss

2023-12-27 · Jing Xu, Andrew Lee, Sainbayar Sukhbaatar, Jason Weston

Practitioners commonly align large language models using pairwise preferences, i.e., given labels of the type response A is preferred to response B for a given input. Perhaps less commonly, methods have also been develop…

LeTI: Learning to Generate from Textual Interactions

2023-05-17 · Xingyao Wang, Hao Peng, Reyhaneh Jabbarvand, Heng Ji

Fine-tuning pre-trained language models (LMs) is essential for enhancing their capabilities. Existing techniques commonly fine-tune on input-output pairs (e.g., instruction tuning) or with numerical rewards that gauge th…

Code GenerationEvent Argument ExtractionHumanEvalmbpp

WhittleSearch: Interactive Image Search with Relative Attribute Feedback

2015-05-15 · Adriana Kovashka, Devi Parikh, Kristen Grauman

We propose a novel mode of feedback for image search, where a user describes which properties of exemplar images should be adjusted in order to more closely match his/her mental model of the image sought. For example, pe…

AttributeImage Retrieval

Explanation Augmented Feedback in Human-in-the-Loop Reinforcement Learning

2020-10-15 · NeurIPS Workshop HAMLETS 2020 12 · Anonymous

Human-in-the-loop Reinforcement Learning (HRL) aims to integrate human guidance with Reinforcement Learning (RL) algorithms to improve sample efficiency and performance. A common type of human guidance in HRL is binary e…

Atari Gamesreinforcement-learningReinforcement LearningReinforcement Learning (RL)