paper-with-me

홈 › Papers

Exploiting Correlation to Achieve Faster Learning Rates in Low-Rank Preference Bandits

2022-02-23 · Suprovat Ghoshal, Aadirupa Saha

We introduce the \emph{Correlated Preference Bandits} problem with random utility-based choice models (RUMs), where the goal is to identify the best item from a given pool of $n$ items through online subsetwise preference feedback. We investigate whether models with a simple correlation structure, e.g. low rank, can result in faster learning rates. While we show that the problem can be impossible to solve for the general low rank' choice models, faster learning rates can be attained assuming more structured item correlations. In particular, we introduce a new class of \emph{Block-Rank} based RUM model, where the best item is shown to be $(\epsilon,\delta)$-PAC learnable with only $O(r \epsilon^{-2} \log(n/\delta))$ samples. This improves on the standard sample complexity bound of $\tilde{O}(n\epsilon^{-2} \log(1/\delta))$ known for the usual learning algorithms which might not exploit the item-correlations ($r \ll n$). We complement the above sample complexity with a matching lower bound (up to logarithmic factors), justifying the tightness of our analysis. Surprisingly, we also show a lower bound of $\Omega(n\epsilon^{-2}\log(1/\delta))$ when the learner is forced to play just duels instead of larger subsetwise queries. Further, we extend the results to a more general \emph{noisy Block-Rank}' model, which ensures robustness of our techniques. Overall, our results justify the advantage of playing subsetwise queries over pairwise preferences $(k=2)$, we show the latter provably fails to exploit correlation.

📄 PDF Abstract BibTeX arXiv:2202.11795

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Exploiting Multi-Label Correlation in Label Distribution Learning

2023-08-03 · Zhiqiang Kou jing wang yuheng jia xin geng

Label Distribution Learning (LDL) is a novel machine learning paradigm that assigns label distribution to each instance. Many LDL methods proposed to leverage label correlation in the learning process to solve the expone…

Multi-Label Learning

Multi-label Classification with High-rank and High-order Label Correlations

2022-07-09 · Chongjie Si, Yuheng Jia, Ran Wang, Min-Ling Zhang 외

Exploiting label correlations is important to multi-label classification. Previous methods capture the high-order label correlations mainly by transforming the label matrix to a latent label space with low-rank matrix fa…

Common Sense ReasoningMulti-Label ClassificationMUlTI-LABEL-ClASSIFICATIONVocal Bursts Intensity Prediction

Dynamic Low-rank Approximation of Full-Matrix Preconditioner for Training Generalized Linear Models

2025-08-28 · Tatyana Matveeva, Aleksandr Katrutsa, Evgeny Frolov arxiv

Adaptive gradient methods like Adagrad and its variants are widespread in large-scale optimization. However, their use of diagonal preconditioning matrices limits the ability to capture parameter correlations. Full-matri…

Low-Rank Key Value Attention

2026-01-16 · James O'Neill, Robert Clancy, Mariia Matskevichus, Fergal Reid arxiv

The key-value (KV) cache is a primary memory bottleneck in Transformers. We propose Low-Rank Key-Value (LRKV) attention, which reduces KV cache memory by exploiting redundancy across attention heads, while being compute …

AirShot: Efficient Few-Shot Detection for Autonomous Exploration

2024-04-07 · Zihan Wang, Bowen Li, Chen Wang, Sebastian Scherer

Few-shot object detection has drawn increasing attention in the field of robotic exploration, where robots are required to find unseen objects with a few online provided examples. Despite recent efforts have been made to…

Few-Shot Object Detectionobject-detectionObject Detection