paper-with-me

Papers

A Simple Unified Framework for High Dimensional Bandit Problems

2021-02-18 · Wenjie Li, Adarsh Barik, Jean Honorio

Stochastic high dimensional bandit problems with low dimensional structures are useful in different applications such as online advertising and drug discovery. In this work, we propose a simple unified algorithm for such problems and present a general analysis framework for the regret upper bound of our algorithm. We show that under some mild unified assumptions, our algorithm can be applied to different high dimensional bandit problems. Our framework utilizes the low dimensional structure to guide the parameter estimation in the problem, therefore our algorithm achieves the comparable regret bounds in the LASSO bandit, as well as novel bounds in the low-rank matrix bandit, the group sparse matrix bandit, and in a new problem: the multi-agent LASSO bandit.

📄 PDF Abstract BibTeX arXiv:2102.09626

Code (0)

등록된 구현이 없습니다.

Tasks

Drug Discoveryparameter estimationVocal Bursts Intensity Prediction

Similar Papers 제목 키워드 기반

A Unified Regularization Approach to High-Dimensional Generalized Tensor Bandits

2025-01-18 · Jiannan Li, Yiyang Yang, Yao Wang, Shaojie Tang

Modern decision-making scenarios often involve data that is both high-dimensional and rich in higher-order contextual information, where existing bandits algorithms fail to generate effective policies. In response, we pr…

Decision Making

Why Most Optimism Bandit Algorithms Have the Same Regret Analysis: A Simple Unifying Theorem

2025-12-20 · Vikram Krishnamurthy arxiv

Several optimism-based stochastic bandit algorithms -- including UCB, UCB-V, linear UCB, and finite-arm GP-UCB -- achieve logarithmic regret using proofs that, despite superficial differences, follow essentially the same…

Towards Scalable and Robust Structured Bandits: A Meta-Learning Framework

2022-02-26 · Runzhe Wan, Lin Ge, Rui Song

Online learning in large-scale structured bandits is known to be challenging due to the curse of dimensionality. In this paper, we propose a unified meta-learning framework for a general class of structured bandit proble…

Meta-LearningThompson Sampling

Dynamic Batch Learning in High-Dimensional Sparse Linear Contextual Bandits

2020-08-27 · Zhimei Ren, Zhengyuan Zhou

We study the problem of dynamic batch learning in high-dimensional sparse linear contextual bandits, where a decision maker, under a given maximum-number-of-batch constraint and only able to observe rewards at the end of…

Decision MakingMarketingMulti-Armed BanditsVocal Bursts Intensity Prediction

Regret Minimization and Statistical Inference in Online Decision Making with High-dimensional Covariates

2024-11-10 · Congyuan Duan, Wanteng Ma, Jiashuo Jiang, Dong Xia

This paper investigates regret minimization, statistical inference, and their interplay in high-dimensional online decision-making based on the sparse linear context bandit model. We integrate the $\varepsilon$-greedy ba…

Decision Makingvalid