paper-with-me

Papers

Federated Online Sparse Decision Making

2022-02-27 · Chi-Hua Wang, Wenjie Li, Guang Cheng, Guang Lin

This paper presents a novel federated linear contextual bandits model, where individual clients face different K-armed stochastic bandits with high-dimensional decision context and coupled through common global parameters. By leveraging the sparsity structure of the linear reward , a collaborative algorithm named \texttt{Fedego Lasso} is proposed to cope with the heterogeneity across clients without exchanging local decision context vectors or raw reward data. \texttt{Fedego Lasso} relies on a novel multi-client teamwork-selfish bandit policy design, and achieves near-optimal regrets for shared parameter cases with logarithmic communication costs. In addition, a new conceptual tool called federated-egocentric policies is introduced to delineate exploration-exploitation trade-off. Experiments demonstrate the effectiveness of the proposed algorithms on both synthetic and real-world datasets.

📄 PDF Abstract BibTeX arXiv:2202.13448

Code (0)

등록된 구현이 없습니다.

Tasks

Decision MakingMulti-Armed Bandits

Similar Papers 제목 키워드 기반

Sparse Uncertainty-Informed Sampling from Federated Streaming Data

2024-08-30 · Manuel Röder, Frank-Michael Schleif

We present a numerically robust, computationally efficient approach for non-I.I.D. data stream sampling in federated client systems, where resources are limited and labeled data for local model adaptation is sparse and e…

Diversity

Pursuing Overall Welfare in Federated Learning through Sequential Decision Making

2024-05-31 · Seok-Ju Hahn, Gi-Soo Kim, Junghye Lee

In traditional federated learning, a single global model cannot perform equally well for all clients. Therefore, the need to achieve the client-level fairness in federated system has been emphasized, which can be realize…

Decision MakingFairnessFederated LearningSequential Decision Making

Batched Online Contextual Sparse Bandits with Sequential Inclusion of Features

2024-09-13 · Rowan Swiers, Subash Prabanantham, Andrew Maher

Multi-armed Bandits (MABs) are increasingly employed in online platforms and e-commerce to optimize decision making for personalized user experiences. In this work, we focus on the Contextual Bandit problem with linear r…

Decision MakingFairnessMulti-Armed Bandits

Regret Minimization and Statistical Inference in Online Decision Making with High-dimensional Covariates

2024-11-10 · Congyuan Duan, Wanteng Ma, Jiashuo Jiang, Dong Xia

This paper investigates regret minimization, statistical inference, and their interplay in high-dimensional online decision-making based on the sparse linear context bandit model. We integrate the $\varepsilon$-greedy ba…

Decision Makingvalid

Tracking as Online Decision-Making: Learning a Policy from Streaming Videos with Reinforcement Learning

2017-07-17 · ICCV 2017 10 · James Steven Supancic III, Deva Ramanan

We formulate tracking as an online decision-making process, where a tracking agent must follow an object despite ambiguous image frames and a limited computational budget. Crucially, the agent must decide where to look i…

Decision MakingDeep Reinforcement LearningReinforcement LearningReinforcement Learning (RL)