paper-with-me

홈 › Papers

Clustered Multi-Agent Linear Bandits

2023-09-15 · Hamza Cherkaoui, Merwan Barlier, Igor Colin

We address in this paper a particular instance of the multi-agent linear stochastic bandit problem, called clustered multi-agent linear bandits. In this setting, we propose a novel algorithm leveraging an efficient collaboration between the agents in order to accelerate the overall optimization problem. In this contribution, a network controller is responsible for estimating the underlying cluster structure of the network and optimizing the experiences sharing among agents within the same groups. We provide a theoretical analysis for both the regret minimization problem and the clustering quality. Through empirical evaluation against state-of-the-art algorithms on both synthetic and real data, we demonstrate the effectiveness of our approach: our algorithm significantly improves regret minimization while managing to recover the true underlying cluster partitioning.

📄 PDF Abstract BibTeX arXiv:2309.08710

Code (0)

등록된 구현이 없습니다.

Tasks

Clustering

Similar Papers 제목 키워드 기반

Thompson Sampling for Bandits with Clustered Arms

2021-09-06 · Emil Carlsson, Devdatt Dubhashi, Fredrik D. Johansson

We propose algorithms based on a multi-level Thompson sampling scheme, for the stochastic multi-armed bandit and its contextual variant with linear expected rewards, in the setting where arms are clustered. We show, both…

ClusteringThompson Sampling

Clustered Linear Contextual Bandits with Knapsacks

2023-08-21 · Yichuan Deng, Michalis Mamakos, Zhao Song

In this work, we study clustered contextual bandits where rewards and resource consumption are the outcomes of cluster-specific linear models. The arms are divided in clusters, with the cluster memberships being unknown …

EconometricsMulti-Armed Bandits

An Arm-Wise Randomization Approach to Combinatorial Linear Semi-Bandits

2019-09-05 · Kei Takemura, Shinji Ito

Combinatorial linear semi-bandits (CLS) are widely applicable frameworks of sequential decision-making, in which a learner chooses a subset of arms from a given set of arms associated with feature vectors. Existing algor…

Decision MakingRecommendation SystemsSequential Decision MakingThompson Sampling

Adaptive Clustering and Personalization in Multi-Agent Stochastic Linear Bandits

2021-06-15 · Avishek Ghosh, Abishek Sankararaman, Kannan Ramchandran

We consider the problem of minimizing regret in an $N$ agent heterogeneous stochastic linear bandits framework, where the agents (users) are similar but not all identical. We model user heterogeneity using two popularly …

Clustering

Near Optimal Best Arm Identification for Clustered Bandits

2025-05-15 · Yash, Nikhil Karamchandani, Avishek Ghosh

This work investigates the problem of best arm identification for multi-agent multi-armed bandits. We consider $N$ agents grouped into $M$ clusters, where each cluster solves a stochastic bandit problem. The mapping betw…

ClusteringComputational EfficiencyMulti-Armed Bandits