paper-with-me

Papers

Distributed Consensus Algorithm for Decision-Making in Multi-agent Multi-armed Bandit

2023-06-09 · Xiaotong Cheng, Setareh Maghsudi

We study a structured multi-agent multi-armed bandit (MAMAB) problem in a dynamic environment. A graph reflects the information-sharing structure among agents, and the arms' reward distributions are piecewise-stationary with several unknown change points. The agents face the identical piecewise-stationary MAB problem. The goal is to develop a decision-making policy for the agents that minimizes the regret, which is the expected total loss of not playing the optimal arm at each time step. Our proposed solution, Restarted Bayesian Online Change Point Detection in Cooperative Upper Confidence Bound Algorithm (RBO-Coop-UCB), involves an efficient multi-agent UCB algorithm as its core enhanced with a Bayesian change point detector. We also develop a simple restart decision cooperation that improves decision-making. Theoretically, we establish that the expected group regret of RBO-Coop-UCB is upper bounded by $\mathcal{O}(KNM\log T + K\sqrt{MT\log T})$, where K is the number of agents, M is the number of arms, and T is the number of time steps. Numerical experiments on synthetic and real-world datasets demonstrate that our proposed method outperforms the state-of-the-art algorithms.

📄 PDF Abstract BibTeX arXiv:2306.05998

Code (0)

등록된 구현이 없습니다.

Tasks

Change Point DetectionDecision Making

Similar Papers 제목 키워드 기반

On Distributed Cooperative Decision-Making in Multiarmed Bandits

2015-12-21 · Peter Landgren, Vaibhav Srivastava, Naomi Ehrich Leonard

We study the explore-exploit tradeoff in distributed cooperative decision-making using the context of the multiarmed bandit (MAB) problem. For the distributed cooperative MAB problem, we design the cooperative UCB algori…

Decision Making

DANCeRS: A Distributed Algorithm for Negotiating Consensus in Robot Swarms with Gaussian Belief Propagation

2025-08-25 · Aalok Patwardhan, Andrew J. Davison arxiv

Robot swarms require cohesive collective behaviour to address diverse challenges, including shape formation and decision-making. Existing approaches often treat consensus in discrete and continuous decision spaces as dis…

Collision Avoidance

Distributed Cooperative Decision-Making in Multiarmed Bandits: Frequentist and Bayesian Algorithms

2016-06-02 · Peter Landgren, Vaibhav Srivastava, Naomi Ehrich Leonard

We study distributed cooperative decision-making under the explore-exploit tradeoff in the multiarmed bandit (MAB) problem. We extend the state-of-the-art frequentist and Bayesian algorithms for single-agent MAB problems…

Decision Making

Risk-Aware Distributed Multi-Agent Reinforcement Learning

2023-04-04 · Abdullah Al Maruf, Luyao Niu, Bhaskar Ramasubramanian, Andrew Clark 외

Autonomous cyber and cyber-physical systems need to perform decision-making, learning, and control in unknown environments. Such decision-making can be sensitive to multiple factors, including modeling errors, changes in…

Decision MakingMulti-agent Reinforcement Learningreinforcement-learningReinforcement Learning

A privacy-preserving distributed credible evidence fusion algorithm for collective decision-making

2024-12-03 · Chaoxiong Ma, Yan Liang, Xinyu Yang, Han Wu 외

The theory of evidence reasoning has been applied to collective decision-making in recent years. However, existing distributed evidence fusion methods lead to participants' preference leakage and fusion failures as they …

Decision MakingLow-Rank Matrix CompletionMatrix CompletionPrivacy Preserving