paper-with-me

홈 › Papers

A Dynamic Observation Strategy for Multi-agent Multi-armed Bandit Problem

2020-04-08 · Udari Madhushani, Naomi Ehrich Leonard

We define and analyze a multi-agent multi-armed bandit problem in which decision-making agents can observe the choices and rewards of their neighbors under a linear observation cost. Neighbors are defined by a network graph that encodes the inherent observation constraints of the system. We define a cost associated with observations such that at every instance an agent makes an observation it receives a constant observation regret. We design a sampling algorithm and an observation protocol for each agent to maximize its own expected cumulative reward through minimizing expected cumulative sampling regret and expected cumulative observation regret. For our proposed protocol, we prove that total cumulative regret is logarithmically bounded. We verify the accuracy of analytical bounds using numerical simulations.

📄 PDF Abstract BibTeX arXiv:2004.03793

Code (0)

등록된 구현이 없습니다.

Tasks

Decision Making

Similar Papers 제목 키워드 기반

Uncoupled Learning of Differential Stackelberg Equilibria with Commitments

2023-02-07 · Robert Loftin, Mustafa Mert Çelikok, Herke van Hoof, Samuel Kaski 외

In multi-agent problems requiring a high degree of cooperation, success often depends on the ability of the agents to adapt to each other's behavior. A natural solution concept in such settings is the Stackelberg equilib…

Deep Reinforcement LearningMulti-agent Reinforcement Learning

Centralized control for multi-agent RL in a complex Real-Time-Strategy game

2023-04-25 · Roger Creus Castanyer

Multi-agent Reinforcement learning (MARL) studies the behaviour of multiple learning agents that coexist in a shared environment. MARL is more challenging than single-agent RL because it involves more complex learning dy…

Multi-agent Reinforcement Learning

Belief sharing: a blessing or a curse

2024-07-02 · Ozan Catal, Toon Van de Maele, Riddhi J. Pitliya, Mahault Albarracin 외

When collaborating with multiple parties, communicating relevant information is of utmost importance to efficiently completing the tasks at hand. Under active inference, communication can be cast as sharing beliefs betwe…

MagNet: Discovering Multi-agent Interaction Dynamics using Neural Network

2020-01-24 · Priyabrata Saha, Arslan Ali, Burhan A. Mudassar, Yun Long 외

We present the MagNet, a neural network-based multi-agent interaction model to discover the governing dynamics and predict evolution of a complex multi-agent system from observations. We formulate a multi-agent system as…

Audio-Visual World Models: Learning Physically Grounded Multisensory Dynamics

2025-11-30 · Jiahua Wang, Leqi Zheng, Jialong Wu, Yaoxin Mao 외 arxiv

World models simulate environmental dynamics to enable embodied agents to plan and reason about future states. While real-world perception is inherently multimodal, existing approaches focus primarily on visual observati…

Sound Source LocalizationVisual Navigation