paper-with-me

Papers

Revisiting the Master-Slave Architecture in Multi-Agent Deep Reinforcement Learning

2017-12-20 · ICLR 2018 1 · Xiangyu Kong, Bo Xin, Fangchen Liu, Yizhou Wang

Many tasks in artificial intelligence require the collaboration of multiple agents. We exam deep reinforcement learning for multi-agent domains. Recent research efforts often take the form of two seemingly conflicting perspectives, the decentralized perspective, where each agent is supposed to have its own controller; and the centralized perspective, where one assumes there is a larger model controlling all agents. In this regard, we revisit the idea of the master-slave architecture by incorporating both perspectives within one framework. Such a hierarchical structure naturally leverages advantages from one another. The idea of combining both perspectives is intuitive and can be well motivated from many real world systems, however, out of a variety of possible realizations, we highlights three key ingredients, i.e. composed action representation, learnable communication and independent reasoning. With network designs to facilitate these explicitly, our proposal consistently outperforms latest competing methods both in synthetic experiments and when applied to challenging StarCraft micromanagement tasks.

📄 PDF Abstract BibTeX arXiv:1712.07305

Code (0)

등록된 구현이 없습니다.

Tasks

Deep Reinforcement Learningreinforcement-learningReinforcement LearningReinforcement Learning (RL)Starcraft

Similar Papers 제목 키워드 기반

Crossover-BPSO Driven Multi-Agent Technology for Managing Local Energy Systems

2025-01-16 · Hafiz Majid Hussain, Ashfaq Ahmad. Pedro H. J. Nardelli

This article presents a new hybrid algorithm, crossover binary particle swarm optimization (crBPSO), for allocating resources in local energy systems via multi-agent (MA) technology. Initially, a hierarchical MA-based ar…

Scheduling

Master-slave Deep Architecture for Top-K Multi-armed Bandits with Non-linear Bandit Feedback and Diversity Constraints

2023-08-24 · Hanchi Huang, Li Shen, Deheng Ye, Wei Liu

We propose a novel master-slave architecture to solve the top-$K$ combinatorial multi-armed bandits problem with non-linear bandit feedback and diversity constraints, which, to the best of our knowledge, is the first com…

DiversityMulti-Armed Bandits

Transfer Learning for sEMG-based Hand Gesture Classification using Deep Learning in a Master-Slave Architecture

2020-04-27 · Karush Suri, Rinki Gupta

Recent advancements in diagnostic learning and development of gesture-based human machine interfaces have driven surface electromyography (sEMG) towards significant importance. Analysis of hand gestures requires an accur…

DiagnosticGeneral ClassificationTransfer Learning

MA-CoNav: A Master-Slave Multi-Agent Framework with Hierarchical Collaboration and Dual-Level Reflection for Long-Horizon Embodied VLN

2026-03-03 · Ling Luo, Qianqian Bai arxiv

Vision-Language Navigation (VLN) aims to empower robots with the ability to perform long-horizon navigation in unfamiliar environments based on complex linguistic instructions. Its success critically hinges on establishi…

Vision-Language Navigation

The Master-Slave Encoder Model for Improving Patent Text Summarization: A New Approach to Combining Specifications and Claims

2024-11-21 · Shu Zhou, Xin Wang, Zhengda Zhou, Haohan Yi 외

In order to solve the problem of insufficient generation quality caused by traditional patent text abstract generation models only originating from patent specifications, the problem of new terminology OOV caused by rapi…

Abstract generationText GenerationText Summarization