paper-with-me

Papers

Macro-Action-Based Multi-Agent/Robot Deep Reinforcement Learning under Partial Observability

2022-09-20 · Yuchen Xiao

The state-of-the-art multi-agent reinforcement learning (MARL) methods have provided promising solutions to a variety of complex problems. Yet, these methods all assume that agents perform synchronized primitive-action executions so that they are not genuinely scalable to long-horizon real-world multi-agent/robot tasks that inherently require agents/robots to asynchronously reason about high-level action selection at varying time durations. The Macro-Action Decentralized Partially Observable Markov Decision Process (MacDec-POMDP) is a general formalization for asynchronous decision-making under uncertainty in fully cooperative multi-agent tasks. In this thesis, we first propose a group of value-based RL approaches for MacDec-POMDPs, where agents are allowed to perform asynchronous learning and decision-making with macro-action-value functions in three paradigms: decentralized learning and control, centralized learning and control, and centralized training for decentralized execution (CTDE). Building on the above work, we formulate a set of macro-action-based policy gradient algorithms under the three training paradigms, where agents are allowed to directly optimize their parameterized policies in an asynchronous manner. We evaluate our methods both in simulation and on real robots over a variety of realistic domains. Empirical results demonstrate the superiority of our approaches in large multi-agent problems and validate the effectiveness of our algorithms for learning high-quality and asynchronous solutions with macro-actions.

📄 PDF Abstract BibTeX arXiv:2209.10003

Code (0)

등록된 구현이 없습니다.

Tasks

Decision MakingDecision Making Under UncertaintyDeep Reinforcement LearningMulti-agent Reinforcement Learningreinforcement-learningReinforcement Learning (RL)

Similar Papers 제목 키워드 기반

Macro-Action-Based Deep Multi-Agent Reinforcement Learning

2020-04-18 · Yuchen Xiao, Joshua Hoffman, Christopher Amato

In real-world multi-robot systems, performing high-quality, collaborative behaviors requires robots to asynchronously reason about high-level action selection at varying time durations. Macro-Action Decentralized Partial…

Decision MakingDecision Making Under UncertaintyDeep Reinforcement LearningMulti-agent Reinforcement Learning+3

Learning Multi-Robot Decentralized Macro-Action-Based Policies via a Centralized Q-Net

2019-09-19 · Yuchen Xiao, Joshua Hoffman, Tian Xia, Christopher Amato

In many real-world multi-robot tasks, high-quality solutions often require a team of robots to perform asynchronous actions under decentralized control. Decentralized multi-agent reinforcement learning methods have diffi…

Multi-agent Reinforcement LearningReinforcement Learning

Asynchronous Multi-Agent Actor-Critic with Macro-Actions

2021-09-29 · Yuchen Xiao, Weihao Tan, Christopher Amato

Many realistic multi-agent problems naturally require agents to be capable of performing asynchronously without waiting for other agents to terminate (e.g., multi-robot domains). Such problems can be modeled as Macro-Act…

Decision MakingPolicy Gradient Methods

Reusability and Transferability of Macro Actions for Reinforcement Learning

2019-08-05 · Yi-Hsiang Chang, Kuan-Yu Chang, Henry Kuo, Chun-Yi Lee

Conventional reinforcement learning (RL) typically determines an appropriate primitive action at each timestep. However, by using a proper macro action, defined as a sequence of primitive actions, an agent is able to byp…

reinforcement-learningReinforcement LearningReinforcement Learning (RL)

MaMiC: Macro and Micro Curriculum for Robotic Reinforcement Learning

2019-05-17 · Manan Tomar, Akhil Sathuluri, Balaraman Ravindran

Shaping in humans and animals has been shown to be a powerful tool for learning complex tasks as compared to learning in a randomized fashion. This makes the problem less complex and enables one to solve the easier sub t…

reinforcement-learningReinforcement LearningReinforcement Learning (RL)Robot Manipulation