paper-with-me

홈 › Papers

${\rm E}(3)$-Equivariant Actor-Critic Methods for Cooperative Multi-Agent Reinforcement Learning

2023-08-23 · Dingyang Chen, Qi Zhang

Identification and analysis of symmetrical patterns in the natural world have led to significant discoveries across various scientific fields, such as the formulation of gravitational laws in physics and advancements in the study of chemical structures. In this paper, we focus on exploiting Euclidean symmetries inherent in certain cooperative multi-agent reinforcement learning (MARL) problems and prevalent in many applications. We begin by formally characterizing a subclass of Markov games with a general notion of symmetries that admits the existence of symmetric optimal values and policies. Motivated by these properties, we design neural network architectures with symmetric constraints embedded as an inductive bias for multi-agent actor-critic methods. This inductive bias results in superior performance in various cooperative MARL benchmarks and impressive generalization capabilities such as zero-shot learning and transfer learning in unseen scenarios with repeated symmetric patterns. The code is available at: https://github.com/dchen48/E3AC.

📄 PDF Abstract BibTeX arXiv:2308.11842

Code (1)

dchen48/e3ac 공식 구현 pytorch

Tasks

Inductive BiasMulti-agent Reinforcement LearningTransfer LearningZero-Shot Learning

Methods 이 논문이 사용한 방법론

Focus 설명 없음

Similar Papers 제목 키워드 기반

Centralized Permutation Equivariant Policy for Cooperative Multi-Agent Reinforcement Learning

2025-08-13 · Zhuofan Xu, Benedikt Bollig, Matthias Függer, Thomas Nowak 외 arxiv

The Centralized Training with Decentralized Execution (CTDE) paradigm has gained significant attention in multi-agent reinforcement learning (MARL) and is the foundation of many recent algorithms. However, decentralized …

Multi-agent Reinforcement Learning

Multi-Agent MDP Homomorphic Networks

2021-10-09 · ICLR 2022 4 · Elise van der Pol, Herke van Hoof, Frans A. Oliehoek, Max Welling

This paper introduces Multi-Agent MDP Homomorphic Networks, a class of networks that allows distributed execution using only local information, yet is able to share experience between global symmetries in the joint state…

Strategic learning for disturbance rejection in multi-agent systems: Nash and Minmax in graphical games

2025-04-10 · Xinyang Wang, Martin Guay, Shimin Wang, Hongwei Zhang

This article investigates the optimal control problem with disturbance rejection for discrete-time multi-agent systems under cooperative and non-cooperative graphical games frameworks. Given the practical challenges of o…

F2A2: Flexible Fully-decentralized Approximate Actor-critic for Cooperative Multi-agent Reinforcement Learning

2020-04-17 · Wenhao Li, Bo Jin, Xiangfeng Wang, Junchi Yan 외

Traditional centralized multi-agent reinforcement learning (MARL) algorithms are sometimes unpractical in complicated applications, due to non-interactivity between agents, curse of dimensionality and computation complex…

Multi-agent Reinforcement LearningReinforcement LearningReinforcement Learning (RL)Starcraft+1

Parameter Sharing Deep Deterministic Policy Gradient for Cooperative Multi-agent Reinforcement Learning

2017-10-01 · Xiangxiang Chu, Hangjun Ye

Deep reinforcement learning for multi-agent cooperation and competition has been a hot topic recently. This paper focuses on cooperative multi-agent problem based on actor-critic methods under local observations settings…

Deep Reinforcement LearningMulti-agent Reinforcement Learningreinforcement-learningReinforcement Learning+1