paper-with-me

홈 › Papers

ModGNN: Expert Policy Approximation in Multi-Agent Systems with a Modular Graph Neural Network Architecture

2021-03-24 · Ryan Kortvelesy, Amanda Prorok

Recent work in the multi-agent domain has shown the promise of Graph Neural Networks (GNNs) to learn complex coordination strategies. However, most current approaches use minor variants of a Graph Convolutional Network (GCN), which applies a convolution to the communication graph formed by the multi-agent system. In this paper, we investigate whether the performance and generalization of GCNs can be improved upon. We introduce ModGNN, a decentralized framework which serves as a generalization of GCNs, providing more flexibility. To test our hypothesis, we evaluate an implementation of ModGNN against several baselines in the multi-agent flocking problem. We perform an ablation analysis to show that the most important component of our framework is one that does not exist in a GCN. By varying the number of agents, we also demonstrate that an application-agnostic implementation of ModGNN possesses an improved ability to generalize to new environments.

📄 PDF Abstract BibTeX arXiv:2103.13446

Code (1)

proroklab/ModGNN 공식 구현 pytorch

Tasks

Graph Neural Network

Methods 이 논문이 사용한 방법론

GCN A Graph Convolutional Network, or GCN, is an approach for semi-supervised learning on graph-structured data. It is based on an efficient variant of [convolutional neural…
Convolution A convolution is a type of matrix operation, consisting of a kernel, a small matrix of weights, that slides over input data performing element-wise multiplication with the…

Similar Papers 제목 키워드 기반

Provably Efficient Generative Adversarial Imitation Learning for Online and Offline Setting with Linear Function Approximation

2021-08-19 · Zhihan Liu, Yufeng Zhang, Zuyue Fu, Zhuoran Yang 외

In generative adversarial imitation learning (GAIL), the agent aims to learn a policy from an expert demonstration so that its performance cannot be discriminated from the expert policy on a certain predefined reward set…

Imitation Learning

Off-Policy Adversarial Inverse Reinforcement Learning

2020-05-03 · ICML Workshop LifelongML 2020 7 · Samin Yeasar Arnob

Adversarial Imitation Learning (AIL) is a class of algorithms in Reinforcement learning (RL), which tries to imitate an expert without taking any reward from the environment and does not provide expert behavior directly …

continuous-controlContinuous ControlImitation Learningreinforcement-learning+3

Imitation Learning for Multi-turn LM Agents via On-policy Expert Corrections

2025-12-16 · Niklas Lauffer, Xiang Deng, Srivatsa Kundurthy, Brad Kenstler 외 arxiv

A popular paradigm for training LM agents relies on imitation learning, fine-tuning on expert trajectories. However, we show that the off-policy nature of imitation learning for multi-turn LM agents suffers from the fund…

Agentic-DPO: From Imitation to Agentic Policy Optimization on Expert Trajectories

2026-07-12 · Yixiong Chen, Alan Yuille arxiv

Large Language Model (LLM) agents are commonly trained from expert trajectories using supervised fine-tuning (SFT), which treats multi-turn agent behavior as ordinary text imitation. This recipe is simple and low-cost, b…

Reinforcement Learning

A Multi-Agent Off-Policy Actor-Critic Algorithm for Distributed Reinforcement Learning

2019-03-15 · Wesley Suttle, Zhuoran Yang, Kaiqing Zhang, Zhaoran Wang 외

This paper extends off-policy reinforcement learning to the multi-agent case in which a set of networked agents communicating with their neighbors according to a time-varying graph collaboratively evaluates and improves …

reinforcement-learningReinforcement LearningReinforcement Learning (RL)