paper-with-me

홈 › Papers

Foresight of Graph Reinforcement Learning Latent Permutations Learnt by Gumbel Sinkhorn Network

2021-10-23 · Tianqi Shen, Hong Zhang, Ding Yuan, Jiaping Xiao, Yifan Yang

Vital importance has necessity to be attached to cooperation in multi-agent environments, as a result of which some reinforcement learning algorithms combined with graph neural networks have been proposed to understand the mutual interplay between agents. However, highly complicated and dynamic multi-agent environments require more ingenious graph neural networks, which can comprehensively represent not only the graph topology structure but also evolution process of the structure due to agents emerging, disappearing and moving. To tackle these difficulties, we propose Gumbel Sinkhorn graph attention reinforcement learning, where a graph attention network highly represents the underlying graph topology structure of the multi-agent environment, and can adapt to the dynamic topology structure of graph better with the help of Gumbel Sinkhorn network by learning latent permutations. Empirically, simulation results show how our proposed graph reinforcement learning methodology outperforms existing methods in the PettingZoo multi-agent environment by learning latent permutations.

📄 PDF Abstract BibTeX arXiv:2110.12144

Code (0)

등록된 구현이 없습니다.

Tasks

Graph Attentionreinforcement-learningReinforcement LearningReinforcement Learning (RL)

Similar Papers 제목 키워드 기반

Decoding Polar Codes with Reinforcement Learning

2020-09-15 · Nghia Doan, Seyyed Ali Hashemi, Warren Gross

In this paper we address the problem of selecting factor-graph permutations of polar codes under belief propagation (BP) decoding to significantly improve the error-correction performance of the code. In particular, we f…

Decoderreinforcement-learningReinforcement LearningReinforcement Learning (RL)

Learning for Visual Navigation by Imagining the Success

2021-02-28 · Mahdi Kazemi Moghaddam, Ehsan Abbasnejad, Qi Wu, Javen Shi 외

Visual navigation is often cast as a reinforcement learning (RL) problem. Current methods typically result in a suboptimal policy that learns general obstacle avoidance and search behaviours. For example, in the target-o…

NavigateReinforcement Learning (RL)Visual Navigation

Mingling Foresight with Imagination: Model-Based Cooperative Multi-Agent Reinforcement Learning

2022-04-20 · Zhiwei Xu, Dapeng Li, Bin Zhang, Yuan Zhan 외

Recently, model-based agents have achieved better performance than model-free ones using the same computational budget and training time in single-agent environments. However, due to the complexity of multi-agent systems…

Multi-agent Reinforcement Learningreinforcement-learningReinforcement Learning (RL)

TacForeSight: Force-Guided Tactile World Model for Contact-Rich Manipulation

2026-06-09 · Yujie Zang, Yuhang Zheng, Xian Nie, Yupeng Zheng 외 arxiv

Contact-rich manipulation requires robots to continuously perceive and regulate evolving physical interactions under dynamic contact transitions or complex surface geometries. Recent imitation learning methods improve co…

CLaD: Planning with Grounded Foresight via Cross-Modal Latent Dynamics

2026-03-31 · Andrew Jeong, Jaemin Kim, Sebin Lee, Sung-Eui Yoon arxiv

Robotic manipulation involves kinematic and semantic transitions that are inherently coupled via underlying actions. However, existing approaches plan within either semantic or latent space without explicitly aligning th…