paper-with-me

Papers

Structured Diversification Emergence via Reinforced Organization Control and Hierarchical Consensus Learning

2021-02-09 · Wenhao Li, Xiangfeng Wang, Bo Jin, Junjie Sheng, Yun Hua, Hongyuan Zha

When solving a complex task, humans will spontaneously form teams and to complete different parts of the whole task, respectively. Meanwhile, the cooperation between teammates will improve efficiency. However, for current cooperative MARL methods, the cooperation team is constructed through either heuristics or end-to-end blackbox optimization. In order to improve the efficiency of cooperation and exploration, we propose a structured diversification emergence MARL framework named {\sc{Rochico}} based on reinforced organization control and hierarchical consensus learning. {\sc{Rochico}} first learns an adaptive grouping policy through the organization control module, which is established by independent multi-agent reinforcement learning. Further, the hierarchical consensus module based on the hierarchical intentions with consensus constraint is introduced after team formation. Simultaneously, utilizing the hierarchical consensus module and a self-supervised intrinsic reward enhanced decision module, the proposed cooperative MARL algorithm {\sc{Rochico}} can output the final diversified multi-agent cooperative policy. All three modules are organically combined to promote the structured diversification emergence. Comparative experiments on four large-scale cooperation tasks show that {\sc{Rochico}} is significantly better than the current SOTA algorithms in terms of exploration efficiency and cooperation strength.

📄 PDF Abstract BibTeX arXiv:2102.04775

Code (0)

등록된 구현이 없습니다.

Tasks

Multi-agent Reinforcement Learning

Similar Papers 제목 키워드 기반

Long-run patterns in the discovery of the adjacent possible

2022-08-01 · Josef Taalbi

The notion of the "adjacent possible" has been advanced to theorize the generation of novelty across many different research domains. This study is an attempt to examine in what way the notion can be made empirically use…

Filtering Patent Maps for Visualization of Diversification Paths of Inventors and Organizations

2016-06-25 · Yan Bowen, Luo Jianxi

In the information science literature, recent studies have used patent databases and patent classification information to construct network maps of patent technology classes. In such a patent technology map, almost all p…

Patent classification

Will enterprise digital transformation affect diversification strategy?

2021-12-13 · Ge-zhi Wu, Da-ming You

This paper empirically examines the impact of enterprise digital transformation on the level of enterprise diversification. It is found that the digital transformation of enterprises has significantly improved the level …

Reinforced Data Sampling for Model Diversification

2020-06-12 · Hoang D. Nguyen, Xuan-Son Vu, Quoc-Tuan Truong, Duc-Trong Le

With the rising number of machine learning competitions, the world has witnessed an exciting race for the best algorithms. However, the involved data selection process may fundamentally suffer from evidence ambiguity and…

BIG-bench Machine Learningmodel

Hierarchical Reinforced Trader (HRT): A Bi-Level Approach for Optimizing Stock Selection and Execution

2024-10-19 · Zijie Zhao, Roy E. Welsch

Leveraging Deep Reinforcement Learning (DRL) in automated stock trading has shown promising results, yet its application faces significant challenges, including the curse of dimensionality, inertia in trading actions, an…

Deep Reinforcement LearningHierarchical Reinforcement Learningreinforcement-learningReinforcement Learning