paper-with-me

홈 › Papers

LAMARL: LLM-Aided Multi-Agent Reinforcement Learning for Cooperative Policy Generation

2025-06-02 · Guobin Zhu, Rui Zhou, Wenkang Ji, Shiyu Zhao

Although Multi-Agent Reinforcement Learning (MARL) is effective for complex multi-robot tasks, it suffers from low sample efficiency and requires iterative manual reward tuning. Large Language Models (LLMs) have shown promise in single-robot settings, but their application in multi-robot systems remains largely unexplored. This paper introduces a novel LLM-Aided MARL (LAMARL) approach, which integrates MARL with LLMs, significantly enhancing sample efficiency without requiring manual design. LAMARL consists of two modules: the first module leverages LLMs to fully automate the generation of prior policy and reward functions. The second module is MARL, which uses the generated functions to guide robot policy training effectively. On a shape assembly benchmark, both simulation and real-world experiments demonstrate the unique advantages of LAMARL. Ablation studies show that the prior policy improves sample efficiency by an average of 185.9% and enhances task completion, while structured prompts based on Chain-of-Thought (CoT) and basic APIs improve LLM output success rates by 28.5%-67.5%. Videos and code are available at https://windylab.github.io/LAMARL/

📄 PDF Abstract BibTeX arXiv:2506.01538

Code (0)

등록된 구현이 없습니다.

Tasks

Multi-agent Reinforcement Learning

Similar Papers 제목 키워드 기반

LlamaRL: A Distributed Asynchronous Reinforcement Learning Framework for Efficient Large-scale LLM Trainin

2025-05-29 · Bo Wu, Sid Wang, Yunhao Tang, Jia Ding 외

Reinforcement Learning (RL) has become the most effective post-training approach for improving the capabilities of Large Language Models (LLMs). In practice, because of the high demands on latency and memory, it is parti…

GPUReinforcement Learning (RL)

Coordinated Anti-Jamming Resilience in Swarm Networks via Multi-Agent Reinforcement Learning

2025-12-18 · Bahman Abolhassani, Tugba Erpek, Kemal Davaslioglu, Yalin E. Sagduyu 외 arxiv

Reactive jammers pose a severe security threat to robotic-swarm networks by selectively disrupting inter-agent communications and undermining formation integrity and mission success. Conventional countermeasures such as …

Multi-agent Reinforcement Learning

Interruption-Aware Cooperative Perception for V2X Communication-Aided Autonomous Driving

2023-04-24 · Shunli Ren, Zixing Lei, Zi Wang, Mehrdad Dianati 외

Cooperative perception can significantly improve the perception performance of autonomous vehicles beyond the limited perception ability of individual vehicles by exchanging information with neighbor agents through V2X c…

Autonomous DrivingAutonomous VehiclesKnowledge Distillation

Multi-AUV Cooperative Target Tracking Based on Supervised Diffusion-Aided Multi-Agent Reinforcement Learning

2026-03-31 · Jiaao Ma, Chuan Lin, Guangjie Han, Shengchao Zhu 외 arxiv

In recent years, advances in underwater networking and multi-agent reinforcement learning (MARL) have significantly expanded multi-autonomous underwater vehicle (AUV) applications in marine exploration and target trackin…

Multi-agent Reinforcement Learning

Provably Efficient Cooperative Multi-Agent Reinforcement Learning with Function Approximation

2021-03-08 · Abhimanyu Dubey, Alex Pentland

Reinforcement learning in cooperative multi-agent settings has recently advanced significantly in its scope, with applications in cooperative estimation for advertising, dynamic treatment regimes, distributed control, an…

Federated LearningMulti-agent Reinforcement Learningreinforcement-learningReinforcement Learning+1