paper-with-me

홈 › Papers

REMAC: Self-Reflective and Self-Evolving Multi-Agent Collaboration for Long-Horizon Robot Manipulation

2025-03-28 · Puzhen Yuan, Angyuan Ma, Yunchao Yao, Huaxiu Yao, Masayoshi Tomizuka, Mingyu Ding

Vision-language models (VLMs) have demonstrated remarkable capabilities in robotic planning, particularly for long-horizon tasks that require a holistic understanding of the environment for task decomposition. Existing methods typically rely on prior environmental knowledge or carefully designed task-specific prompts, making them struggle with dynamic scene changes or unexpected task conditions, e.g., a robot attempting to put a carrot in the microwave but finds the door was closed. Such challenges underscore two critical issues: adaptability and efficiency. To address them, in this work, we propose an adaptive multi-agent planning framework, termed REMAC, that enables efficient, scene-agnostic multi-robot long-horizon task planning and execution through continuous reflection and self-evolution. REMAC incorporates two key modules: a self-reflection module performing pre-condition and post-condition checks in the loop to evaluate progress and refine plans, and a self-evolvement module dynamically adapting plans based on scene-specific reasoning. It offers several appealing benefits: 1) Robots can initially explore and reason about the environment without complex prompt design. 2) Robots can keep reflecting on potential planning errors and adapting the plan based on task-specific insights. 3) After iterations, a robot can call another one to coordinate tasks in parallel, maximizing the task execution efficiency. To validate REMAC's effectiveness, we build a multi-agent environment for long-horizon robot manipulation and navigation based on RoboCasa, featuring 4 task categories with 27 task styles and 50+ different objects. Based on it, we further benchmark state-of-the-art reasoning models, including DeepSeek-R1, o3-mini, QwQ, and Grok3, demonstrating REMAC's superiority by boosting average success rates by 40% and execution efficiency by 52.7% over the single robot baseline.

📄 PDF Abstract BibTeX arXiv:2503.22122

Code (0)

등록된 구현이 없습니다.

Tasks

Robot ManipulationTask Planning

Similar Papers 제목 키워드 기반

Self-evolving Agents with reflective and memory-augmented abilities

2024-09-01 · Xuechen Liang, Yangfan He, Yinghui Xia, Xinyuan Song 외

Large language models (LLMs) have made significant advances in the field of natural language processing, but they still face challenges such as continuous decision-making. In this research, we propose a novel framework b…

Decision Making

VoTranhOmniLearner - A Self-Evolving Digital Universe for Simulating Consciousness and Societal Dynamics

2025-04-26 · Independent publication 2025 4 · Vi Nhat Son

VoTranhOmniLearner is an innovative simulation framework designed to model a dynamic digital universe with thousands of entities, intricate social interactions, and self-reflective processes. By integrating machine learn…

Management

Learn Like Humans: Use Meta-cognitive Reflection for Efficient Self-Improvement

2026-01-17 · Xinmeng Hou, Peiliang Gong, Bohao Qu, Wuqi Wang 외 arxiv

While Large Language Models (LLMs) enable complex autonomous behavior, current agents remain constrained by static, human-designed prompts that limit adaptability. Existing self-improving frameworks attempt to bridge thi…

Quantum Semi-Supervised Learning with Quantum Supremacy

2021-10-05 · Zhou Shangnan

Quantum machine learning promises to efficiently solve important problems. There are two persistent challenges in classical machine learning: the lack of labeled data, and the limit of computational power. We propose a n…

BIG-bench Machine LearningClusteringQuantum Machine Learning

Be Your Own Red Teamer: Safety Alignment via Self-Play and Reflective Experience Replay

2026-01-15 · Hao Wang, Yanting Wang, Hao Li, Rui Li 외 arxiv

Large Language Models (LLMs) have achieved remarkable capabilities but remain vulnerable to adversarial ``jailbreak'' attacks designed to bypass safety guardrails. Current safety alignment methods depend heavily on stati…

Reinforcement LearningRed Teaming