paper-with-me

Papers

Toward Open-ended Embodied Tasks Solving

2023-12-10 · William Wei Wang, Dongqi Han, Xufang Luo, Yifei Shen, Charles Ling, Boyu Wang, Dongsheng Li

Empowering embodied agents, such as robots, with Artificial Intelligence (AI) has become increasingly important in recent years. A major challenge is task open-endedness. In practice, robots often need to perform tasks with novel goals that are multifaceted, dynamic, lack a definitive "end-state", and were not encountered during training. To tackle this problem, this paper introduces \textit{Diffusion for Open-ended Goals} (DOG), a novel framework designed to enable embodied AI to plan and act flexibly and dynamically for open-ended task goals. DOG synergizes the generative prowess of diffusion models with state-of-the-art, training-free guidance techniques to adaptively perform online planning and control. Our evaluations demonstrate that DOG can handle various kinds of novel task goals not seen during training, in both maze navigation and robot control problems. Our work sheds light on enhancing embodied AI's adaptability and competency in tackling open-ended goals.

📄 PDF Abstract BibTeX arXiv:2312.05822

Code (0)

등록된 구현이 없습니다.

Methods 이 논문이 사용한 방법론

Diffusion Diffusion models generate samples by gradually removing noise from a signal, and their training objective can be expressed as a reweighted variational lower-bound…

Similar Papers 제목 키워드 기반

Do We Really Need a Complex Agent System? Distill Embodied Agent into a Single Model

2024-04-06 · Zhonghan Zhao, Ke Ma, Wenhao Chai, Xuan Wang 외

With the power of large language models (LLMs), open-ended embodied agents can flexibly understand human instructions, generate interpretable guidance strategies, and output executable actions. Nowadays, Multi-modal Lang…

Knowledge Distillation

MP5: A Multi-modal Open-ended Embodied System in Minecraft via Active Perception

2023-12-12 · CVPR 2024 1 · Yiran Qin, Enshen Zhou, Qichang Liu, Zhenfei Yin 외

It is a long-lasting goal to design an embodied system that can solve long-horizon open-world tasks in human-like ways. However, existing approaches usually struggle with compound difficulties caused by the logic-aware d…

Minecraft

Building Open-Ended Embodied Agent via Language-Policy Bidirectional Adaptation

2023-12-12 · Shaopeng Zhai, Jie Wang, Tianyi Zhang, Fuxian Huang 외

Building embodied agents on integrating Large Language Models (LLMs) and Reinforcement Learning (RL) have revolutionized human-AI interaction: researchers can now leverage language instructions to plan decision-making fo…

Decision MakingLanguage ModellingReinforcement Learning (RL)

Luban: Building Open-Ended Creative Agents via Autonomous Embodied Verification

2024-05-24 · Yuxuan Guo, Shaohui Peng, Jiaming Guo, Di Huang 외

Building open agents has always been the ultimate goal in AI research, and creative agents are the more enticing. Existing LLM agents excel at long-horizon tasks with well-defined goals (e.g., `mine diamonds' in Minecraf…

Minecraft

FOUNDER: Grounding Foundation Models in World Models for Open-Ended Embodied Decision Making

2025-07-15 · Yucen Wang, Rui Yu, Shenghua Wan, Le Gan 외

Foundation Models (FMs) and World Models (WMs) offer complementary strengths in task generalization at different levels. In this work, we propose FOUNDER, a framework that integrates the generalizable knowledge embedded …

Decision Making