paper-with-me

홈 › Papers

H-WM: Robotic Task and Motion Planning Guided by Hierarchical World Model

2026-02-11 · Jinbang Huang, Wenyuan Chen, Zhiyuan Li, Oscar Pang, Xiao Hu, Lingfeng Zhang, Yuanzhao Hu, Zhanguang Zhang, Mark Coates, Tongtong Cao, Xingyue Quan, Yingxue Zhang arxiv

World models are becoming central to robotic planning and control as they enable prediction of future state transitions. Existing approaches often emphasize video generation or natural-language prediction, which are difficult to ground in robot actions and suffer from compounding errors over long horizons. Classic task and motion planning models world transitions in logical space, enabling robot-executable and robust long-horizon reasoning. However, they typically operate independently of visual perception, preventing synchronized symbolic and visual state prediction. We propose a Hierarchical World Model (H-WM) that jointly predicts logical and visual state transitions within a unified framework. H-WM combines a high-level logical world model with a low-level visual world model, integrating the long-horizon robustness of symbolic reasoning with visual grounding. The hierarchical outputs provide stable intermediate guidance for long-horizon tasks, mitigating error accumulation and enabling robust execution across extended task sequences. Experiments across multiple vision-language-action (VLA) control policies demonstrate the effectiveness and generality of H-WM's guidance.

📄 PDF Abstract BibTeX arXiv:2602.11291

Code (0)

등록된 구현이 없습니다.

Tasks

Visual GroundingVideo GenerationMotion Planning

Similar Papers 제목 키워드 기반

EmbodiedUS-FS: Fast Slow Intelligence for Ultrasound Robotics

2026-06-21 · Fangzhuo Zhang, Xinyu Wang, Xiao Yang, Jinchang Zhang arxiv

Robotic ultrasound scanning in real clinical environments requires both high-level clinical workflow reasoning and low-level closed-loop execution. Physicians natural-language instructions often contain implicit anatomic…

Motion Planning in Compressed Representation Spaces

2026-06-29 · Lukas Lao Beyer, Sertac Karaman arxiv

Deep learning methods have vastly expanded the capabilities of motion planning in robotics applications, as learning priors from large-scale data has been shown to be essential in capturing the highly complex behavior re…

Dimensionality ReductionAutonomous VehiclesMotion Planning

Physics-Guided Hierarchical Reward Mechanism for Learning-Based Robotic Grasping

2022-05-26 · Yunsik Jung, Lingfeng Tao, Michael Bowman, Jiucai Zhang 외

Learning-based grasping can afford real-time grasp motion planning of multi-fingered robotics hands thanks to its high computational efficiency. However, learning-based methods are required to explore large search spaces…

Computational EfficiencyDeep Reinforcement LearningMotion Planningreinforcement-learning+3

Feasibility-Guided Planning over Multi-Specialized Locomotion Policies

2026-02-08 · Ying-Sheng Luo, Lu-Ching Wang, Hanjaya Mandala, Yu-Lun Chou 외 arxiv

Planning over unstructured terrain presents a significant challenge in the field of legged robotics. Although recent works in reinforcement learning have yielded various locomotion strategies, planning over multiple expe…

Reinforcement Learning

From Motion to Behavior: Hierarchical Modeling of Humanoid Generative Behavior Control

2025-05-28 · Jusheng Zhang, Jinzhou Tang, Sidi Liu, Mingyan Li 외

Human motion generative modeling or synthesis aims to characterize complicated human motions of daily activities in diverse real-world environments. However, current research predominantly focuses on either low-level, sh…

Motion GenerationMotion PlanningTask and Motion Planning