paper-with-me

Papers

Cascaded Diffusion Models for Neural Motion Planning

2025-05-21 · Mohit Sharma, Adam Fishman, Vikash Kumar, Chris Paxton, Oliver Kroemer

Robots in the real world need to perceive and move to goals in complex environments without collisions. Avoiding collisions is especially difficult when relying on sensor perception and when goals are among clutter. Diffusion policies and other generative models have shown strong performance in solving local planning problems, but often struggle at avoiding all of the subtle constraint violations that characterize truly challenging global motion planning problems. In this work, we propose an approach for learning global motion planning using diffusion policies, allowing the robot to generate full trajectories through complex scenes and reasoning about multiple obstacles along the path. Our approach uses cascaded hierarchical models which unify global prediction and local refinement together with online plan repair to ensure the trajectories are collision free. Our method outperforms (by ~5%) a wide variety of baselines on challenging tasks in multiple domains including navigation and manipulation.

📄 PDF Abstract BibTeX arXiv:2505.15157

Code (0)

등록된 구현이 없습니다.

Tasks

Motion Planning

Methods 이 논문이 사용한 방법론

Diffusion Diffusion models generate samples by gradually removing noise from a signal, and their training objective can be expressed as a reweighted variational lower-bound…

Similar Papers 제목 키워드 기반

DiffDance: Cascaded Human Motion Diffusion Model for Dance Generation

2023-08-05 · Qiaosong Qi, Le Zhuo, Aixi Zhang, Yue Liao 외

When hearing music, it is natural for people to dance to its rhythm. Automatic dance generation, however, is a challenging task due to the physical constraints of human motion and rhythmic alignment with target music. Co…

Representation LearningRhythmSuper-Resolution

LUVE : Latent-Cascaded Ultra-High-Resolution Video Generation with Dual Frequency Experts

2026-02-12 · Chen Zhao, Jiawei Chen, Hongyu Li, Zhuoliang Kang 외 arxiv

Recent advances in video diffusion models have significantly improved visual quality, yet ultra-high-resolution (UHR) video generation remains a formidable challenge due to the compounded difficulties of motion modeling,…

Video Generation

DreamMotion: Space-Time Self-Similar Score Distillation for Zero-Shot Video Editing

2024-03-18 · Hyeonho Jeong, Jinho Chang, Geon Yeong Park, Jong Chul Ye

Text-driven diffusion-based video editing presents a unique challenge not encountered in image editing literature: establishing real-world motion. Unlike existing video editing approaches, here we focus on score distilla…

Video Editing

Dynamics Learning with Cascaded Variational Inference for Multi-Step Manipulation

2019-10-29 · Kuan Fang, Yuke Zhu, Animesh Garg, Silvio Savarese 외

The fundamental challenge of planning for multi-step manipulation is to find effective and plausible action sequences that lead to the task goal. We present Cascaded Variational Inference (CAVIN) Planner, a model-based m…

Variational Inference

Text-based Talking Video Editing with Cascaded Conditional Diffusion

2024-07-20 · Bo Han, Heqing Zou, Haoyang Li, Guangcong Wang 외

Text-based talking-head video editing aims to efficiently insert, delete, and substitute segments of talking videos through a user-friendly text editing approach. It is challenging because of \textbf{1)} generalizable ta…

Video Editing