paper-with-me

홈 › Papers

Compositional Motion Generation from Demonstration with Object-Centric Neural Fields

2026-07-08 · Ahmet Ercan Tekden, Yasemin Bekiroglu arxiv

Compositionality, by organizing complex behavior as combinations of simpler elements, enables robot learning that is scalable and data efficient. Leveraging this principle, we propose a generative learning-from-demonstration framework that enables compositional modeling of robotic behavior by connecting perception and motion through shared object-level representations. We render scenes from object-centric neural representations that integrate canonical neural fields with latent-conditioned deformations, capturing positional and geometric variations in a smooth, consistent, and interpretable way. For motion generation, a temporal mixture-of-experts (MoE) employs a gating mechanism to combine object-conditioned movement primitives over time, producing complete trajectories. This spatial-temporal compositionality maintains the data efficiency of movement primitives while grounding motion in visual structure, enabling systematic generalization across diverse scene configurations. In simulation, long-horizon manipulation tasks are successfully completed using the proposed model, which requires significantly less training data than other image-based baselines. Real-world experiments further demonstrate the method's robustness to noise, its ability to generalize at the category level through language-based segmentation models, and its capacity to operate directly on 3D scene representations.

📄 PDF Abstract BibTeX arXiv:2607.07129

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

RoboDream: Compositional World Models for Scalable Robot Data Synthesis

2026-06-01 · Junjie Ye, Rong Xue, Basile Van Hoorick, Runhao Li 외 arxiv

Scaling robot learning requires large-scale, diverse demonstrations, yet real-world data collection via teleoperation remains prohibitively expensive and time-consuming. While video diffusion models offer a promising ave…

Human2Any: Human-to-Robot Transfer via Constraint-Aware Compositional Planning

2026-06-27 · Shuo Cheng, Chuye Zhang, Alfred Cueva, Caelan Garrett 외 arxiv

Human videos are a scalable source of supervision for robot manipulation, as they are abundant and naturally capture rich object interactions. However, transferring human demonstrations to robots remains challenging due …

Robot ManipulationMotion Planning

Joint Flow Trajectory Optimization For Feasible Robot Motion Generation from Video Demonstrations

2025-09-25 · Xiaoxiang Dong, Matthew Johnson-Roberson, Weiming Zhi arxiv

Learning from human video demonstrations offers a scalable alternative to teleoperation or kinesthetic teaching, but poses challenges for robot manipulators due to embodiment differences and joint feasibility constraints…

Comp4D: LLM-Guided Compositional 4D Scene Generation

2024-03-25 · Dejia Xu, Hanwen Liang, Neel P. Bhatt, Hezhen Hu 외

Recent advancements in diffusion models for 2D and 3D content creation have sparked a surge of interest in generating 4D content. However, the scarcity of 3D scene datasets constrains current methodologies to primarily o…

ObjectScene GenerationText to 3D

ROOTS: Object-Centric Representation and Rendering of 3D Scenes

2020-06-11 · Chang Chen, Fei Deng, Sungjin Ahn

A crucial ability of human intelligence is to build up models of individual 3D objects from partial scene observations. Recent works achieve object-centric generation but without the ability to infer the representation, …

ObjectRepresentation LearningScene Generation