paper-with-me

홈 › Papers

FoldPath: End-to-End Object-Centric Motion Generation via Modulated Implicit Paths

2025-11-03 · Paolo Rabino, Gabriele Tiboni, Tatiana Tommasi arxiv

Object-Centric Motion Generation (OCMG) is instrumental in advancing automated manufacturing processes, particularly in domains requiring high-precision expert robotic motions, such as spray painting and welding. To realize effective automation, robust algorithms are essential for generating extended, object-aware trajectories across intricate 3D geometries. However, contemporary OCMG techniques are either based on ad-hoc heuristics or employ learning-based pipelines that are still reliant on sensitive post-processing steps to generate executable paths. We introduce FoldPath, a novel, end-to-end, neural field based method for OCMG. Unlike prior deep learning approaches that predict discrete sequences of end-effector waypoints, FoldPath learns the robot motion as a continuous function, thus implicitly encoding smooth output paths. This paradigm shift eliminates the need for brittle post-processing steps that concatenate and order the predicted discrete waypoints. Particularly, our approach demonstrates superior predictive performance compared to recently proposed learning-based methods, and attains generalization capabilities even in real industrial settings, where only a limited amount of 70 expert samples are provided. We validate FoldPath through comprehensive experiments in a realistic simulation environment and introduce new, rigorous metrics designed to comprehensively evaluate long-horizon robotic paths, thus advancing the OCMG task towards practical maturity.

📄 PDF Abstract BibTeX arXiv:2511.01407

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

TIV-Diffusion: Towards Object-Centric Movement for Text-driven Image to Video Generation

2024-12-13 · Xingrui Wang, Xin Li, Yaosi Hu, Hanxin Zhu 외

Text-driven Image to Video Generation (TI2V) aims to generate controllable video given the first frame and corresponding textual description. The primary challenges of this task lie in two parts: (i) how to identify the …

Image to Video GenerationObjectVideo Generation

GIMO: Gaze-Informed Human Motion Prediction in Context

2022-04-20 · Yang Zheng, Yanchao Yang, Kaichun Mo, Jiaman Li 외

Predicting human motion is critical for assistive robots and AR/VR applications, where the interaction with humans needs to be safe and comfortable. Meanwhile, an accurate prediction depends on understanding both the sce…

Human motion predictionmotion predictionPrediction

EgoFlow: Gradient-Guided Flow Matching for Egocentric 6DoF Object Motion Generation

2026-04-01 · Abhishek Saroha, Huajian Zeng, Xingxing Zuo, Daniel Cremers 외 arxiv

Understanding and predicting object motion from egocentric video is fundamental to embodied perception and interaction. However, generating physically consistent 6DoF trajectories remains challenging due to occlusions, f…

Collision Avoidance

Generating Human Motion in 3D Scenes from Text Descriptions

2024-05-13 · CVPR 2024 1 · Zhi Cen, Huaijin Pi, Sida Peng, Zehong Shen 외

Generating human motions from textual descriptions has gained growing research interest due to its wide range of applications. However, only a few works consider human-scene interactions together with text conditions, wh…

Motion GenerationObjectSpatial Reasoning

Benchmarks and Challenges in Pose Estimation for Egocentric Hand Interactions with Objects

2024-03-25 · Zicong Fan, Takehiko Ohkawa, Linlin Yang, Nie Lin 외

We interact with the world with our hands and see it through our own (egocentric) perspective. A holistic 3Dunderstanding of such interactions from egocentric views is important for tasks in robotics, AR/VR, action recog…

Action RecognitionMotion GenerationObjectObject Reconstruction+1