paper-with-me

Papers

SE3-Pose-Nets: Structured Deep Dynamics Models for Visuomotor Planning and Control

2017-10-02 · Arunkumar Byravan, Felix Leeb, Franziska Meier, Dieter Fox

In this work, we present an approach to deep visuomotor control using structured deep dynamics models. Our deep dynamics model, a variant of SE3-Nets, learns a low-dimensional pose embedding for visuomotor control via an encoder-decoder structure. Unlike prior work, our dynamics model is structured: given an input scene, our network explicitly learns to segment salient parts and predict their pose-embedding along with their motion modeled as a change in the pose space due to the applied actions. We train our model using a pair of point clouds separated by an action and show that given supervision only in the form of point-wise data associations between the frames our network is able to learn a meaningful segmentation of the scene along with consistent poses. We further show that our model can be used for closed-loop control directly in the learned low-dimensional pose space, where the actions are computed by minimizing error in the pose space using gradient-based methods, similar to traditional model-based control. We present results on controlling a Baxter robot from raw depth data in simulation and in the real world and compare against two baseline deep networks. Our method runs in real-time, achieves good prediction of scene dynamics and outperforms the baseline methods on multiple control runs. Video results can be found at: https://rse-lab.cs.washington.edu/se3-structured-deep-ctrl/

📄 PDF Abstract BibTeX arXiv:1710.00489

Code (0)

등록된 구현이 없습니다.

Tasks

Decoder

Similar Papers 제목 키워드 기반

DexFuture: Hierarchical Future-State Visuomotor Targeting for Bimanual Dexterous Tool Use

2026-06-04 · Runfa Blark Li, Kuang-Ting Tu, Nikola Raicevic, Dwait Bhatt 외 arxiv

Bimanual dexterous tool use remains challenging for robots due to high-dimensional hand configurations and complex hand-tool-object dynamics and contact. Most existing control policies depend on future configuration refe…

Goal2Skill: Long-Horizon Manipulation with Adaptive Planning and Reflection

2026-04-15 · Zhen Liu, Xinyu Ning, Zhe Hu, Xinxin Xie 외 arxiv

Recent vision-language-action (VLA) systems have demonstrated strong capabilities in embodied manipulation. However, most existing VLA policies rely on limited observation windows and end-to-end action prediction, which …

Universal Planning Networks: Learning Generalizable Representations for Visuomotor Control

2018-07-01 · ICML 2018 7 · Aravind Srinivas, Allan Jabri, Pieter Abbeel, Sergey Levine 외

A key challenge in complex visuomotor control is learning abstract representations that are effective for specifying goals, planning, and generalization. To this end, we introduce universal planning networks (UPN). …

Imitation LearningReinforcement Learning

Hindsight for Foresight: Unsupervised Structured Dynamics Models from Physical Interaction

2020-08-02 · Iman Nematollahi, Oier Mees, Lukas Hermann, Wolfram Burgard

A key challenge for an agent learning to interact with the world is to reason about physical properties of objects and to foresee their dynamics under the effect of applied forces. In order to scale learning through inte…

ObjectOptical Flow EstimationRobotic Grasping

Learning Predictive Visuomotor Coordination

2025-03-30 · Wenqi Jia, Bolin Lai, Miao Liu, Danfei Xu 외

Understanding and predicting human visuomotor coordination is crucial for applications in robotics, human-computer interaction, and assistive technologies. This work introduces a forecasting-based task for visuomotor mod…