paper-with-me

홈 › Papers

Tube-NeRF: Efficient Imitation Learning of Visuomotor Policies from MPC using Tube-Guided Data Augmentation and NeRFs

2023-11-23 · Andrea Tagliabue, Jonathan P. How

Imitation learning (IL) can train computationally-efficient sensorimotor policies from a resource-intensive Model Predictive Controller (MPC), but it often requires many samples, leading to long training times or limited robustness. To address these issues, we combine IL with a variant of robust MPC that accounts for process and sensing uncertainties, and we design a data augmentation (DA) strategy that enables efficient learning of vision-based policies. The proposed DA method, named Tube-NeRF, leverages Neural Radiance Fields (NeRFs) to generate novel synthetic images, and uses properties of the robust MPC (the tube) to select relevant views and to efficiently compute the corresponding actions. We tailor our approach to the task of localization and trajectory tracking on a multirotor, by learning a visuomotor policy that generates control actions using images from the onboard camera as only source of horizontal position. Numerical evaluations show 80-fold increase in demonstration efficiency and a 50% reduction in training time over current IL methods. Additionally, our policies successfully transfer to a real multirotor, achieving low tracking errors despite large disturbances, with an onboard inference time of only 1.5 ms. Video: https://youtu.be/_W5z33ZK1m4

📄 PDF Abstract BibTeX arXiv:2311.14153

Code (0)

등록된 구현이 없습니다.

Tasks

Data AugmentationImitation LearningNeRF

Similar Papers 제목 키워드 기반

Output Feedback Tube MPC-Guided Data Augmentation for Robust, Efficient Sensorimotor Policy Learning

2022-10-18 · Andrea Tagliabue, Jonathan P. How

Imitation learning (IL) can generate computationally efficient sensorimotor policies from demonstrations provided by computationally expensive model-based sensing and control algorithms. However, commonly employed IL met…

Data AugmentationImitation Learning

Learning to Drive by Watching YouTube Videos: Action-Conditioned Contrastive Policy Pretraining

2022-04-05 · Qihang Zhang, Zhenghao Peng, Bolei Zhou

Deep visuomotor policy learning, which aims to map raw visual observation to action, achieves promising results in control tasks such as robotic manipulation and autonomous driving. However, it requires a huge number of …

Autonomous DrivingImitation Learning

Dynamic Tube MPC for Nonlinear Systems

2019-07-15

Modeling error or external disturbances can severely degrade the performance of Model Predictive Control (MPC) in real-world scenarios. Robust MPC (RMPC) addresses this limitation by optimizing over feedback policies but…

Model Predictive Control

A Comprehensive General Model of Tendon-Actuated Concentric Tube Robots with Multiple Tubes and Tendons

2025-10-28 · Pejman Kheradmand, Behnam Moradkhani, Raghavasimhan Sankaranarayanan, Kent K. Yamamoto 외 arxiv

Tendon-actuated concentric tube mechanisms combine the advantages of tendon-driven continuum robots and concentric tube robots while addressing their respective limitations. They overcome the restricted degrees of freedo…

TubeRMC: Tube-conditioned Reconstruction with Mutual Constraints for Weakly-supervised Spatio-Temporal Video Grounding

2025-11-13 · Jinxuan Li, Yi Zhang, Jian-Fang Hu, Chaolei Tan 외 arxiv

Spatio-Temporal Video Grounding (STVG) aims to localize a spatio-temporal tube that corresponds to a given language query in an untrimmed video. This is a challenging task since it involves complex vision-language unders…

Spatio-Temporal Video GroundingVisual Grounding