paper-with-me

홈 › Papers

Data-efficient visuomotor policy training using reinforcement learning and generative models

2020-07-26 · Ali Ghadirzadeh, Petra Poklukar, Ville Kyrki, Danica Kragic, Mårten Björkman

We present a data-efficient framework for solving visuomotor sequential decision-making problems which exploits the combination of reinforcement learning (RL) and latent variable generative models. Our framework trains deep visuomotor policies by introducing an action latent variable such that the feed-forward policy search can be divided into three parts: (i) training a sub-policy that outputs a distribution over the action latent variable given a state of the system, (ii) unsupervised training of a generative model that outputs a sequence of motor actions conditioned on the latent action variable, and (iii) supervised training of the deep visuomotor policy in an end-to-end fashion. Our approach enables safe exploration and alleviates the data-inefficiency problem as it exploits prior knowledge about valid sequences of motor actions. Moreover, we provide a set of measures for evaluation of generative models such that we are able to predict the performance of the RL policy training prior to the actual training on a physical robot. We define two novel measures of disentanglement and local linearity for assessing the quality of latent representations, and complement them with existing measures for assessment of the learned distribution. We experimentally determine the characteristics of different generative models that have the most influence on performance of the final policy training on a robotic picking task.

📄 PDF Abstract BibTeX arXiv:2007.13134

Code (0)

등록된 구현이 없습니다.

Tasks

Decision MakingDisentanglementreinforcement-learningReinforcement Learning (RL)Safe ExplorationSequential Decision Makingvalid

Similar Papers 제목 키워드 기반

Adversarial Feature Training for Generalizable Robotic Visuomotor Control

2019-09-17 · Xi Chen, Ali Ghadirzadeh, Mårten Björkman, Patric Jensfelt

Deep reinforcement learning (RL) has enabled training action-selection policies, end-to-end, by learning a function which maps image pixels to action outputs. However, it's application to visuomotor robotic policy traini…

Deep Reinforcement LearningReinforcement LearningReinforcement Learning (RL)Transfer Learning

Oracle-Guided Masked Contrastive Reinforcement Learning for Visuomotor Policies

2025-10-07 · Yuhang Zhang, Jiaping Xiao, Chao Yan, Mir Feroskhan arxiv

A prevailing approach for learning visuomotor policies is to employ reinforcement learning to map high-dimensional visual observations directly to action commands. However, the combination of high-dimensional visual inpu…

Representation LearningReinforcement LearningContrastive Learning

High-Fidelity One-Step Generative Visuomotor Policy via Recursive Correction, Frequency Consistency, and Contrastive Flow Matching

2026-07-04 · Yuran Chen, Xinye Cai, Zhonglin Gong, Yang Huang arxiv

Generative models such as diffusion and flow matching have advanced robotic visuomotor policies by modeling multimodal action distributions, but their multi-step sampling or ODE solving introduces inference latency. Exis…

Normalizing Flows are Capable Models for Bi-manual Visuomotor Policy

2025-09-25 · Jialong Li, Simon Kristoffersson Lind, Wenrui Xie, Maj Stenmark 외 arxiv

The field of general-purpose robotics has recently embraced powerful probabilistic diffusion-based models to learn the complex embodiment behaviours. However, existing models often come with significant trade-offs, namel…

Learning Visuomotor Policy for Multi-Robot Laser Tag Game

2026-03-12 · Kai Li, Shiyu Zhao arxiv

In this paper, we study multi robot laser tag, a simplified yet practical shooting-game-style task. Classic modular approaches on these tasks face challenges such as limited observability and reliance on depth mapping an…

Reinforcement LearningCollision Avoidance