paper-with-me

홈 › Papers

MELD: Meta-Reinforcement Learning from Images via Latent State Models

2020-10-26 · Tony Z. Zhao, Anusha Nagabandi, Kate Rakelly, Chelsea Finn, Sergey Levine

Meta-reinforcement learning algorithms can enable autonomous agents, such as robots, to quickly acquire new behaviors by leveraging prior experience in a set of related training tasks. However, the onerous data requirements of meta-training compounded with the challenge of learning from sensory inputs such as images have made meta-RL challenging to apply to real robotic systems. Latent state models, which learn compact state representations from a sequence of observations, can accelerate representation learning from visual inputs. In this paper, we leverage the perspective of meta-learning as task inference to show that latent state models can \emph{also} perform meta-learning given an appropriately defined observation space. Building on this insight, we develop meta-RL with latent dynamics (MELD), an algorithm for meta-RL from images that performs inference in a latent state model to quickly acquire new skills given observations and rewards. MELD outperforms prior meta-RL methods on several simulated image-based robotic control problems, and enables a real WidowX robotic arm to insert an Ethernet cable into new locations given a sparse task completion signal after only $8$ hours of real world meta-training. To our knowledge, MELD is the first meta-RL algorithm trained in a real-world robotic control setting from images.

📄 PDF Abstract BibTeX arXiv:2010.13957

Code (1)

tonyzhaozh/meld 공식 구현 tf

Tasks

Meta-LearningMeta Reinforcement Learningreinforcement-learningReinforcement LearningReinforcement Learning (RL)Representation Learning

Similar Papers 제목 키워드 기반

TimeLDM: Latent Diffusion Model for Unconditional Time Series Generation

2024-07-05 · Jian Qian, Bingyu Xie, Biao Wan, Minhao Li 외

Time series generation is a crucial research topic in the area of decision-making systems, which can be particularly important in domains like autonomous driving, healthcare, and, notably, robotics. Recent approaches foc…

Autonomous DrivingData AugmentationMissing ValuesTime Series+1

Latent Bayesian melding for integrating individual and population models

2015-10-30 · NeurIPS 2015 12 · Mingjun Zhong, Nigel Goddard, Charles Sutton

In many statistical problems, a more coarse-grained model may be suitable for population-level behaviour, whereas a more detailed model is appropriate for accurate modelling of individual behaviour. This raises the quest…

blind source separation

Fast moment estimation for generalized latent Dirichlet models

2016-03-17 · Shiwen Zhao, Barbara E. Engelhardt, Sayan Mukherjee, David B. Dunson

We develop a generalized method of moments (GMM) approach for fast parameter estimation in a new class of Dirichlet latent variable models with mixed data types. Parameter estimation via GMM has been demonstrated to have…

parameter estimationVariational Inference

Enhancing experimental signals in single-cell RNA-sequencing data using graph signal processing

2019-03-24 · ICLR Workshop LLD 2019 · Daniel B. Burkhardt, Jay S. Stanley III, Ana Luisa Perdigoto, Scott A. Gigante 외

Single-cell RNA-sequencing (scRNA-seq) is a powerful tool for analyzing biological systems. However, due to biological and technical noise, quantifying the effects of multiple experimental conditions presents an analytic…

GeoMeld: Toward Semantically Grounded Foundation Models for Remote Sensing

2026-04-12 · Maram Hasan, Md Aminur Hossain, Savitra Roy, Souparna Bhowmik 외 arxiv

Effective foundation modeling in remote sensing requires spatially aligned heterogeneous modalities coupled with semantically grounded supervision, yet such resources remain limited at scale. We present GeoMeld, a large-…

Representation Learning