paper-with-me

Papers

Example-Driven Model-Based Reinforcement Learning for Solving Long-Horizon Visuomotor Tasks

2021-09-21 · Bohan Wu, Suraj Nair, Li Fei-Fei, Chelsea Finn

In this paper, we study the problem of learning a repertoire of low-level skills from raw images that can be sequenced to complete long-horizon visuomotor tasks. Reinforcement learning (RL) is a promising approach for acquiring short-horizon skills autonomously. However, the focus of RL algorithms has largely been on the success of those individual skills, more so than learning and grounding a large repertoire of skills that can be sequenced to complete extended multi-stage tasks. The latter demands robustness and persistence, as errors in skills can compound over time, and may require the robot to have a number of primitive skills in its repertoire, rather than just one. To this end, we introduce EMBER, a model-based RL method for learning primitive skills that are suitable for completing long-horizon visuomotor tasks. EMBER learns and plans using a learned model, critic, and success classifier, where the success classifier serves both as a reward function for RL and as a grounding mechanism to continuously detect if the robot should retry a skill when unsuccessful or under perturbations. Further, the learned model is task-agnostic and trained using data from all skills, enabling the robot to efficiently learn a number of distinct primitives. These visuomotor primitive skills and their associated pre- and post-conditions can then be directly combined with off-the-shelf symbolic planners to complete long-horizon tasks. On a Franka Emika robot arm, we find that EMBER enables the robot to complete three long-horizon visuomotor tasks at 85% success rate, such as organizing an office desk, a file cabinet, and drawers, which require sequencing up to 12 skills, involve 14 unique learned primitives, and demand generalization to novel objects.

📄 PDF Abstract BibTeX arXiv:2109.10312

Code (0)

등록된 구현이 없습니다.

Tasks

Model-based Reinforcement Learningreinforcement-learningReinforcement Learning (RL)

Similar Papers 제목 키워드 기반

O3D: Offline Data-driven Discovery and Distillation for Sequential Decision-Making with Large Language Models

2023-10-22 · Yuchen Xiao, Yanchao Sun, Mengda Xu, Udari Madhushani 외

Recent advancements in large language models (LLMs) have exhibited promising performance in solving sequential decision-making problems. By imitating few-shot examples provided in the prompts (i.e., in-context learning),…

Decision MakingIn-Context LearningSequential Decision Making

DISCOVER: Automated Curricula for Sparse-Reward Reinforcement Learning

2025-05-26 · Leander Diaz-Bone, Marco Bagatella, Jonas Hübotter, Andreas Krause

Sparse-reward reinforcement learning (RL) can model a wide range of highly complex tasks. Solving sparse-reward tasks is RL's core premise - requiring efficient exploration coupled with long-horizon credit assignment - a…

Efficient Explorationreinforcement-learningReinforcement LearningReinforcement Learning (RL)

Relay Policy Learning: Solving Long-Horizon Tasks via Imitation and Reinforcement Learning

2019-10-25 · Abhishek Gupta, Vikash Kumar, Corey Lynch, Sergey Levine 외

We present relay policy learning, a method for imitation and reinforcement learning that can solve multi-stage, long-horizon robotic tasks. This general and universally-applicable, two-phase approach consists of an imita…

Imitation Learningreinforcement-learningReinforcement LearningReinforcement Learning (RL)

Data-Driven Mean Field Equilibrium Computation in Large-Population LQG Games

2025-02-27 · Zhenhui Xu, Jiayu Chen, Bing-Chang Wang, Tielong Shen

This paper presents a novel data-driven approach for approximating the $\varepsilon$-Nash equilibrium in continuous-time linear quadratic Gaussian (LQG) games, where multiple agents interact with each other through their…

Reinforcement Learning with Anticipation: A Hierarchical Approach for Long-Horizon Tasks

2025-09-06 · Yang Yu arxiv

Solving long-horizon goal-conditioned tasks remains a significant challenge in reinforcement learning (RL). Hierarchical reinforcement learning (HRL) addresses this by decomposing tasks into more manageable sub-tasks, bu…

Hierarchical Reinforcement Learning