paper-with-me

홈 › Papers

Automating Reinforcement Learning with Example-based Resets

2022-04-05 · Jigang Kim, J. Hyeon Park, Daesol Cho, H. Jin Kim

Deep reinforcement learning has enabled robots to learn motor skills from environmental interactions with minimal to no prior knowledge. However, existing reinforcement learning algorithms assume an episodic setting, in which the agent resets to a fixed initial state distribution at the end of each episode, to successfully train the agents from repeated trials. Such reset mechanism, while trivial for simulated tasks, can be challenging to provide for real-world robotics tasks. Resets in robotic systems often require extensive human supervision and task-specific workarounds, which contradicts the goal of autonomous robot learning. In this paper, we propose an extension to conventional reinforcement learning towards greater autonomy by introducing an additional agent that learns to reset in a self-supervised manner. The reset agent preemptively triggers a reset to prevent manual resets and implicitly imposes a curriculum for the forward agent. We apply our method to learn from scratch on a suite of simulated and real-world continuous control tasks and demonstrate that the reset agent successfully learns to reduce manual resets whilst also allowing the forward policy to improve gradually over time.

📄 PDF Abstract BibTeX arXiv:2204.02041

Code (1)

jigangkim/autoreset_rl 공식 구현 tf

Tasks

continuous-controlContinuous ControlDeep Reinforcement Learningreinforcement-learningReinforcement LearningReinforcement Learning (RL)

Similar Papers 제목 키워드 기반

Mini-batch Coresets for Memory-efficient Training of Large Language Models

2024-07-28 · Dang Nguyen, Wenhan Yang, Rathul Anand, Yu Yang 외

Training with larger mini-batches improves the convergence rate and can yield superior performance. However, training with large mini-batches becomes prohibitive for Large Language Models (LLMs), due to the large GPU mem…

GPUNetwork Pruning

Staggered Environment Resets Improve Massively Parallel On-Policy Reinforcement Learning

2025-11-26 · Sid Bharthulwar, Stone Tao, Hao Su arxiv

Massively parallel GPU simulation environments have accelerated reinforcement learning (RL) research by enabling fast data collection for on-policy RL algorithms like Proximal Policy Optimization (PPO). To maximize throu…

Reinforcement Learning

When Learning Is Out of Reach, Reset: Generalization in Autonomous Visuomotor Reinforcement Learning

2023-03-30 · Zichen Zhang, Luca Weihs

Episodic training, where an agent's environment is reset after every success or failure, is the de facto standard when training embodied reinforcement learning (RL) agents. The underlying assumption that the environment …

Reinforcement Learning (RL)

Coresets for Near-Convex Functions

2020-06-09 · NeurIPS 2020 12 · Murad Tukan, Alaa Maalouf, Dan Feldman

Coreset is usually a small weighted subset of $n$ input points in $\mathbb{R}^d$, that provably approximates their loss function for a given set of queries (models, classifiers, etc.). Coresets become increasingly common…

regressionSensitivity

Coresets for Gaussian Mixture Models of Any Shape

2019-06-12 · Dan Feldman, Zahi Kfir, Xuan Wu

An $\varepsilon$-coreset for a given set $D$ of $n$ points, is usually a small weighted set, such that querying the coreset \emph{provably} yields a $(1+\varepsilon)$-factor approximation to the original (full) dataset, …

ClusteringGPU