Low-cost Real-world Implementation of the Swing-up Pendulum for Deep Reinforcement Learning Experiments
Deep reinforcement learning (DRL) has had success in virtual and simulated domains, but due to key differences between simulated and real-world environments, DRL-trained policies have had limited success in real-world applications. To assist researchers to bridge the \textit{sim-to-real gap}, in this paper, we describe a low-cost physical inverted pendulum apparatus and software environment for exploring sim-to-real DRL methods. In particular, the design of our apparatus enables detailed examination of the delays that arise in physical systems when sensing, communicating, learning, inferring and actuating. Moreover, we wish to improve access to educational systems, so our apparatus uses readily available materials and parts to reduce cost and logistical barriers. Our design shows how commercial, off-the-shelf electronics and electromechanical and sensor systems, combined with common metal extrusions, dowel and 3D printed couplings provide a pathway for affordable physical DRL apparatus. The physical apparatus is complemented with a simulated environment implemented using a high-fidelity physics engine and OpenAI Gym interface.
Code (0)
등록된 구현이 없습니다.
Tasks
Deep Reinforcement LearningOpenAI GymSimilar Papers 제목 키워드 기반
Swing-Up of a Weakly Actuated Double Pendulum via Nonlinear Normal Modes
We identify the nonlinear normal modes spawning from the stable equilibrium of a double pendulum under gravity, and we establish their connection to homoclinic orbits through the unstable upright position as energy incre…
PositionA Deep Reinforcement Learning Approach towards Pendulum Swing-up Problem based on TF-Agents
Adapting the idea of training CartPole with Deep Q-learning agent, we are able to find a promising result that prevent the pole from falling down. The capacity of reinforcement learning (RL) to learn from the interaction…
Deep Reinforcement LearningPositionQ-Learningreinforcement-learning+1Payload Swing Estimation and Damping Without Payload Parameters for Multirotor UAVs
Cable-suspended payload transport by multirotor UAVs is flexible but generates periodic swing disturbance that degrades tracking and risks instability. Existing anti-swing methods require additional sensors or precise id…
Real-Time Model Predictive Control for the Swing-Up Problem of an Underactuated Double Pendulum
The 3rd AI Olympics with RealAIGym competition poses the challenge of developing a global policy that can swing up and stabilize an underactuated 2-link system Acrobot and/or Pendubot from any configuration in the state …
AcrobotModel Predictive ControlEnhancement of Energy-Based Swing-Up Controller via Entropy Search
An energy based approach for stabilizing a mechanical system has offered a simple yet powerful control scheme. However, since it does not impose such strong constraints on parameter space of the controller, finding appro…
Bayesian Optimization