A Comparison of Reward Functions in Q-Learning Applied to a Cart Position Problem
Growing advancements in reinforcement learning has led to advancements in control theory. Reinforcement learning has effectively solved the inverted pendulum problem and more recently the double inverted pendulum problem. In reinforcement learning, our agents learn by interacting with the control system with the goal of maximizing rewards. In this paper, we explore three such reward functions in the cart position problem. This paper concludes that a discontinuous reward function that gives non-zero rewards to agents only if they are within a given distance from the desired position gives the best results.
Code (1)
Tasks
PositionQ-Learningreinforcement-learningReinforcement LearningReinforcement Learning (RL)Similar Papers 제목 키워드 기반
Design and comparison of two linear controllers with precompensation gain for the Quadruple inverted pendulum
In this work we present a workflow for designing two linear control techniques applied to the dynamic system quadruple inverted pendulum mounted on a cart (QIP) where the steady state error on cart position is eliminated…
PositionDynamics-Aware Comparison of Learned Reward Functions
The ability to learn reward functions plays an important role in enabling the deployment of intelligent agents in the real world. However, comparing reward functions, for example as a means of evaluating reward learning …
Knee Cartilage Segmentation Using Diffusion-Weighted MRI
The integrity of articular cartilage is a crucial aspect in the early diagnosis of osteoarthritis (OA). Many novel MRI techniques have the potential to assess compositional changes of the cartilage extracellular matrix. …
SegmentationBalancing a Stick with Eyes Shut: Inverted Pendulum on a Cart without Angle Measurement
We consider linear time-invariant dynamic systems in the single-input, single-output (SISO) framework. In particular, we consider stabilization of an inverted pendulum on a cart using a force on the cart. This system is …
PositionSensitivityA Unified Framework for Zero-Shot Reinforcement Learning
Zero-shot reinforcement learning (RL) has emerged as a setting for developing general agents, capable of solving downstream tasks without additional training or planning at test-time. While conventional RL optimizes poli…
Reinforcement Learning