Information-theoretic Model Identification and Policy Search using Physics Engines with Application to Robotic Manipulation
We consider the problem of a robot learning the mechanical properties of objects through physical interaction with the object, and introduce a practical, data-efficient approach for identifying the motion models of these objects. The proposed method utilizes a physics engine, where the robot seeks to identify the inertial and friction parameters of the object by simulating its motion under different values of the parameters and identifying those that result in a simulation which matches the observed real motions. The problem is solved in a Bayesian optimization framework. The same framework is used for both identifying the model of an object online and searching for a policy that would minimize a given cost function according to the identified model. Experimental results both in simulation and using a real robot indicate that the proposed method outperforms state-of-the-art model-free reinforcement learning approaches.
Code (0)
등록된 구현이 없습니다.
Tasks
Bayesian OptimizationFrictionObjectReinforcement LearningSimilar Papers 제목 키워드 기반
Fast Model Identification via Physics Engines for Data-Efficient Policy Search
This paper presents a method for identifying mechanical parameters of robots or objects, such as their mass and friction coefficients. Key features are the use of off-the-shelf physics engines and the adaptation of a Bay…
Bayesian OptimizationFrictionModel-based Reinforcement LearningReinforcement LearningPhysics-Regulated Deep Reinforcement Learning: Invariant Embeddings
This paper proposes the Phy-DRL: a physics-regulated deep reinforcement learning (DRL) framework for safety-critical autonomous systems. The Phy-DRL has three distinguished invariant-embedding designs: i) residual action…
Deep Reinforcement Learningreinforcement-learningReinforcement LearningPhysics-Informed Causal MDPs for Sequential Constraint Repair in Engineering Simulation Pipelines
Off-policy learning in constrained MDPs with large binary state spaces faces a fundamental tension: causal identification of transition dynamics requires structural assumptions, while sample-efficient policy learning req…
Reinforcement Learning with Continuous Actions Under Unmeasured Confounding
This paper addresses the challenge of offline policy learning in reinforcement learning with continuous action spaces when unmeasured confounders are present. While most existing research focuses on policy evaluation wit…
reinforcement-learningReinforcement LearningBlessing from Human-AI Interaction: Super Reinforcement Learning in Confounded Environments
As AI becomes more prevalent throughout society, effective methods of integrating humans and AI systems that leverage their respective strengths and mitigate risk have become an important priority. In this paper, we intr…
Decision Makingreinforcement-learningReinforcement LearningReinforcement Learning (RL)+1