paper-with-me

홈 › Papers

Efficient Empowerment Estimation for Unsupervised Stabilization

2020-07-14 · ICLR 2021 1 · Ruihan Zhao, Kevin Lu, Pieter Abbeel, Stas Tiomkin

Intrinsically motivated artificial agents learn advantageous behavior without externally-provided rewards. Previously, it was shown that maximizing mutual information between agent actuators and future states, known as the empowerment principle, enables unsupervised stabilization of dynamical systems at upright positions, which is a prototypical intrinsically motivated behavior for upright standing and walking. This follows from the coincidence between the objective of stabilization and the objective of empowerment. Unfortunately, sample-based estimation of this kind of mutual information is challenging. Recently, various variational lower bounds (VLBs) on empowerment have been proposed as solutions; however, they are often biased, unstable in training, and have high sample complexity. In this work, we propose an alternative solution based on a trainable representation of a dynamical system as a Gaussian channel, which allows us to efficiently calculate an unbiased estimator of empowerment by convex optimization. We demonstrate our solution for sample-based unsupervised stabilization on different dynamical control systems and show the advantages of our method by comparing it to the existing VLB approaches. Specifically, we show that our method has a lower sample complexity, is more stable in training, possesses the essential properties of the empowerment function, and allows estimation of empowerment from images. Consequently, our method opens a path to wider and easier adoption of empowerment for various applications.

📄 PDF Abstract BibTeX arXiv:2007.07356

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Unsupervised Real-Time Control through Variational Empowerment

2017-10-13 · Maximilian Karl, Maximilian Soelch, Philip Becker-Ehmck, Djalel Benbouzid 외

We introduce a methodology for efficiently computing a lower bound to empowerment, allowing it to be used as an unsupervised cost function for policy learning in real-time control. Empowerment, being the channel capacity…

Unsupervised Skill Discovery through Skill Regions Differentiation

2025-06-17 · Ting Xiao, Jiakun Zheng, Rushuai Yang, Kang Xu 외

Unsupervised Reinforcement Learning (RL) aims to discover diverse behaviors that can accelerate the learning of downstream tasks. Previous methods typically focus on entropy-based exploration or empowerment-driven skill …

Density EstimationReinforcement Learning (RL)Unsupervised Reinforcement Learning

Information-Theoretic Policy Pre-Training with Empowerment

2025-10-07 · Moritz Schneider, Robert Krug, Narunas Vaskevicius, Luigi Palmieri 외 arxiv

Empowerment, an information-theoretic measure of an agent's potential influence on its environment, has emerged as a powerful intrinsic motivation and exploration framework for reinforcement learning (RL). Besides for un…

Reinforcement Learning

Learning to Perceive the World Through Control: Empowerment-Based Representation Learning

2026-05-28 · Mahsa Bastankhah, Sophie Broderick, Benjamin Eysenbach arxiv

In many practical reinforcement learning environments, observations are far higher-dimensional than the variables that matter for control. In this work, we ask: can we learn representations that capture only control-rele…

Representation LearningReinforcement Learning

Intrinsically motivated option learning: a comparative study of recent methods

2022-06-13 · Djordje Božić, Predrag Tadić, Mladen Nikolić

Options represent a framework for reasoning across multiple time scales in reinforcement learning (RL). With the recent active interest in the unsupervised learning paradigm in the RL research community, the option frame…

reinforcement-learningReinforcement Learning (RL)