paper-with-me

홈 › Papers

Variational Intrinsic Control Revisited

2020-10-07 · ICLR 2021 1 · Taehwan Kwon

In this paper, we revisit variational intrinsic control (VIC), an unsupervised reinforcement learning method for finding the largest set of intrinsic options available to an agent. In the original work by Gregor et al. (2016), two VIC algorithms were proposed: one that represents the options explicitly, and the other that does it implicitly. We show that the intrinsic reward used in the latter is subject to bias in stochastic environments, causing convergence to suboptimal solutions. To correct this behavior and achieve the maximal empowerment, we propose two methods respectively based on the transitional probability model and Gaussian mixture model. We substantiate our claims through rigorous mathematical derivations and experimental analyses.

📄 PDF Abstract BibTeX arXiv:2010.03281

Code (0)

등록된 구현이 없습니다.

Tasks

reinforcement-learningReinforcement Learning (RL)Unsupervised Reinforcement Learning

Similar Papers 제목 키워드 기반

Intrinsic Control of Variational Beliefs in Dynamic Partially-Observed Visual Environments

2021-06-13 · ICML Workshop URL 2021 7 · Nicholas Rhinehart, Jenny Wang, Glen Berseth, John D Co-Reyes 외

Humans and animals explore their environment and acquire useful skills even in the absence of clear goals, exhibiting intrinsic motivation. The study of intrinsic motivation in artificial agents is concerned with the fol…

Deep Reinforcement Learning

Systematic and multifactor risk models revisited

2013-12-18 · Michel Fliess, Cédric Join

Systematic and multifactor risk models are revisited via methods which were already successfully developed in signal processing and in automatic control. The results, which bypass the usual criticisms on those risk model…

Universal AI maximizes Variational Empowerment

2025-02-20 · Yusuke Hayashi, Koichi Takahashi

This paper presents a theoretical framework unifying AIXI -- a model of universal AI -- with variational empowerment as an intrinsic drive for exploration. We build on the existing framework of Self-AIXI -- a universal l…

VASE: Variational Assorted Surprise Exploration for Reinforcement Learning

2019-10-31 · Haitao Xu, Brendan McCane, Lech Szymanski

Exploration in environments with continuous control and sparse rewards remains a key challenge in reinforcement learning (RL). Recently, surprise has been used as an intrinsic reward that encourages systematic and effici…

continuous-controlContinuous ControlEfficient Explorationreinforcement-learning+3

Provably Safe Stein Variational Clarity-Aware Informative Planning

2025-11-13 · Kaleb Ben Naveed, Utkrisht Sahai, Anouck Girard, Dimitra Panagou arxiv

Autonomous robots are increasingly deployed for information-gathering tasks in environments that vary across space and time. Planning informative and safe trajectories in such settings is challenging because information …

Bayesian Inference