paper-with-me

Papers

VIME: Variational Information Maximizing Exploration

2016-05-31 · NeurIPS 2016 12 · Rein Houthooft, Xi Chen, Yan Duan, John Schulman, Filip De Turck, Pieter Abbeel

Scalable and effective exploration remains a key challenge in reinforcement learning (RL). While there are methods with optimality guarantees in the setting of discrete state and action spaces, these methods cannot be applied in high-dimensional deep RL scenarios. As such, most contemporary RL relies on simple heuristics such as epsilon-greedy exploration or adding Gaussian noise to the controls. This paper introduces Variational Information Maximizing Exploration (VIME), an exploration strategy based on maximization of information gain about the agent's belief of environment dynamics. We propose a practical implementation, using variational inference in Bayesian neural networks which efficiently handles continuous state and action spaces. VIME modifies the MDP reward function, and can be applied with several different underlying RL algorithms. We demonstrate that VIME achieves significantly better performance compared to heuristic exploration methods across a variety of continuous control tasks and algorithms, including tasks with very sparse rewards.

📄 PDF Abstract BibTeX arXiv:1605.09674

Code (2)

alec-tschantz/vime pytorch
openai/vime

Tasks

continuous-controlContinuous ControlReinforcement LearningReinforcement Learning (RL)Variational Inference

Similar Papers 제목 키워드 기반

Models of Vimentin Organization Under Actin-Driven Transport

2022-09-06 · Youngmin Park, Cécile Leduc, Sandrine Etienne-Manneville, Stéphanie Portet

Intermediate filaments form an essential structural network, spread throughout the cytoplasm and play a key role in cell mechanics, intracellular organization and molecular signaling. The maintenance of the network and i…

2D Basement Relief Inversion using Sparse Regularization

2024-10-19 · Francisco Márcio Barboza, Arthur Anthony da Cunha Romão E Silva, Bruno Motta de Carvalho

Basement relief gravimetry is crucial in geophysics, especially for oil exploration and mineral prospecting. It involves solving an inverse problem to infer geological model parameters from observed data. The model repre…

Geophysics

Machine Learning Algorithms for Transplanting Accelerometer Observations in Future Satellite Gravimetry Missions

2025-08-05 · Mohsen Romeshkani, Jürgen Müller, Sahar Ebadi, Alexey Kupriyanov 외 arxiv

Accurate and continuous monitoring of Earth's gravity field is essential for tracking mass redistribution processes linked to climate variability, hydrological cycles, and geodynamic phenomena. While the GRACE and GRACE …

EviMem: Evidence-Gap-Driven Iterative Retrieval for Long-Term Conversational Memory

2026-04-30 · Yuyang Li, Yime He, Zeyu Zhang, Dong Gong arxiv

Long-term conversational memory requires retrieving evidence scattered across multiple sessions, yet single-pass retrieval fails on temporal and multi-hop questions. Existing iterative methods refine queries via generate…

Information Gain Is Not All You Need

2025-03-28 · Ludvig Ericson, José Pedro, Patric Jensfelt

Autonomous exploration in mobile robotics often involves a trade-off between two objectives: maximizing environmental coverage and minimizing the total path length. In the widely used information gain paradigm, explorati…

AllCommon Sense Reasoning