paper-with-me

홈 › Papers

Dynamic Control of Explore/Exploit Trade-Off In Bayesian Optimization

2018-07-03 · Dipti Jasrasaria, Edward O. Pyzer-Knapp

Bayesian optimization offers the possibility of optimizing black-box operations not accessible through traditional techniques. The success of Bayesian optimization methods such as Expected Improvement (EI) are significantly affected by the degree of trade-off between exploration and exploitation. Too much exploration can lead to inefficient optimization protocols, whilst too much exploitation leaves the protocol open to strong initial biases, and a high chance of getting stuck in a local minimum. Typically, a constant margin is used to control this trade-off, which results in yet another hyper-parameter to be optimized. We propose contextual improvement as a simple, yet effective heuristic to counter this - achieving a one-shot optimization strategy. Our proposed heuristic can be swiftly calculated and improves both the speed and robustness of discovery of optimal solutions. We demonstrate its effectiveness on both synthetic and real world problems and explore the unaccounted for uncertainty in the pre-determination of search hyperparameters controlling explore-exploit trade-off.

📄 PDF Abstract BibTeX arXiv:1807.01279

Code (0)

등록된 구현이 없습니다.

Tasks

Bayesian Optimization

Methods 이 논문이 사용한 방법론

SPEED The monocular depth estimation (MDE) is the task of estimating depth from a single frame. This information is an essential knowledge in many computer vision tasks such as scene…

Similar Papers 제목 키워드 기반

Dynamic Exploration-Exploitation Trade-Off in Active Learning Regression with Bayesian Hierarchical Modeling

2023-04-16 · Upala Junaida Islam, Kamran Paynabar, George Runger, Ashif Sikandar Iquebal

Active learning provides a framework to adaptively query the most informative experiments towards learning an unknown black-box function. Various approaches of active learning have been proposed in the literature, howeve…

Active Learningregression

Dual Control for Approximate Bayesian Reinforcement Learning

2015-10-13 · Edgar D. Klenske, Philipp Hennig

Control of non-episodic, finite-horizon dynamical systems with uncertain dynamics poses a tough and elementary case of the exploration-exploitation trade-off. Bayesian reinforcement learning, reasoning about the effect o…

regressionreinforcement-learningReinforcement LearningReinforcement Learning (RL)

Expected Diverse Utility (EDU): Diverse Bayesian Optimization of Expensive Computer Simulators

2024-10-02 · John Joshua Miller, Simon Mak, Benny Sun, Sai Ranjeet Narayanan 외

The optimization of expensive black-box simulators arises in a myriad of modern scientific and engineering applications. Bayesian optimization provides an appealing solution, by leveraging a fitted surrogate model to gui…

Bayesian OptimizationDiversity

Information Maximizing Exploration with a Latent Dynamics Model

2018-04-04 · Trevor Barron, Oliver Obst, Heni Ben Amor

All reinforcement learning algorithms must handle the trade-off between exploration and exploitation. Many state-of-the-art deep reinforcement learning methods use noise in the action selection, such as Gaussian noise in…

continuous-controlContinuous ControlDeep Reinforcement Learningmodel+5

Multi-Agent LLMs for Adaptive Acquisition in Bayesian Optimization

2026-03-30 · Andrea Carbonati, Mohammadsina Almasi, Hadis Anahideh arxiv

The exploration-exploitation trade-off is central to sequential decision-making and black-box optimization, yet how Large Language Models (LLMs) reason about and manage this trade-off remains poorly understood. Unlike Ba…