Optimizing Closed-Loop Performance with Data from Similar Systems: A Bayesian Meta-Learning Approach
Bayesian optimization (BO) has demonstrated potential for optimizing control performance in data-limited settings, especially for systems with unknown dynamics or unmodeled performance objectives. The BO algorithm efficiently trades-off exploration and exploitation by leveraging uncertainty estimates using surrogate models. These surrogates are usually learned using data collected from the target dynamical system to be optimized. Intuitively, the convergence rate of BO is better for surrogate models that can accurately predict the target system performance. In classical BO, initial surrogate models are constructed using very limited data points, and therefore rarely yield accurate predictions of system performance. In this paper, we propose the use of meta-learning to generate an initial surrogate model based on data collected from performance optimization tasks performed on a variety of systems that are different to the target system. To this end, we employ deep kernel networks (DKNs) which are simple to train and which comprise encoded Gaussian process models that integrate seamlessly with classical BO. The effectiveness of our proposed DKN-BO approach for speeding up control system performance optimization is demonstrated using a well-studied nonlinear system with unknown dynamics and an unmodeled performance function.
Code (0)
등록된 구현이 없습니다.
Tasks
Bayesian OptimizationMeta-LearningMethods 이 논문이 사용한 방법론
Similar Papers 제목 키워드 기반
Stability-informed Bayesian Optimization for MPC Cost Function Learning
Designing predictive controllers towards optimal closed-loop performance while maintaining safety and stability is challenging. This work explores closed-loop learning for predictive control parameters under imperfect in…
Bayesian Optimizationglobal-optimizationQuasi-static closed-loop wind-farm control for combined power and fatigue optimization
We develop a methodology for combined power and loads optimization by coupling a surrogate loads model with an analytical quasi-static Gaussian wake merging model. The look-up table based fatigue model is developed offli…
COMO: Closed-Loop Optical Molecule Recognition with Minimum Risk Training
Optical chemical structure recognition (OCSR) translates molecular images into machine-readable representations like SMILES strings or molecular graphs, but remains challenging in real-world documents due to inexhaustibl…
Multi-agent reinforcement learning for intent-based service assurance in cellular networks
Recently, intent-based management has received good attention in telecom networks owing to stringent performance requirements for many of the use cases. Several approaches in the literature employ traditional closed-loop…
ManagementMulti-agent Reinforcement Learningreinforcement-learningReinforcement Learning (RL)Learning Model Predictive Control Parameters via Bayesian Optimization for Battery Fast Charging
Tuning parameters in model predictive control (MPC) presents significant challenges, particularly when there is a notable discrepancy between the controller's predictions and the actual behavior of the closed-loop plant.…
Bayesian OptimizationModel Predictive Control