paper-with-me

Papers

Bayesian Algorithms Learn to Stabilize Unknown Continuous-Time Systems

2021-12-30 · Mohamad Kazem Shirani Faradonbeh, Mohamad Sadegh Shirani Faradonbeh

Linear dynamical systems are canonical models for learning-based control of plants with uncertain dynamics. The setting consists of a stochastic differential equation that captures the state evolution of the plant understudy, while the true dynamics matrices are unknown and need to be learned from the observed data of state trajectory. An important issue is to ensure that the system is stabilized and destabilizing control actions due to model uncertainties are precluded as soon as possible. A reliable stabilization procedure for this purpose that can effectively learn from unstable data to stabilize the system in a finite time is not currently available. In this work, we propose a novel Bayesian learning algorithm that stabilizes unknown continuous-time stochastic linear systems. The presented algorithm is flexible and exposes effective stabilization performance after a remarkably short time period of interacting with the system.

📄 PDF Abstract BibTeX arXiv:2112.15094

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Stochastic Bayesian Optimization with Unknown Continuous Context Distribution via Kernel Density Estimation

2023-12-16 · Xiaobin Huang, Lei Song, Ke Xue, Chao Qian

Bayesian optimization (BO) is a sample-efficient method and has been widely used for optimizing expensive black-box functions. Recently, there has been a considerable interest in BO literature in optimizing functions tha…

Bayesian OptimizationDensity Estimation

A Lipschitz Exploration-Exploitation Scheme for Bayesian Optimization

2012-03-30 · Ali Jalali, Javad Azimi, Xiaoli Fern, Ruofei Zhang

The problem of optimizing unknown costly-to-evaluate functions has been studied for a long time in the context of Bayesian Optimization. Algorithms in this field aim to find the optimizer of the function by asking only a…

Bayesian Optimization

Surveillance Evasion Through Bayesian Reinforcement Learning

2021-09-30 · Dongping Qi, David Bindel, Alexander Vladimirsky

We consider a task of surveillance-evading path-planning in a continuous setting. An Evader strives to escape from a 2D domain while minimizing the risk of detection (and immediate capture). The probability of detection …

regressionreinforcement-learningReinforcement LearningReinforcement Learning (RL)

Learning and Testing Causal Models with Interventions

2018-05-24 · NeurIPS 2018 12 · Jayadev Acharya, Arnab Bhattacharyya, Constantinos Daskalakis, Saravanan Kandasamy

We consider testing and learning problems on causal Bayesian networks as defined by Pearl (Pearl, 2009). Given a causal Bayesian network $\mathcal{M}$ on a graph with $n$ discrete variables and bounded in-degree and boun…

Convergence of Expectation-Maximization Algorithm with Mixed-Integer Optimization

2024-01-31 · Geethu Joseph

The convergence of expectation-maximization (EM)-based algorithms typically requires continuity of the likelihood function with respect to all the unknown parameters (optimization variables). The requirement is not met w…