A time-optimal feedback control for a particular case of the game of two cars
In this paper, a time-optimal feedback solution to the game of two cars, for the case where the pursuer is faster and more agile than the evader, is presented. The concept of continuous subsets of the reachable set is introduced to characterize the time-optimal pursuit-evasion game under feedback strategies. Using these subsets it is shown that, if initially the pursuer is distant enough from the evader, then the feedback saddle point strategies for both the pursuer and the evader are coincident with one of the common tangents from the minimum radius turning circles of the pursuer to the minimum radius turning circles of the evader. Using geometry, four feasible tangents are identified and the feedback min-max strategy for the pursuer and the max-min strategy for the evader are derived by solving a $2 \times 2$ matrix game at each instant. Insignificant computational effort is involved in evaluating the pursuer and evader inputs using the proposed feedback control law and hence it is suitable for real-time implementation. The proposed law is validated further by comparing the resulting trajectories with those obtained by solving the differential game using numerical techniques.
Code (0)
등록된 구현이 없습니다.
Similar Papers 제목 키워드 기반
Optimal Output Feedback Learning Control for Discrete-Time Linear Quadratic Regulation
This paper studies the linear quadratic regulation (LQR) problem of unknown discrete-time systems via dynamic output feedback learning control. In contrast to the state feedback, the optimality of the dynamic output feed…
Machine Learning based Optimal Feedback Control for Microgrid Stabilization
Microgrids have more operational flexibilities as well as uncertainties than conventional power grids, especially when renewable energy resources are utilized. An energy storage based feedback controller can compensate u…
BIG-bench Machine LearningLearning the optimal state-feedback via supervised imitation learning
Imitation learning is a control design paradigm that seeks to learn a control policy reproducing demonstrations from expert agents. By substituting expert demonstrations for optimal behaviours, the same paradigm leads to…
Imitation LearningDiscrete-time Flatness-based Control of a Gantry Crane
This article addresses the design of a discrete-time flatness-based tracking control for a gantry crane and demonstrates the practical applicability of the approach by measurement results. The required sampled-data model…
Learning Decentralized Linear Quadratic Regulators with $\sqrt{T}$ Regret
We propose an online learning algorithm that adaptively designs a decentralized linear quadratic regulator when the system model is unknown a priori and new data samples from a single system trajectory become progressive…