Decentralized linear quadratic systems with major and minor agents and non-Gaussian noise
A decentralized linear quadratic system with a major agent and a collection of minor agents is considered. The major agent affects the minor agents, but not vice versa. The state of the major agent is observed by all agents. In addition, the minor agents have a noisy observation of their local state. The noise processes is \emph{not} assumed to be Gaussian. The structures of the optimal strategy and the best linear strategy are characterized. It is shown that major agent's optimal control action is a linear function of the major agent's MMSE (minimum mean squared error) estimate of the system state while the minor agent's optimal control action is a linear function of the major agent's MMSE estimate of the system state and a "correction term" which depends on the difference of the minor agent's MMSE estimate of its local state and the major agent's MMSE estimate of the minor agent's local state. Since the noise is non-Gaussian, the minor agent's MMSE estimate is a non-linear function of its observation. It is shown that replacing the minor agent's MMSE estimate by its LLMS (linear least mean square) estimate gives the best linear control strategy. The results are proved using a direct method based on conditional independence, common-information-based splitting of state and control actions, and simplifying the per-step cost based on conditional independence, orthogonality principle, and completion of squares.
Code (0)
등록된 구현이 없습니다.
Similar Papers 제목 키워드 기반
Guaranteed Stability Margins for Decentralized Linear Quadratic Regulators
It is well-known that linear quadratic regulators (LQR) enjoy guaranteed stability margins, whereas linear quadratic Gaussian regulators (LQG) do not. In this letter, we consider systems and compensators defined over dir…
On the Sample Complexity of Decentralized Linear Quadratic Regulator with Partially Nested Information Structure
We study the problem of control policy design for decentralized state-feedback linear quadratic control with a partially nested information structure, when the system model is unknown. We propose a model-based learning s…
Synthèse non quadratique H$\infty$ de contrôleurs décentralisés pour un ensemble de descripteurs flous T-S interconnectés
This paper deals with the non-quadratic decentralized stabilization of a set of n Takagi-Sugeno descriptors. To ensure the stability of the whole closed-loop dynamics and to minimize interconnection effects between subsy…
Positionality-Weighted Aggregation Methods for Cumulative Voting
Respecting minority opinions is vital in solving social problems. However, minority opinions are often ignored in general majority rules. To build consensus on pluralistic values and make social choices that consider min…
Distributed Reinforcement Learning for Decentralized Linear Quadratic Control: A Derivative-Free Policy Optimization Approach
This paper considers a distributed reinforcement learning problem for decentralized linear quadratic control with partial state observations and local costs. We propose a Zero-Order Distributed Policy Optimization algori…
Reinforcement LearningReinforcement Learning (RL)