Linear-Quadratic Zero-Sum Mean-Field Type Games: Optimality Conditions and Policy Optimization
In this paper, zero-sum mean-field type games (ZSMFTG) with linear dynamics and quadratic cost are studied under infinite-horizon discounted utility function. ZSMFTG are a class of games in which two decision makers whose utilities sum to zero, compete to influence a large population of indistinguishable agents. In particular, the case in which the transition and utility functions depend on the state, the action of the controllers, and the mean of the state and the actions, is investigated. The optimality conditions of the game are analysed for both open-loop and closed-loop controls, and explicit expressions for the Nash equilibrium strategies are derived. Moreover, two policy optimization methods that rely on policy gradient are proposed for both model-based and sample-based frameworks. In the model-based case, the gradients are computed exactly using the model, whereas they are estimated using Monte-Carlo simulations in the sample-based case. Numerical experiments are conducted to show the convergence of the utility function as well as the two players' controls.
Code (0)
등록된 구현이 없습니다.
Similar Papers 제목 키워드 기반
Policy Optimization for Linear-Quadratic Zero-Sum Mean-Field Type Games
In this paper, zero-sum mean-field type games (ZSMFTG) with linear dynamics and quadratic utility are studied under infinite-horizon discounted utility function. ZSMFTG are a class of games in which two decision makers w…
Vocal Bursts Type PredictionAn Addendum to the Problem of Zero-Sum LQ Stochastic Mean-Field Dynamic Games\\ (Extended version)
In this paper, we first address a linear quadratic mean-field game problem with a leader-follower structure. By adopting a Riccati-type approach, we show how one can obtain a state-feedback representation of the pairs of…
Stability Via Adversarial Training of Neural Network Stochastic Control of Mean-Field Type
In this paper, we present an approach to neural network mean-field-type control and its stochastic stability analysis by means of adversarial inputs (aka adversarial attacks). This is a class of data-driven mean-field-ty…
Vocal Bursts Type PredictionThompson sampling for linear quadratic mean-field teams
We consider optimal control of an unknown multi-agent linear quadratic (LQ) system where the dynamics and the cost are coupled across the agents through the mean-field (i.e., empirical mean) of the states and controls. D…
Thompson SamplingLQG Graphon Mean Field Games: Analysis via Graphon Invariant Subspaces
This paper studies approximate solutions to large-scale linear quadratic stochastic games with homogeneous nodal dynamics parameters and heterogeneous network couplings within the graphon mean field game framework in [2]…