paper-with-me

홈 › Papers

Learning Distributed Equilibria in Linear-Quadratic Stochastic Differential Games: An $α$-Potential Approach

2026-02-18 · Philipp Plank, Yufei Zhang arxiv

We analyze independent policy-gradient (PG) learning in $N$-player linear-quadratic (LQ) stochastic differential games. Each player employs a distributed policy that depends only on its own state and updates the policy independently using the gradient of its own objective. We establish global linear convergence of these methods to an equilibrium by showing that the LQ game admits an $α$-potential structure, with $α$ determined by the degree of pairwise interaction asymmetry. For pairwise-symmetric interactions, we construct an affine distributed equilibrium by minimizing the potential function and show that independent PG methods converge globally to this equilibrium, with complexity scaling linearly in the population size and logarithmically in the desired accuracy. For asymmetric interactions, we prove that independent projected PG algorithms converge linearly to an approximate equilibrium, with suboptimality proportional to the degree of asymmetry. Numerical experiments confirm the theoretical results across both symmetric and asymmetric interaction networks.

📄 PDF Abstract BibTeX arXiv:2602.16555

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Time-Inconsistent Stochastic Linear--Quadratic Control: Characterization and Uniqueness of Equilibrium

2015-05-26

In this paper, we continue our study on a general time-inconsistent stochastic linear--quadratic (LQ) control problem originally formulated in [6]. We derive a necessary and sufficient condition for equilibrium controls …

Entropy-Regularized Reinforcement Learning for Linear-Quadratic Stackelberg Differential Games in Regime-Switching Diffusion Models

2026-06-27 · Congde Hu, Danping Li, Lin Xu, Wenying Xu arxiv

Stackelberg differential games (SDGs) provide a powerful framework for hierarchical decision-making in stochastic and continuous-time environments, yet their solution remains computationally challenging due to the comple…

Reinforcement Learning

Existence and uniqueness of quadratic and linear mean-variance equilibria in general semimartingale markets

2024-08-06 · Christoph Czichowsky, Martin Herdegen, David Martins

We revisit the classical topic of quadratic and linear mean-variance equilibria with both financial and real assets. The novelty of our results is that they are the first allowing for equilibrium prices driven by general…

Optimal Investment in a Large Population of Competitive and Heterogeneous Agents

2022-02-23 · Ludovic Tangpi, Xuchen Zhou

This paper studies a stochastic utility maximization game under relative performance concerns in finite agent and infinite agent settings, where a continuum of agents interact through a graphon (see definition below). We…

Finite-Time Error Bounds for Distributed Linear Stochastic Approximation

2021-11-24 · NeurIPS 2021 12 · Yixuan Lin, Vijay Gupta, Ji Liu

This paper considers a novel multi-agent linear stochastic approximation algorithm driven by Markovian noise and general consensus-type interaction, in which each agent evolves according to its local stochastic approxima…