paper-with-me

홈 › Papers

Learning Nash Equilibrial Hamiltonian for Two-Player Collision-Avoiding Interactions

2025-03-10 · Lei Zhang, Siddharth Das, Tanner Merry, Wenlong Zhang, Yi Ren

We consider the problem of learning Nash equilibrial policies for two-player risk-sensitive collision-avoiding interactions. Solving the Hamilton-Jacobi-Isaacs equations of such general-sum differential games in real time is an open challenge due to the discontinuity of equilibrium values on the state space. A common solution is to learn a neural network that approximates the equilibrium Hamiltonian for given system states and actions. The learning, however, is usually supervised and requires a large amount of sample equilibrium policies from different initial states in order to mitigate the risks of collisions. This paper claims two contributions towards more data-efficient learning of equilibrium policies: First, instead of computing Hamiltonian through a value network, we show that the equilibrium co-states have simple structures when collision avoidance dominates the agents' loss functions and system dynamics is linear, and therefore are more data-efficient to learn. Second, we introduce theory-driven active learning to guide data sampling, where the acquisition function measures the compliance of the predicted co-states to Pontryagin's Maximum Principle. On an uncontrolled intersection case, the proposed method leads to more generalizable approximation of the equilibrium policies, and in turn, lower collision probabilities, than the state-of-the-art under the same data acquisition budget.

📄 PDF Abstract BibTeX arXiv:2503.07013

Code (0)

등록된 구현이 없습니다.

Tasks

Active LearningCollision Avoidance

Similar Papers 제목 키워드 기반

Approximating Discontinuous Nash Equilibrial Values of Two-Player General-Sum Differential Games

2022-07-05 · Lei Zhang, Mukesh Ghimire, Wenlong Zhang, Zhe Xu 외

Finding Nash equilibrial policies for two-player differential games requires solving Hamilton-Jacobi-Isaacs (HJI) PDEs. Self-supervised learning has been used to approximate solutions of such PDEs while circumventing the…

Autonomous DrivingSelf-Supervised Learning

Nash Equilibria of Static Prediction Games

2009-12-01 · NeurIPS 2009 12 · Michael Brückner, Tobias Scheffer

The standard assumption of identically distributed training and test data can be violated when an adversary can exercise some control over the generation of the test data. In a prediction game, a learner produces a predi…

Prediction

Markov Potential Game with Final-time Reach-Avoid Objectives

2024-10-23 · Sarah H. Q. Li, Abraham P. Vinod

We formulate a Markov potential game with final-time reach-avoid objectives by integrating potential game theory with stochastic reach-avoid control. Our focus is on multi-player trajectory planning where players maximiz…

Motion PlanningTrajectory Planning

Multiplayer bandits without observing collision information

2018-08-25 · Gabor Lugosi, Abbas Mehrabian

We study multiplayer stochastic multi-armed bandit problems in which the players cannot communicate and if two or more players pull the same arm, a collision occurs and the involved players receive zero reward. We consid…

Distributed Nash Equilibrium Seeking with Stochastic Event-Triggered Mechanism

2023-04-20 · Wei Huo, Kam Fai Elvis Tsang, Yamin Yan, Karl Henrik Johansson 외

In this paper, we study the problem of consensus-based distributed Nash equilibrium (NE) seeking where a network of players, abstracted as a directed graph, aim to minimize their own local cost functions non-cooperativel…