paper-with-me

홈 › Papers

Conjectural Online Learning with First-order Beliefs in Asymmetric Information Stochastic Games

2024-02-29 · Tao Li, Kim Hammar, Rolf Stadler, Quanyan Zhu

Asymmetric information stochastic games (AISGs) arise in many complex socio-technical systems, such as cyber-physical systems and IT infrastructures. Existing computational methods for AISGs are primarily offline and can not adapt to equilibrium deviations. Further, current methods are limited to particular information structures to avoid belief hierarchies. Considering these limitations, we propose conjectural online learning (COL), an online learning method under generic information structures in AISGs. COL uses a forecaster-actor-critic (FAC) architecture, where subjective forecasts are used to conjecture the opponents' strategies within a lookahead horizon, and Bayesian learning is used to calibrate the conjectures. To adapt strategies to nonstationary environments based on information feedback, COL uses online rollout with cost function approximation (actor-critic). We prove that the conjectures produced by COL are asymptotically consistent with the information feedback in the sense of a relaxed Bayesian consistency. We also prove that the empirical strategy profile induced by COL converges to the Berk-Nash equilibrium, a solution concept characterizing rationality under subjectivity. Experimental results from an intrusion response use case demonstrate COL's {faster convergence} over state-of-the-art reinforcement learning methods against nonstationary attacks.

📄 PDF Abstract BibTeX arXiv:2402.18781

Code (0)

등록된 구현이 없습니다.

Tasks

Decision Making

Similar Papers 제목 키워드 기반

From Knowledge to Conjectures: A Modal Framework for Reasoning about Hypotheses

2025-08-10 · Fabio Vitali arxiv

This paper introduces a new family of cognitive modal logics designed to formalize conjectural reasoning: modal systems in which cognitive contexts extend known facts with hypothetical assumptions in order to explore the…

Online Adaptive Asymmetric Active Learning with Limited Budgets

2019-11-18 · Yifan Zhang, Peilin Zhao, Shuaicheng Niu, Qingyao Wu 외

Online Active Learning (OAL) aims to manage unlabeled datastream by selectively querying the label of data. OAL is applicable to many real-world problems, such as anomaly detection in health-care and finance. In these pr…

Active LearningAnomaly Detection

LLM-Stackelberg Games: Conjectural Reasoning Equilibria and Their Applications to Spearphishing

2025-07-12 · Quanyan Zhu

We introduce the framework of LLM-Stackelberg games, a class of sequential decision-making models that integrate large language models (LLMs) into strategic interactions between a leader and a follower. Departing from cl…

Decision MakingMisinformationRecommendation SystemsSequential Decision Making

Mind the Perspective: Let's Reason Recursively for Theory of Mind

2026-06-10 · Chao Lei, Guang Hu, Meng Yang, Yanbei Jiang 외 arxiv

Theory of Mind (ToM) reasoning requires inferring agents' beliefs from partial and asymmetric observations, which remains an open challenge for LLMs. Existing prompting-based approaches improve ToM reasoning through obse…

Learning and Selfconfirming Equilibria in Network Games

2018-12-31 · Pierpaolo Battigalli, Fabrizio Panebianco, Paolo Pin

Consider a set of agents who play a network game repeatedly. Agents may not know the network. They may even be unaware that they are interacting with other agents in a network. Possibly, they just understand that their p…