paper-with-me

Papers

Bayes' Bluff: Opponent Modelling in Poker

2012-07-04 · Finnegan Southey, Michael P. Bowling, Bryce Larson, Carmelo Piccione, Neil Burch, Darse Billings, Chris Rayner

Poker is a challenging problem for artificial intelligence, with non-deterministic dynamics, partial observability, and the added difficulty of unknown adversaries. Modelling all of the uncertainties in this domain is not an easy task. In this paper we present a Bayesian probabilistic model for a broad class of poker games, separating the uncertainty in the game dynamics from the uncertainty of the opponent's strategy. We then describe approaches to two key subproblems: (i) inferring a posterior over opponent strategies given a prior distribution and observations of their play, and (ii) playing an appropriate response to that distribution. We demonstrate the overall approach on a reduced version of poker using Dirichlet priors and then on the full game of Texas hold'em using a more informed prior. We demonstrate methods for playing effective responses to the opponent, based on the posterior.

📄 PDF Abstract BibTeX arXiv:1207.1411

Code (1)

sotetsuk/pgx jax

Similar Papers 제목 키워드 기반

Analysis of Bluffing by DQN and CFR in Leduc Hold'em Poker

2025-09-04 · Tarik Zaciragic, Aske Plaat, K. Joost Batenburg arxiv

In the game of poker, being unpredictable, or bluffing, is an essential skill. When humans play poker, they bluff. However, most works on computer-poker focus on performance metrics such as win rates, while bluffing is o…

Reinforcement Learning

Beyond Game Theory Optimal: Profit-Maximizing Poker Agents for No-Limit Holdem

2025-09-28 · SeungHyun Yi, Seungjun Yi arxiv

Game theory has grown into a major field over the past few decades, and poker has long served as one of its key case studies. Game-Theory-Optimal (GTO) provides strategies to avoid loss in poker, but pure GTO does not gu…

AlphaExploitem: Going Beyond the Nash Equilibrium in Poker by Learning to Exploit Suboptimal Play

2026-05-09 · Vlad Murgoci, Matthijs Spaan, Yaniv Oren arxiv

Poker is an imperfect information game that has served as a long-standing benchmark for decision-making under uncertainty. To maximize utility beyond the Nash equilibrium, an agent can deviate from Nash-equilibrium polic…

On self-play computation of equilibrium in poker

2018-05-23 · Mikhail Goykhman

We compare performance of the genetic algorithm and the counterfactual regret minimization algorithm in computing the near-equilibrium strategies in the simplified poker games. We focus on the von Neumann poker and the s…

counterfactual

Outbidding and Outbluffing Elite Humans: Mastering Liar's Poker via Self-Play and Reinforcement Learning

2025-11-05 · Richard Dewey, Janos Botyanszki, Ciamac C. Moallemi, Andrew T. Zheng arxiv

AI researchers have long focused on poker-like games as a testbed for environments characterized by multi-player dynamics, imperfect information, and reasoning under uncertainty. While recent breakthroughs have matched e…

Reinforcement Learning