paper-with-me

홈 › Papers

The Power of Perturbation under Sampling in Solving Extensive-Form Games

2025-01-28 · Wataru Masaka, Mitsuki Sakamoto, Kenshi Abe, Kaito Ariu, Tuomas Sandholm, Atsushi Iwasaki

This paper investigates how perturbation does and does not improve the Follow-the-Regularized-Leader (FTRL) algorithm in imperfect-information extensive-form games. Perturbing the expected payoffs guarantees that the FTRL dynamics reach an approximate equilibrium, and proper adjustments of the magnitude of the perturbation lead to a Nash equilibrium (\textit{last-iterate convergence}). This approach is robust even when payoffs are estimated using sampling -- as is the case for large games -- while the optimistic approach often becomes unstable. Building upon those insights, we first develop a general framework for perturbed FTRL algorithms under \textit{sampling}. We then empirically show that in the last-iterate sense, the perturbed FTRL consistently outperforms the non-perturbed FTRL. We further identify a divergence function that reduces the variance of the estimates for perturbed payoffs, with which it significantly outperforms the prior algorithms on Leduc poker (whose structure is more asymmetric in a sense than that of the other benchmark games) and consistently performs smooth convergence behavior on all the benchmark games.

📄 PDF Abstract BibTeX arXiv:2501.16600

Code (0)

등록된 구현이 없습니다.

Tasks

Form

Similar Papers 제목 키워드 기반

CCS: Controllable and Constrained Sampling with Diffusion Models via Initial Noise Perturbation

2025-02-07 · Bowen Song, Zecheng Zhang, ZhaoXu Luo, Jason Hu 외

Diffusion models have emerged as powerful tools for generative tasks, producing high-quality outputs across diverse domains. However, how the generated data responds to the initial noise perturbation in diffusion models …

Diversity

SURE Guided Posterior Sampling: Trajectory Correction for Diffusion-Based Inverse Problems

2025-12-29 · Minwoo Kim, Hongki Lim arxiv

Diffusion models have emerged as powerful learned priors for solving inverse problems. However, current iterative solving approaches which alternate between diffusion sampling and data consistency steps typically require…

Noise Estimation

On Sampling from the Gibbs Distribution with Random Maximum A-Posteriori Perturbations

2013-09-29 · NeurIPS 2013 12 · Tamir Hazan, Subhransu Maji, Tommi Jaakkola

In this paper we describe how MAP inference can be used to sample efficiently from Gibbs distributions. Specifically, we provide means for drawing either approximate or unbiased samples from Gibbs' distributions by intro…

A study of Thompson Sampling with Parameter h

2017-10-05 · Qiang Ha

Thompson Sampling algorithm is a well known Bayesian algorithm for solving stochastic multi-armed bandit. At each time step the algorithm chooses each arm with probability proportional to it being the current best arm. W…

Thompson Sampling

Active Learning of Spin Network Models

2019-03-25 · Jialong Jiang, David A. Sivak, Matt Thomson

The inverse statistical problem of finding direct interactions in complex networks is difficult. In the natural sciences, well-controlled perturbation experiments are widely used to probe the structure of complex network…

Active Learning