paper-with-me

Papers

Beyond Outcome-Based Imperfect-Recall: Higher-Resolution Abstractions for Imperfect-Information Games

2025-10-16 · Yanchang Fu, Qiyue Yin, Shengda Liu, Pei Xu, Kaiqi Huang arxiv

Hand abstraction is crucial for scaling imperfect-information games (IIGs) such as Texas Hold'em, yet progress is limited by the lack of a formal task model and by evaluations that require resource-intensive strategy solving. We introduce signal observation ordered games (SOOGs), a subclass of IIGs tailored to hold'em-style games that cleanly separates signal from player action sequences, providing a precise mathematical foundation for hand abstraction. Within this framework, we define a resolution bound-an information-theoretic upper bound on achievable performance under a given signal abstraction. Using the bound, we show that mainstream outcome-based imperfect-recall algorithms suffer substantial losses by arbitrarily discarding historical information; we formalize this behavior via potential-aware outcome Isomorphism (PAOI) and prove that PAOI characterizes their resolution bound. To overcome this limitation, we propose full-recall outcome isomorphism (FROI), which integrates historical information to raise the bound and improve policy quality. Experiments on hold'em-style benchmarks confirm that FROI consistently outperforms outcome-based imperfect-recall baselines. Our results provide a unified formal treatment of hand abstraction and practical guidance for designing higher-resolution abstractions in IIGs.

📄 PDF Abstract BibTeX arXiv:2510.15094

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Beyond Outcome Reward: Decoupling Search and Answering Improves LLM Agents

2025-10-06 · Yiding Wang, Zhepei Wei, Xinyu Zhu, Yu Meng arxiv

Enabling large language models (LLMs) to utilize search tools offers a promising path to overcoming fundamental limitations such as knowledge cutoffs and hallucinations. Recent work has explored reinforcement learning (R…

Reinforcement LearningAnswer Generation

Decision Making under Imperfect Recall: Algorithms and Benchmarks

2026-02-16 · Emanuel Tewolde, Brian Hu Zhang, Ioannis Anagnostides, Tuomas Sandholm 외 arxiv

In game theory, imperfect-recall decision problems model situations in which an agent forgets information it held before. They encompass games such as the ``absentminded driver'' and team games with limited communication…

Decision Making

KnowMe-Bench: Benchmarking Person Understanding for Lifelong Digital Companions

2026-01-08 · Tingyu Wu, Zhisheng Chen, Ziyan Weng, Shuhe Wang 외 arxiv

Existing long-horizon memory benchmarks mostly use multi-turn dialogues or synthetic user histories, which makes retrieval performance an imperfect proxy for person understanding. We present \BenchName, a publicly releas…

Adaptive User Pairing for Downlink NOMA System with Imperfect SIC

2020-12-13 · Nemalidinne Siva Mouni, Abhinav Kumar, Prabhat K. Upadhyay

Non-orthogonal multiple access (NOMA) has been recognized as a key driving technology for the fifth generation (5G) and beyond 5G cellular networks. For a practical dowlink NOMA system with imperfect successive interfere…

BeyondScene: Higher-Resolution Human-Centric Scene Generation With Pretrained Diffusion

2024-04-06 · Gwanghyun Kim, Hayeon Kim, Hoigi Seo, Dong Un Kang 외

Generating higher-resolution human-centric scenes with details and controls remains a challenge for existing text-to-image diffusion models. This challenge stems from limited training image size, text encoder capacity (l…

8kScene Generation