paper-with-me

홈 › Papers

Equilibrium Residuals Expose Three Regimes of Matrix-Game Strategic Reasoning in Language Models

2026-05-11 · Wenhua Nie, Binhan Luo, Zijie Meng, Jyh-Shing Roger Jang, Ching-Wen Ma arxiv

Large language models can score well on named game-theory benchmarks while failing on the same strategic computation once semantic cues are removed. We show this gap with procedurally generated zero-sum matrix games: a model that recognizes familiar games drops to 34%, 18%, and 2% success on anonymous $2{\times}2$, $3{\times}3$, and $5{\times}5$ payoff matrices. The benchmark separates semantic recall, learned approximate Nash computation, and an output-interface bottleneck that limits scale. Training only on $2{\times}2$ and $3{\times}3$ games, supervised fine-tuning raises unseen $5{\times}5$--$7{\times}7$ success from 2% to 61%, while exploitability-reward training averages 37% with high seed variance. We prove that the exploitability residual is $2$-Lipschitz in payoff perturbations, unlike discontinuous vertex-returning LP equilibrium selectors, explaining why residual training can transfer under payoff shifts even when formatting instability limits mean performance. A dominated-action padding experiment provides causal evidence: trained models solve $3{\times}3$ games embedded in much larger matrices, while random-padded controls fail and dense $12{\times}12$ games remain near failure. Procedural evaluation is therefore necessary for measuring strategic reasoning, and residual rewards expose a real but format-limited route to approximate equilibrium computation.

📄 PDF Abstract BibTeX arXiv:2605.10410

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

PTL-PINNs: Perturbation-Guided Transfer Learning with Physics- Informed Neural Networks for Nonlinear Systems

2026-01-17 · Duarte Alexandrino, Ben Moseley, Pavlos Protopapas arxiv

Accurately and efficiently solving nonlinear differential equations is crucial for modeling dynamic behavior across science and engineering. Physics-Informed Neural Networks (PINNs) have emerged as a powerful solution th…

Transfer Learning

Capturing reduced-order quantum many-body dynamics out of equilibrium via neural ordinary differential equations

2025-12-15 · Patrick Egenlauf, Iva Březinová, Sabine Andergassen, Miriam Klopotek arxiv

Out-of-equilibrium quantum many-body systems exhibit rapid correlation buildup that underlies many emerging phenomena. Exact wave-function methods to describe this scale exponentially with particle number; simpler mean-f…

Dimensionality Reduction

Equilibrium and non-Equilibrium regimes in the learning of Restricted Boltzmann Machines

2021-05-28 · NeurIPS 2021 12 · Aurélien Decelle, Cyril Furtlehner, Beatriz Seoane

Training Restricted Boltzmann Machines (RBMs) has been challenging for a long time due to the difficulty of computing precisely the log-likelihood gradient. Over the past decades, many works have proposed more or less su…

Inferring Synaptic Structure in presence of Neural Interaction Time Scales

2015-02-18

Biological networks display a variety of activity patterns reflecting a web of interactions that is complex both in space and time. Yet inference methods have mainly focused on reconstructing, from the network's activity…

Human-AI Co-Evolution and Epistemic Collapse: A Dynamical Systems Perspective

2026-05-07 · Xuening Wu, Yanlan Kang, Qianya Xu, Kexuan Xie 외 arxiv

Large language models (LLMs) are reshaping how knowledge is produced, with increasing reliance on AI systems for generation, summarization, and reasoning. While prior work has studied cognitive offloading in humans and m…