paper-with-me

홈 › Papers

Simulated Annealing Enhances Theory-of-Mind Reasoning in Autoregressive Language Models

2026-01-18 · Xucong Hu, Jian-Qiao Zhu arxiv

Autoregressive language models are next-token predictors and have been criticized for only optimizing surface plausibility (i.e., local coherence) rather than maintaining correct latent-state representations (i.e., global coherence). Because Theory of Mind (ToM) tasks crucially depend on reasoning about latent mental states of oneself and others, such models are therefore often thought to fail at ToM. While post-training methods can improve ToM performance, we show that strong ToM capability can be recovered directly from the base model without any additional weight updates or verifications. Our approach builds on recent power-sampling methods (Karan & Du, 2025) that use Markov chain Monte Carlo (MCMC) to sample from sharpened sequence-level (rather than token-level) probability distributions of autoregressive language models. We further find that incorporating annealing, where the tempered distribution is gradually shifted from high to low temperature, substantially improves ToM performance over fixed-temperature power sampling. Together, these results suggest that sampling-based optimization provides a powerful way to extract latent capabilities from language models without retraining.

📄 PDF Abstract BibTeX arXiv:2601.12269

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Minding Language Models' (Lack of) Theory of Mind: A Plug-and-Play Multi-Character Belief Tracker

2023-06-01 · Melanie Sclar, Sachin Kumar, Peter West, Alane Suhr 외

Theory of Mind (ToM)$\unicode{x2014}$the ability to reason about the mental states of other people$\unicode{x2014}$is a key element of our social intelligence. Yet, despite their ever more impressive performance, large-s…

Reading Comprehension

Annealed Multiple Choice Learning: Overcoming limitations of Winner-takes-all with annealing

2024-07-22 · David Perera, Victor Letzelter, Théo Mariotte, Adrien Cortés 외

We introduce Annealed Multiple Choice Learning (aMCL) which combines simulated annealing with MCL. MCL is a learning framework handling ambiguous tasks by predicting a small set of plausible hypotheses. These hypotheses …

AllDiversityMultiple-choiceSpeech Separation

Search Algorithms for Mastermind

2019-08-16 · Anthony D. Rhodes

his paper presents two novel approaches to solving the classic board game mastermind, including a variant of simulated annealing (SA) and a technique we term maximum expected reduction in consistency (MERC). In addition,…

Shedding some light on Light Up with Artificial Intelligence

2021-07-22 · Libo Sun, James Browning, Roberto Perera

The Light-Up puzzle, also known as the AKARI puzzle, has never been solved using modern artificial intelligence (AI) methods. Currently, the most widely used computational technique to autonomously develop solutions invo…

DPMT: Dual Process Multi-scale Theory of Mind Framework for Real-time Human-AI Collaboration

2025-07-18 · Xiyun Li, Yining Ding, Yuhua Jiang, Yunlong Zhao 외 arxiv

Real-time human-artificial intelligence (AI) collaboration is crucial yet challenging, especially when AI agents must adapt to diverse and unseen human behaviors in dynamic scenarios. Existing large language model (LLM) …