paper-with-me

홈 › Papers

Evidence for Limited Metacognition in LLMs

2025-09-25 · Christopher Ackerman arxiv

The possibility of LLM self-awareness and even sentience is gaining increasing public attention and has major safety and policy implications, but the science of measuring them is still in a nascent state. Here we introduce a novel methodology for quantitatively evaluating metacognitive abilities in LLMs. Taking inspiration from research on metacognition in nonhuman animals, our approach eschews model self-reports and instead tests to what degree models can strategically deploy knowledge of internal states. Using two experimental paradigms, we demonstrate that frontier LLMs introduced since early 2024 show increasingly strong evidence of certain metacognitive abilities, specifically the ability to assess and utilize their own confidence in their ability to answer factual and reasoning questions correctly and the ability to anticipate what answers they would give and utilize that information appropriately. We buttress these behavioral findings with an analysis of the token probabilities returned by the models, which suggests the presence of an upstream internal signal that could provide the basis for metacognition. We further find that these abilities 1) are limited in resolution, 2) emerge in context-dependent manners, and 3) seem to be qualitatively different from those of humans. We also report intriguing differences across models of similar capabilities, suggesting that LLM post-training may have a role in developing metacognitive abilities.

📄 PDF Abstract BibTeX arXiv:2509.21545

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Metacognition is all you need? Using Introspection in Generative Agents to Improve Goal-directed Behavior

2024-01-09 · Jason Toy, Josh MacAdam, Phil Tabor

Recent advances in Large Language Models (LLMs) have shown impressive capabilities in various applications, yet LLMs face challenges such as limited context windows and difficulties in generalization. In this paper, we i…

All

Metacognition in LLMs: Foundations, Progress, and Opportunities

2026-07-13 · Gabrielle Kaili-May Liu, Areeb Gani, Jacqueline Lu, Jordan Thomas 외 hf

Metacognition is a foundational component of intelligence critical to effective learning, problem solving, decision-making, communication, and more. In recent years, it has become increasingly recognized as a cornerstone…

Metacognition and Uncertainty Communication in Humans and Large Language Models

2025-04-18 · Mark Steyvers, Megan A. K. Peters

Metacognition, the capacity to monitor and evaluate one's own knowledge and performance, is foundational to human decision-making, learning, and communication. As large language models (LLMs) become increasingly embedded…

Decision Making

Confidence is detection-like in high-dimensional spaces

2024-10-24 · Wiktoria Łuczak, Kevin O'Neill, Stephen M. Fleming

Confidence estimates are often "detection-like" - driven by positive evidence in favour of a decision. This empirical observation has been interpreted as showing human metacognition is limited by biases or heuristics. He…

Sensitivity

LLMs Show No Signs Of Individuated Metacognition

2026-05-22 · M. Moran, Mark Whiting arxiv

Confidence-weighted routing, selective abstention, and ensemble weighting all assume that a model's stated confidence is informative about its capability on the question being asked. They presume functional metacognition…

Mathematical ReasoningInformation Retrieval