paper-with-me

홈 › Papers

LLMs Show No Signs Of Individuated Metacognition

2026-05-22 · M. Moran, Mark Whiting arxiv

Confidence-weighted routing, selective abstention, and ensemble weighting all assume that a model's stated confidence is informative about its capability on the question being asked. They presume functional metacognition, the capacity to assess one's own capabilities, without exercising them. Aggregate calibration is well studied, with mixed results, but the underlying structure of elicited confidence is less well understood. We decompose binary confidence judgements from 20 frontier Large Language Models (LLMs) across six benchmarks using tetrachoric factor analysis paired with pairwise calibration, asking whether two models that differ in confidence also differ in performance. On factual recall and information retrieval benchmarks the cross-model confidence matrix is approximately rank-one and a single dominant factor captures most of the latent variance. Models retrieving facts share an item-level difficulty axis and differ mainly in their decision thresholds along it. Across all benchmarks the relationship between confidence and performance collapses once items that all models agree on are removed. Inter-model pairwise calibration is small even where statistically significant, and what remains shrinks to nothing once base-rate differences along the shared factor are controlled for. Mathematical reasoning is the apparent exception, but this turns out to be a confound where reasoning models answer questions about their confidence by trying to solve them in their chain of thought, bypassing the sub-symbolic self-knowledge we seek to measure. We find no evidence for significant verbalised individuated metacognition in any tested domain.

📄 PDF Abstract BibTeX arXiv:2605.24299

Code (0)

등록된 구현이 없습니다.

Tasks

Mathematical ReasoningInformation Retrieval

Similar Papers 제목 키워드 기반

Metacognition and Uncertainty Communication in Humans and Large Language Models

2025-04-18 · Mark Steyvers, Megan A. K. Peters

Metacognition, the capacity to monitor and evaluate one's own knowledge and performance, is foundational to human decision-making, learning, and communication. As large language models (LLMs) become increasingly embedded…

Decision Making

Metacognition in LLMs: Foundations, Progress, and Opportunities

2026-07-13 · Gabrielle Kaili-May Liu, Areeb Gani, Jacqueline Lu, Jordan Thomas 외 hf

Metacognition is a foundational component of intelligence critical to effective learning, problem solving, decision-making, communication, and more. In recent years, it has become increasingly recognized as a cornerstone…

Metacognition is all you need? Using Introspection in Generative Agents to Improve Goal-directed Behavior

2024-01-09 · Jason Toy, Josh MacAdam, Phil Tabor

Recent advances in Large Language Models (LLMs) have shown impressive capabilities in various applications, yet LLMs face challenges such as limited context windows and difficulties in generalization. In this paper, we i…

All

Metacognition as Reward: Reinforcing LLM Reasoning via Knowledge and Regulation Signals

2026-05-22 · Sirui Chen, Lei Xu, Yuying Zhao, Yutian Chen 외 arxiv

Recent RL methods have substantially improved the reasoning abilities of LLMs. Existing reward designs mainly follow two paradigms: (1) Reinforcement learning with verifiable rewards (RLVR) derives outcome signals from e…

Reinforcement Learning

Evidence for Limited Metacognition in LLMs

2025-09-25 · Christopher Ackerman arxiv

The possibility of LLM self-awareness and even sentience is gaining increasing public attention and has major safety and policy implications, but the science of measuring them is still in a nascent state. Here we introdu…