paper-with-me

홈 › Papers

Measuring the metacognition of AI

2026-03-31 · Richard Servajean, Philippe Servajean arxiv

A robust decision-making process must take into account uncertainty, especially when the choice involves inherent risks. Because artificial intelligence (AI) systems are increasingly integrated into decision-making workflows, managing uncertainty relies more and more on the metacognitive capabilities of these systems; i.e, their ability to assess the reliability of and regulate their own decisions. Hence, it is crucial to employ robust methods to measure the metacognitive abilities of AI. This paper is primarily a methodological contribution arguing for the adoption of the meta-d' framework as the gold standard for assessing the metacognitive sensitivity of AIs--the ability to generate confidence ratings that distinguish correct from incorrect responses. Moreover, we propose to leverage signal detection theory (SDT) to measure the ability of AIs to spontaneously regulate their decisions based on uncertainty and risk. To demonstrate the practical utility of these psychophysical frameworks, we conduct two series of experiments on three large language models (LLMs)--GPT-5, DeepSeek-V3.2-Exp, and Mistral-Medium-2508.

📄 PDF Abstract BibTeX arXiv:2603.29693

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

MEDLEY-BENCH: Scale Buys Evaluation but Not Control in AI Metacognition

2026-04-17 · Farhad Abtahi, Abdolamir Karbalaie, Eduardo Illueca-Fernandez, Fernando Seoane arxiv

Metacognition, the ability to monitor and regulate one's own reasoning, remains under-evaluated in AI benchmarking. We introduce MEDLEY-BENCH, a benchmark of behavioural metacognition that separates independent reasoning…

Evidence for Limited Metacognition in LLMs

2025-09-25 · Christopher Ackerman arxiv

The possibility of LLM self-awareness and even sentience is gaining increasing public attention and has major safety and policy implications, but the science of measuring them is still in a nascent state. Here we introdu…

Learning a metacognition for object perception

2020-11-30 · NeurIPS Workshop SVRHM 2020 12 · Marlene Berke, Mario Belledonne, Julian Jara-Ettinger

Beyond representing the external world, humans also represent their own cognitive processes. In the context of perception, this metacognition helps us identify unreliable percepts, such as when we recognize that we are s…

Object

Computational Metacognition

2022-01-30 · Michael Cox, Zahiduddin Mohammad, Sravya Kondrakunta, Ventaksamapth Raja Gogineni 외

Computational metacognition represents a cognitive systems perspective on high-order reasoning in integrated artificial systems that seeks to leverage ideas from human metacognition and from metareasoning approaches in a…

A Proposal to Extend the Common Model of Cognition with Metacognition

2025-06-09 · John Laird, Christian Lebiere, Paul Rosenbloom, Andrea Stocco

The Common Model of Cognition (CMC) provides an abstract characterization of the structure and processing required by a cognitive architecture for human-like minds. We propose a unified approach to integrating metacognit…