paper-with-me

Papers

Monitor-Generate-Verify (MGV): Formalising Metacognitive Theory for Language Model Reasoning

2025-11-06 · Nick Oh, Fernand Gobet arxiv

Test-time reasoning architectures such as those following the Generate-Verify paradigm, where a model iteratively refines or verifies its own generated outputs, prioritise generation and verification but exclude the monitoring processes that determine when and how reasoning should begin. This omission may contribute to the prefix dominance trap, in which models commit early to suboptimal reasoning paths and seldom recover, yielding roughly 20% accuracy loss. We address this architectural gap by proposing the Monitor-Generate-Verify (MGV) framework, a computational translation of Flavell's and Nelson and Narens' metacognitive theories that preserves their psychological detail. MGV extends the Generate-Verify paradigm by adding explicit monitoring that captures metacognitive experiences (from difficulty assessments to confidence judgements) before generation begins and refines future monitoring through verification feedback. Though we present no empirical validation, MGV provides a vocabulary for diagnosing component-level failures in reasoning systems, suggests specific architectural interventions for future designs, and identifies connections to resource-rational analysis that may ground its mechanisms in normative principles.

📄 PDF Abstract BibTeX arXiv:2511.04341

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Before you <think>, monitor: Implementing Flavell's metacognitive framework in LLMs

2025-10-18 · Nick Oh arxiv

Current approaches to enhancing LLM reasoning follows two isolated paradigms: Monitor-Generate methods like Plan-and-Solve (Wang et al., 2023) and SELF-DISCOVER (Zhou et al., 2024) excel at strategic planning but lack me…

Arithmetic Reasoning

Think$^{2}$: Grounded Metacognitive Reasoning in Large Language Models

2026-02-21 · Abraham Paul Elenjical, Vivek Hruday Kavuri, Vasudeva Varma arxiv

Large Language Models (LLMs) demonstrate strong reasoning performance, yet their ability to reliably monitor, diagnose, and correct their own errors remains limited. We introduce a psychologically grounded metacognitive …

LLMs Know When They Know, but Do Not Act on It: A Metacognitive Harness for Test-time Scaling

2026-05-13 · Qi Cao, Yufan Wang, Peijia Qin, Shuhao Zhang 외 arxiv

Large language models (LLMs) often expose useful signals of self-monitoring: before solving a problem, they can estimate whether they are likely to succeed, and after solving it, they can judge whether their answer is li…

Multimodal Reasoning

The Metacognitive Monitoring Battery: A Cross-Domain Benchmark for LLM Self-Monitoring

2026-04-17 · Jon-Paul Cacioli arxiv

We introduce a cross-domain behavioural assay of monitoring-control coupling in LLMs, grounded in the Nelson and Narens (1990) metacognitive framework and applying human psychometric methodology to LLM evaluation. The ba…

Metacognitive Learning Approach for Online Tool Condition Monitoring

2017-05-06 · Mahardhika Pratama, Eric Dimla, Chow Yin Lai, Edwin Lughofer

As manufacturing processes become increasingly automated, so should tool condition monitoring (TCM) as it is impractical to have human workers monitor the state of the tools continuously. Tool condition is crucial to ens…