paper-with-me

Papers

Indications of Belief-Guided Agency and Meta-Cognitive Monitoring in Large Language Models

2026-02-02 · Noam Steinmetz Yalon, Ariel Goldstein, Liad Mudrik, Mor Geva arxiv

Rapid advancements in large language models (LLMs) have sparked the question whether these models possess some form of consciousness. To tackle this challenge, Butlin et al. (2023) introduced a list of indicators for consciousness in artificial systems based on neuroscientific theories. In this work, we evaluate a key indicator from this list, called HOT-3, which tests for agency guided by a general belief-formation and action selection system that updates beliefs based on meta-cognitive monitoring. We view beliefs as representations in the model's latent space that emerge in response to a given input, and introduce a metric to quantify their dominance during generation. Analyzing the dynamics between competing beliefs across models and tasks reveals three key findings: (1) external manipulations systematically modulate internal belief formation, (2) belief formation causally drives the model's action selection, and (3) models can monitor and report their own belief states. Together, these results provide empirical support for the existence of belief-guided agency and meta-cognitive monitoring in LLMs. More broadly, our work lays methodological groundwork for investigating the emergence of agency, beliefs, and meta-cognition in LLMs.

📄 PDF Abstract BibTeX arXiv:2602.02467

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Metacognitive particles, mental action and the sense of agency

2024-05-21 · Lars Sandved-Smith, Lancelot Da Costa

This paper articulates metacognition using the language of statistical physics and Bayesian mechanics. Metacognitive beliefs, defined as beliefs about beliefs, find a natural description within this formalism, which allo…

In-context learning agents are asymmetric belief updaters

2024-02-06 · Johannes A. Schubert, Akshay K. Jagadish, Marcel Binz, Eric Schulz

We study the in-context learning dynamics of large language models (LLMs) using three instrumental learning tasks adapted from cognitive psychology. We find that LLMs update their beliefs in an asymmetric manner and lear…

counterfactualIn-Context LearningMeta Reinforcement Learning

The Belief-Desire-Intention Ontology for modelling mental reality and agency

2025-11-21 · Sara Zuppiroli, Carmelo Fabio Longo, Anna Sofia Lippolis, Rocco Paolillo 외 arxiv

The Belief-Desire-Intention (BDI) model is a cornerstone for representing rational agency in artificial intelligence and cognitive sciences. Yet, its integration into structured, semantically interoperable knowledge repr…

MetaMind: General and Cognitive World Models in Multi-Agent Systems by Meta-Theory of Mind

2026-02-28 · Lingyi Wang, Rashed Shelim, Walid Saad, Naren Ramakrishna arxiv

A major challenge for world models in multi-agent systems is to understand interdependent agent dynamics, predict interactive multi-agent trajectories, and plan over long horizons with collective awareness, without centr…

Investigating Agency of LLMs in Human-AI Collaboration Tasks

2023-05-22 · ASHISH SHARMA, Sudha Rao, Chris Brockett, Akanksha Malhotra 외

Agency, the capacity to proactively shape events, is central to how humans interact and collaborate. While LLMs are being developed to simulate human behavior and serve as human-like agents, little attention has been giv…