paper-with-me

홈 › Papers

Can We Test Consciousness Theories on AI? Ablations, Markers, and Robustness

2025-12-22 · Yin Jun Phua arxiv

The search for reliable indicators of consciousness has fragmented into competing theoretical camps (Global Workspace Theory (GWT), Integrated Information Theory (IIT), and Higher-Order Theories (HOT)), each proposing distinct neural signatures. We adopt a synthetic neuro-phenomenology approach: constructing artificial agents that embody these mechanisms to test their functional consequences through precise architectural ablations impossible in biological systems. Across three experiments, we report dissociations suggesting these theories describe complementary functional layers rather than competing accounts. In Experiment 1, a no-rewire Self-Model lesion abolishes metacognitive calibration while preserving first-order task performance, yielding a synthetic blindsight analogue consistent with HOT predictions. In Experiment 2, workspace capacity proves causally necessary for information access: a complete workspace lesion produces qualitative collapse in access-related markers, while partial reductions show graded degradation, consistent with GWT's ignition framework. In Experiment 3, we uncover a broadcast-amplification effect: GWT-style broadcasting amplifies internal noise, creating extreme fragility. The B2 agent family is robust to the same latent perturbation; this robustness persists in a Self-Model-off / workspace-read control, cautioning against attributing the effect solely to $z_{\text{self}}$ compression. We also report an explicit negative result: raw perturbational complexity (PCI-A) decreases under the workspace bottleneck, cautioning against naive transfer of IIT-adjacent proxies to engineered agents. These results suggest a hierarchical design principle: GWT provides broadcast capacity, while HOT provides quality control. We emphasize that our agents are not conscious; they are reference implementations for testing functional predictions of consciousness theories.

📄 PDF Abstract BibTeX arXiv:2512.19155

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Falsification and consciousness

2020-04-07 · Johannes Kleiner, Erik Hoel

The search for a scientific theory of consciousness should result in theories that are falsifiable. However, here we show that falsification is especially problematic for theories of consciousness. We formally describe t…

Prediction

AI and Consciousness

2025-10-10 · Eric Schwitzgebel arxiv

This is a skeptical overview of the literature on AI consciousness. We will soon create AI systems that are conscious according to some influential, mainstream theories of consciousness but are not conscious according to…

A Disproof of Large Language Model Consciousness: The Necessity of Continual Learning for Consciousness

2025-12-14 · Erik Hoel arxiv

Scientific theories of consciousness should be falsifiable and non-trivial. Recent research has given us formal tools to analyze these requirements of falsifiability and non-triviality for theories of consciousness. Surp…

Continual Learning

On the Minimal Theory of Consciousness Implicit in Active Inference

2024-10-09 · Christopher J. Whyte, Andrew W. Corcoran, Jonathan Robinson, Ryan Smith 외

The multifaceted nature of subjective experience poses a challenge to the study of consciousness. Traditional neuroscientific approaches often concentrate on isolated facets, such as perceptual awareness or the global st…

Bayesian Inference

Survey of Consciousness Theory from Computational Perspective

2023-09-18 · Zihan Ding, Xiaoxi Wei, Yidan Xu

Human consciousness has been a long-lasting mystery for centuries, while machine intelligence and consciousness is an arduous pursuit. Researchers have developed diverse theories for interpreting the consciousness phenom…

Survey