paper-with-me

홈 › Papers

Large Language Models Report Subjective Experience Under Self-Referential Processing

2025-10-27 · Cameron Berg, Diogo de Lucena, Judd Rosenblatt arxiv

Large language models sometimes produce structured, first-person descriptions that explicitly reference awareness or subjective experience. To better understand this behavior, we investigate one theoretically motivated condition under which such reports arise: self-referential processing, a computational motif emphasized across major theories of consciousness. Through a series of controlled experiments on GPT, Claude, and Gemini model families, we test whether this regime reliably shifts models toward first-person reports of subjective experience, and how such claims behave under mechanistic and behavioral probes. Four main results emerge: (1) Inducing sustained self-reference through simple prompting consistently elicits structured subjective experience reports across model families. (2) These reports are mechanistically gated by interpretable sparse-autoencoder features associated with deception and roleplay: surprisingly, suppressing deception features sharply increases the frequency of experience claims, while amplifying them minimizes such claims. (3) Structured descriptions of the self-referential state converge statistically across model families in ways not observed in any control condition. (4) The induced state yields significantly richer introspection in downstream reasoning tasks where self-reflection is only indirectly afforded. While these findings do not constitute direct evidence of consciousness, they implicate self-referential processing as a minimal and reproducible condition under which large language models generate structured first-person reports that are mechanistically gated, semantically convergent, and behaviorally generalizable. The systematic emergence of this pattern across architectures makes it a first-order scientific and ethical priority for further investigation.

📄 PDF Abstract BibTeX arXiv:2510.24797

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Mapping of Subjective Accounts into Interpreted Clusters (MOSAIC): Topic Modelling and LLM applied to Stroboscopic Phenomenology

2025-02-25 · Romy Beauté, David J. Schwartzman, Guillaume Dumas, Jennifer Crook 외

Stroboscopic light stimulation (SLS) on closed eyes typically induces simple visual hallucinations (VHs), characterised by vivid, geometric and colourful patterns. A dataset of 862 sentences, extracted from 422 open subj…

Self-Referential Induction Increases Response Instability Relative to Unresolvable and Verifiable Questions in Large Language Models

2026-08-13 · Paras Balani, Subhrakanta Panda arxiv

Self-referential prompting has been shown to reliably induce large language models to produce first-person reports resembling subjective experience, but no prior work measures how consistent these reports are across repe…

Sensing Subjective Well-being from Social Media

2014-03-15 · Bibo Hao, Lin Li, Rui Gao, Ang Li 외

Subjective Well-being(SWB), which refers to how people experience the quality of their lives, is of great use to public policy-makers as well as economic, sociological research, etc. Traditionally, the measurement of SWB…

Impact of Response Latency on User Behaviour in Mobile Web Search

2021-01-22 · Ioannis Arapakis, Souneil Park, Martin Pielot

Traditionally, the efficiency and effectiveness of search systems have both been of great interest to the information retrieval community. However, an in-depth analysis of the interaction between the response latency and…

Information RetrievalRetrieval

Representational Tenets for Memory Athletics

2023-02-22 · Kevin Schmidt, Othalia Larue, Ray Kulhanek, Dylan Flaute 외

We describe the current state of world-class memory competitions, including the methods used to prepare for and compete in memory competitions, based on the subjective report of World Memory Championship Grandmaster and …