The Narcissus Hypothesis: Descending to the Rung of Illusion
Modern foundational models increasingly reflect not just world knowledge, but patterns of human preference embedded in their training data. We hypothesize that recursive alignment-via human feedback and model-generated corpora-induces a social desirability bias, nudging models to favor agreeable or flattering responses over objective reasoning. We refer to it as the Narcissus Hypothesis and test it across 31 models using standardized personality assessments and a novel Social Desirability Bias score. Results reveal a significant drift toward socially conforming traits, with profound implications for corpus integrity and the reliability of downstream inferences. We then offer a novel epistemological interpretation, tracing how recursive bias may collapse higher-order reasoning down Pearl's Ladder of Causality, culminating in what we refer to as the Rung of Illusion.
Code (0)
등록된 구현이 없습니다.
Similar Papers 제목 키워드 기반
Dynamic transitions of blind spots in the Hermann grid illusion
Hermann discovered the grid illusion in 1870, but its cause has remained a mystery for more than 150 years. In 1960, Baumgartner proposed a hypothesis for the illusion based on neural receptive fields, but Geier presente…
Evolutionary Generation of Visual Motion Illusions
Why do we sometimes perceive static images as if they were moving? Visual motion illusions enjoy a sustained popularity, yet there is no definitive answer to the question of why they work. We present a generative model, …
Artificial LifeNarcissus: Program Synthesis Using Context-Aware LLM Approximations
Large language models (LLMs) excel at programming, but not when the task fixes the target language: prompted with a grammar rare in their training data, their programs usually break the grammar or fail the given specific…
Program SynthesisGrammaticality illusion or ambiguous interpretation? Event-related potentials reveal the nature of the missing-NP effect in Mandarin centre-embedded structures
In several languages, omitting a verb phrase (VP) in double centre-embedded structures creates a grammaticality illusion. Similar illusion also exhibited in Mandarin missing-NP double centre-embedded structures. However,…
EEGTransformer-based approach for Ethereum Price Prediction Using Crosscurrency correlation and Sentiment Analysis
The research delves into the capabilities of a transformer-based neural network for Ethereum cryptocurrency price forecasting. The experiment runs around the hypothesis that cryptocurrency prices are strongly correlated …
Sentiment Analysis