paper-with-me

홈 › Papers

The Narcissus Hypothesis: Descending to the Rung of Illusion

2025-09-22 · Riccardo Cadei, Christian Internò arxiv

Modern foundational models increasingly reflect not just world knowledge, but patterns of human preference embedded in their training data. We hypothesize that recursive alignment-via human feedback and model-generated corpora-induces a social desirability bias, nudging models to favor agreeable or flattering responses over objective reasoning. We refer to it as the Narcissus Hypothesis and test it across 31 models using standardized personality assessments and a novel Social Desirability Bias score. Results reveal a significant drift toward socially conforming traits, with profound implications for corpus integrity and the reliability of downstream inferences. We then offer a novel epistemological interpretation, tracing how recursive bias may collapse higher-order reasoning down Pearl's Ladder of Causality, culminating in what we refer to as the Rung of Illusion.

📄 PDF Abstract BibTeX arXiv:2509.17999

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Dynamic transitions of blind spots in the Hermann grid illusion

2024-07-17 · Yutaka Nishiyama

Hermann discovered the grid illusion in 1870, but its cause has remained a mystery for more than 150 years. In 1960, Baumgartner proposed a hypothesis for the illusion based on neural receptive fields, but Geier presente…

Evolutionary Generation of Visual Motion Illusions

2021-12-25 · Lana Sinapayen, Eiji Watanabe

Why do we sometimes perceive static images as if they were moving? Visual motion illusions enjoy a sustained popularity, yet there is no definitive answer to the question of why they work. We present a generative model, …

Artificial Life

Narcissus: Program Synthesis Using Context-Aware LLM Approximations

2026-08-26 · Tilman Hinnerichs, Sebastijan Dumancic, Neil Yorke-Smith arxiv

Large language models (LLMs) excel at programming, but not when the task fixes the target language: prompted with a grammar rare in their training data, their programs usually break the grammar or fail the given specific…

Program Synthesis

Grammaticality illusion or ambiguous interpretation? Event-related potentials reveal the nature of the missing-NP effect in Mandarin centre-embedded structures

2024-02-17 · Qihang Yang, Caimei Yang, Yu Liao, Ziman Zhuang

In several languages, omitting a verb phrase (VP) in double centre-embedded structures creates a grammaticality illusion. Similar illusion also exhibited in Mandarin missing-NP double centre-embedded structures. However,…

EEG

Transformer-based approach for Ethereum Price Prediction Using Crosscurrency correlation and Sentiment Analysis

2024-01-16 · Shubham Singh, Mayur Bhat

The research delves into the capabilities of a transformer-based neural network for Ethereum cryptocurrency price forecasting. The experiment runs around the hypothesis that cryptocurrency prices are strongly correlated …

Sentiment Analysis