paper-with-me

홈 › Papers

Breaking the Illusion of Identity in LLM Tooling

2026-04-08 · Marek Miller arxiv

Large language models (LLMs) in research and development toolchains produce output that triggers attribution of agency and understanding -- a cognitive illusion that degrades verification behavior and trust calibration. No existing mitigation provides a systematic, deployable constraint set for output register. This paper proposes seven output-side rules, each targeting a documented linguistic mechanism, and validates them empirically. In 780 two-turn conversations (constrained vs. default register, 30 tasks, 13 replicates, 1560 API calls), anthropomorphic markers dropped from 1233 to 33 (>97% reduction, p < 0.001), outputs were 49% shorter by word count, and adapted AnthroScore confirmed the shift toward machine register (-1.94 vs. -0.96, p < 0.001). The rules are implemented as a configuration-file system prompt requiring no model modification; validation uses a single model (Claude Sonnet 4). Output quality under the constrained register was not evaluated. The mechanism is extensible to other domains.

📄 PDF Abstract BibTeX arXiv:2604.07398

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Breaking the Illusion: Consensus-Based Generative Mitigation of Adversarial Illusions in Multi-Modal Embeddings

2025-11-26 · Fatemeh Akbarian, Anahita Baninajjar, Yingyi Zhang, Ananth Balashankar 외 arxiv

Multi-modal foundation models align images, text, and other modalities in a shared embedding space but remain vulnerable to adversarial illusions [35], where imperceptible perturbations disrupt cross-modal alignment and …

The Grand Illusion: The Myth of Software Portability and Implications for ML Progress

2023-09-12 · Fraser Mince, Dzung Dinh, Jonas Kgomo, Neil Thompson 외

Pushing the boundaries of machine learning often requires exploring different hardware and software combinations. However, the freedom to experiment across different tooling stacks can be at odds with the drive for effic…

Friction

The Grand Illusion: The Myth of Software Portability and Implications for ML Progress.

2023-09-21 · NeurIPS 2023 11

Pushing the boundaries of machine learning often requires exploring different hardware and software combinations. However, this ability to experiment with different systems can be at odds with the drive for efficiency, w…

Trust and Reliance in Consensus-Based Explanations from an Anti-Misinformation Agent

2023-04-22 · Takane Ueno, Yeongdae Kim, Hiroki Oura, Katie Seaborn

The illusion of consensus occurs when people believe there is consensus across multiple sources, but the sources are the same and thus there is no "true" consensus. We explore this phenomenon in the context of an AI-base…

Explainable Artificial Intelligence (XAI)Misinformation

Safety Is Not Universal: The Selective Safety Trap in LLM Alignment

2026-01-07 · Iago Alves Brito, Walcy Santos Rezende Rios, Julia Soares Dollis, Diogo Fernandes Costa Silva 외 arxiv

Current safety evaluations of large language models (LLMs) create a dangerous illusion of universal protection by aggregating harms under generic categories such as "Identity Hate", obscuring vulnerabilities toward speci…