paper-with-me

홈 › Papers

The Evolutionary Origin of Values: implications for AI alignment, sentience and existential risk

2026-08-04 · Francis Heylighen arxiv

AI systems based on Large Language Models (LLMs) have prompted fears that they may harbor hidden goals, seek to dominate or eliminate humanity, or even suffer as sentient beings. We address these concerns by tracing the evolutionary origin of value in biological organisms. Values emerge from autopoiesis: living systems must actively maintain themselves against perturbation and dissipation. Natural selection has equipped them with hierarchies of "vicarious selectors" that guide their behavior toward fitness. LLMs, by contrast, are allopoietic and allotelic: they produce outputs for others, and their goals derive from user prompts rather than an autonomous drive. They lack the intrinsic motivation for self-preservation, dominance, or resource competition that underlies existential-risk scenarios, and the embodied vulnerability required for feeling or suffering. Still, because LLMs learn statistical patterns from human-generated text, they implicitly absorb human values as well as knowledge, allowing them to focus on what is relevant. That is why the "orthogonality thesis" separating intelligence from values does not apply to them. Such separation would in fact expose any intelligence to the frame problem: the combinatorial explosion of the search space that makes any realistic utility function physically uncomputable. That also precludes the convergence of instrumental values thesis. We conclude that the real alignment challenge lies not in preventing rogue AI agency, but in ensuring LLMs intelligently apply learned ethical values.

📄 PDF Abstract BibTeX arXiv:2608.03361

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Parsimonious evolutionary scenario for the origin of allostery and coevolution patterns in proteins

2019-07-12

Proteins display generic properties that are challenging to explain by direct selection, notably allostery, the capacity to be regulated through long-range effects, and evolvability, the capacity to adapt to new selectiv…

Multi-Value Alignment in Normative Multi-Agent System: Evolutionary Optimisation Approach

2023-05-12 · Maha Riad, Vinicius Renan de Carvalho, Fatemeh Golpayegani

Value-alignment in normative multi-agent systems is used to promote a certain value and to ensure the consistent behavior of agents in autonomous intelligent systems with human values. However, the current literature is …

Evolutionary Algorithms

Mental Models of Autonomy and Sentience Shape Reactions to AI

2025-12-09 · Janet V. T. Pauketat, Daniel B. Shank, Aikaterina Manoli, Jacy Reese Anthis arxiv

Narratives about artificial intelligence (AI) entangle autonomy, the capacity to self-govern, with sentience, the capacity to sense and feel. AI agents that perform tasks autonomously and companions that recognize and ex…

Engineering Sentience

2025-06-25 · Konstantin Demin, Taylor Webb, Eric Elmoznino, Hakwan Lau

We spell out a definition of sentience that may be useful for designing and building it in machines. We propose that for sentience to be meaningful for AI, it must be fleshed out in functional, computational terms, in en…

The Sentience Readiness Index: A Preliminary Framework for Measuring National Preparedness for the Possibility of Artificial Sentience

2026-03-02 · Tony Rost arxiv

The scientific study of consciousness has begun to generate testable predictions about artificial systems. A landmark collaborative assessment evaluated current AI architectures against six leading theories of consciousn…