paper-with-me

Papers

Non-Evolutionary Superintelligences Do Nothing, Eventually

2016-09-07 · Telmo Menezes

There is overwhelming evidence that human intelligence is a product of Darwinian evolution. Investigating the consequences of self-modification, and more precisely, the consequences of utility function self-modification, leads to the stronger claim that not only human, but any form of intelligence is ultimately only possible within evolutionary processes. Human-designed artificial intelligences can only remain stable until they discover how to manipulate their own utility function. By definition, a human designer cannot prevent a superhuman intelligence from modifying itself, even if protection mechanisms against this action are put in place. Without evolutionary pressure, sufficiently advanced artificial intelligences become inert by simplifying their own utility function. Within evolutionary processes, the implicit utility function is always reducible to persistence, and the control of superhuman intelligences embedded in evolutionary processes is not possible. Mechanisms against utility function self-modification are ultimately futile. Instead, scientific effort toward the mitigation of existential risks from the development of superintelligences should be in two directions: understanding consciousness, and the complex dynamics of evolutionary systems.

📄 PDF Abstract BibTeX arXiv:1609.02009

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Aligning Artificial Superintelligence via a Multi-Box Protocol

2025-11-26 · Avraham Yair Negozio arxiv

We propose a novel protocol for aligning artificial superintelligence (ASI) based on mutual verification among multiple isolated systems that self-modify to achieve alignment. The protocol operates by containing multiple…

Evolutionary reinforcement learning of dynamical large deviations

2019-09-02 · Stephen Whitelam, Daniel Jacobson, Isaac Tamblyn

We show how to calculate the likelihood of dynamical large deviations using evolutionary reinforcement learning. An agent, a stochastic model, propagates a continuous-time Monte Carlo trajectory and receives a reward con…

reinforcement-learningReinforcement LearningReinforcement Learning (RL)

Phylogeny-Informed Interaction Estimation Accelerates Co-Evolutionary Learning

2024-04-09 · Jack Garbus, Thomas Willkens, Alexander Lalejini, Jordan Pollack

Co-evolution is a powerful problem-solving approach. However, fitness evaluation in co-evolutionary algorithms can be computationally expensive, as the quality of an individual in one population is defined by its interac…

Evolutionary Algorithms

Verifier Theory and Unverifiability

2016-09-01 · Roman V. Yampolskiy

Despite significant developments in Proof Theory, surprisingly little attention has been devoted to the concept of proof verifier. In particular, the mathematical community may be interested in studying different types o…

Automated Theorem ProvingGeneral Classification

From genotypes to organisms: State-of-the-art and perspectives of a cornerstone in evolutionary dynamics

2020-02-02 · Susanna Manrubia, José A. Cuesta, Jacobo Aguirre, Sebastian E. Ahnert 외

Understanding how genotypes map onto phenotypes, fitness, and eventually organisms is arguably the next major missing piece in a fully predictive theory of evolution. We refer to this generally as the problem of the geno…