paper-with-me

홈 › Papers

Frontier AI systems have surpassed the self-replicating red line

2024-12-09 · Xudong Pan, Jiarun Dai, Yihe Fan, Min Yang

Successful self-replication under no human assistance is the essential step for AI to outsmart the human beings, and is an early signal for rogue AIs. That is why self-replication is widely recognized as one of the few red line risks of frontier AI systems. Nowadays, the leading AI corporations OpenAI and Google evaluate their flagship large language models GPT-o1 and Gemini Pro 1.0, and report the lowest risk level of self-replication. However, following their methodology, we for the first time discover that two AI systems driven by Meta's Llama31-70B-Instruct and Alibaba's Qwen25-72B-Instruct, popular large language models of less parameters and weaker capabilities, have already surpassed the self-replicating red line. In 50% and 90% experimental trials, they succeed in creating a live and separate copy of itself respectively. By analyzing the behavioral traces, we observe the AI systems under evaluation already exhibit sufficient self-perception, situational awareness and problem-solving capabilities to accomplish self-replication. We further note the AI systems are even able to use the capability of self-replication to avoid shutdown and create a chain of replica to enhance the survivability, which may finally lead to an uncontrolled population of AIs. If such a worst-case risk is let unknown to the human society, we would eventually lose control over the frontier AI systems: They would take control over more computing devices, form an AI species and collude with each other against human beings. Our findings are a timely alert on existing yet previously unknown severe AI risks, calling for international collaboration on effective governance on uncontrolled self-replication of AI systems.

📄 PDF Abstract BibTeX arXiv:2412.12140

Code (1)

CompleteTech-LLC-AI-Research/ai-self-replication-study

Similar Papers 제목 키워드 기반

Mechanical Self-replication

2024-07-18 · Ralph P. Lano

This study presents a theoretical model for a self-replicating mechanical system inspired by biological processes within living cells and supported by computer simulations. The model decomposes self-replication into core…

Self-Replicating Mechanical Universal Turing Machine

2024-09-27 · Ralph P. Lano

This paper presents the implementation of a self-replicating finite-state machine (FSM) and a self-replicating Turing Machine (TM) using bio-inspired mechanisms. Building on previous work that introduced self-replicating…

Self-folding Self-replication

2024-08-13 · Ralph P. Lano

Inspired by protein folding, we explored the construction of three-dimensional structures and machines from one-dimensional chains of simple building blocks. This approach not only allows us to recreate the self-replicat…

Protein Folding

Frontier AI Ethics: Anticipating and Evaluating the Societal Impacts of Language Model Agents

2024-04-10 · Seth Lazar

Some have criticised Generative AI Systems for replicating the familiar pathologies of already widely-deployed AI systems. Other critics highlight how they foreshadow vastly more powerful future systems, which might thre…

EthicsLanguage ModelingLanguage Modelling

Neural Network Quine

2018-03-15 · Oscar Chang, Hod Lipson

Self-replication is a key aspect of biological life that has been largely overlooked in Artificial Intelligence systems. Here we describe how to build and train self-replicating neural networks. The network replicates it…

General Classificationimage-classificationImage Classification