paper-with-me

홈 › Papers

Solipsistic Superintelligence is Unlikely to be Cooperative

2026-06-02 · Rakshit S Trivedi, Natasha Jaques, Logan Cross, Alexander Sasha Vezhnevets, Joel Z Leibo arxiv

AI's central challenge is shifting from capability to coexistence. The dominant paradigm in AI research focuses on developing powerful agents that treat the world as an exogenous and stationary source of feedback. We contend that superintelligence, an extremely capable task solver, born out of such a solipsistic approach to AI design, is unlikely to be cooperative. Deploying AI systems induces endogenous non-stationarity, resulting in a train-test-deploy gap where historical distributions diverge from the deployment context. We refer to this as the self-undermining property of unilateral optimization. Closing this gap requires AI that participates in cooperation: the equilibrium-selection process through which multiple actors navigate their interdependence. We call for a non-solipsistic research paradigm that treats this interdependence as a core design principle rather than approaching cooperation as a task to solve. This entails building dynamic evaluation testbeds involving adaptive counterparties, treating institutions as design primitives, and preserving human agency as a structural feature of the systems we build.

📄 PDF Abstract BibTeX arXiv:2606.03237

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Solipsistic Reinforcement Learning

2021-03-09 · ICLR Workshop SSL-RL 2021 5 · Mingtian Zhang, Peter Noel Hayes, Tim Z. Xiao, Andi Zhang 외

We introduce a new model-based reinforcement learning framework that aims to tackle environments with high dimensional state spaces. In contrast to existing approaches, agents under our framework learn a low dimensional …

Model-based Reinforcement Learningreinforcement-learningReinforcement LearningReinforcement Learning (RL)

Aligning Artificial Superintelligence via a Multi-Box Protocol

2025-11-26 · Avraham Yair Negozio arxiv

We propose a novel protocol for aligning artificial superintelligence (ASI) based on mutual verification among multiple isolated systems that self-modify to achieve alignment. The protocol operates by containing multiple…

Superintelligence cannot be contained: Lessons from Computability Theory

2016-07-04 · Manuel Alfonseca, Manuel Cebrian, Antonio Fernandez Anta, Lorenzo Coviello 외

Superintelligence is a hypothetical agent that possesses intelligence far surpassing that of the brightest and most gifted human minds. In light of recent advances in machine intelligence, a number of scientists, philoso…

Melting Pot 2.0

2022-11-24 · John P. Agapiou, Alexander Sasha Vezhnevets, Edgar A. Duéñez-Guzmán, Jayd Matyas 외

Multi-agent artificial intelligence research promises a path to develop intelligent technologies that are more human-like and more human-compatible than those produced by "solipsistic" approaches, which do not consider i…

Artificial LifeNavigate

Invasion of cooperative parasites in moderately structured host populations

2022-01-06 · Vianney Brouard, Cornelia Pokalyuk

Certain defense mechanisms of phages against the immune system of their bacterial host rely on cooperation of phages. Motivated by this example we analyse invasion probabilities of cooperative parasites in host populatio…