paper-with-me

홈 › Papers

The AI Risk Spectrum: From Dangerous Capabilities to Existential Threats

2025-08-19 · Markov Grey, Charbel-Raphaël Segerie arxiv

As AI systems become more capable, integrated, and widespread, understanding the associated risks becomes increasingly important. This paper maps the full spectrum of AI risks, from current harms affecting individual users to existential threats that could endanger humanity's survival. We organize these risks into three main causal categories. Misuse risks, which occur when people deliberately use AI for harmful purposes - creating bioweapons, launching cyberattacks, adversarial AI attacks or deploying lethal autonomous weapons. Misalignment risks happen when AI systems pursue outcomes that conflict with human values, irrespective of developer intentions. This includes risks arising through specification gaming (reward hacking), scheming and power-seeking tendencies in pursuit of long-term strategic goals. Systemic risks, which arise when AI integrates into complex social systems in ways that gradually undermine human agency - concentrating power, accelerating political and economic disempowerment, creating overdependence that leads to human enfeeblement, or irreversibly locking in current values curtailing future moral progress. Beyond these core categories, we identify risk amplifiers - competitive pressures, accidents, corporate indifference, and coordination failures - that make all risks more likely and severe. Throughout, we connect today's existing risks and empirically observable AI behaviors to plausible future outcomes, demonstrating how existing trends could escalate to catastrophic outcomes. Our goal is to help readers understand the complete landscape of AI risks. Good futures are possible, but they don't happen by default. Navigating these challenges will require unprecedented coordination, but an extraordinary future awaits if we do.

📄 PDF Abstract BibTeX arXiv:2508.13700

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Artificial General Intelligence, Existential Risk, and Human Risk Perception

2023-11-15 · David R. Mandel

Artificial general intelligence (AGI) does not yet exist, but given the pace of technological development in artificial intelligence, it is projected to reach human-level intelligence within roughly the next two decades.…

Democratising Risk: In Search of a Methodology to Study Existential Risk

2021-12-27 · Carla Zoe Cremer, Luke Kemp

Studying potential global catastrophes is vital. The high stakes of existential risk studies (ERS) necessitate serious scrutiny and self-reflection. We argue that existing approaches to studying existential risk are not …

Ethics

AI Consciousness and Existential Risk

2025-11-24 · Rufin VanRullen arxiv

In AI, the existential risk denotes the hypothetical threat posed by an artificial system that would possess both the capability and the objective, either directly or indirectly, to eradicate humanity. This issue is gain…

Current and Near-Term AI as a Potential Existential Risk Factor

2022-09-21 · Benjamin S. Bucknall, Shiri Dori-Hacohen

There is a substantial and ever-growing corpus of evidence and literature exploring the impacts of Artificial intelligence (AI) technologies on society, politics, and humanity as a whole. A separate, parallel body of wor…

Defining and Evaluating Physical Safety for Large Language Models

2024-11-04 · Yung-Chen Tang, Pin-Yu Chen, Tsung-Yi Ho

Large Language Models (LLMs) are increasingly used to control robotic systems such as drones, but their risks of causing physical threats and harm in real-world applications remain unexplored. Our study addresses the cri…

Code GenerationIn-Context LearningPrompt Engineering