paper-with-me

Papers

Artificial Intelligence: Arguments for Catastrophic Risk

2024-01-27 · Adam Bales, William D'Alessandro, Cameron Domenico Kirk-Giannini

Recent progress in artificial intelligence (AI) has drawn attention to the technology's transformative potential, including what some see as its prospects for causing large-scale harm. We review two influential arguments purporting to show how AI could pose catastrophic risks. The first argument -- the Problem of Power-Seeking -- claims that, under certain assumptions, advanced AI systems are likely to engage in dangerous power-seeking behavior in pursuit of their goals. We review reasons for thinking that AI systems might seek power, that they might obtain it, that this could lead to catastrophe, and that we might build and deploy such systems anyway. The second argument claims that the development of human-level AI will unlock rapid further progress, culminating in AI systems far more capable than any human -- this is the Singularity Hypothesis. Power-seeking behavior on the part of such systems might be particularly dangerous. We discuss a variety of objections to both arguments and conclude by assessing the state of the debate.

📄 PDF Abstract BibTeX arXiv:2401.15487

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Revisiting the shutdown problem

2026-06-06 · David Thorstad arxiv

A key premise in leading arguments for existential risk from artificial intelligence is that malfunctioning artificial agents could not be easily shut down. This motivates the catastrophic shutdown problem of ensuring th…

AI Must not be Fully Autonomous

2025-07-31 · Tosin Adewumi, Lama Alkhaled, Florent Imbert, Hui Han 외 arxiv

Autonomous Artificial Intelligence (AI) has many benefits. It also has many risks. In this work, we identify the 3 levels of autonomous AI. We are of the position that AI must not be fully autonomous because of the many …

Arguments about Highly Reliable Agent Designs as a Useful Path to Artificial Intelligence Safety

2022-01-09 · Issa Rice, David Manheim

Several different approaches exist for ensuring the safety of future Transformative Artificial Intelligence (TAI) or Artificial Superintelligence (ASI) systems, and proponents of different approaches have made different …

Actionable Guidance for High-Consequence AI Risk Management: Towards Standards Addressing AI Catastrophic Risks

2022-06-17 · Anthony M. Barrett, Dan Hendrycks, Jessica Newman, Brandie Nonnecke

Artificial intelligence (AI) systems can provide many beneficial capabilities but also risks of adverse events. Some AI systems could present risks of events with very high or catastrophic consequences at societal scale.…

Management

Why do Experts Disagree on Existential Risk and P(doom)? A Survey of AI Experts

2025-01-25 · Severin Field

The development of artificial general intelligence (AGI) is likely to be one of humanity's most consequential technological advancements. Leading AI labs and scientists have called for the global prioritization of AI saf…