paper-with-me

홈 › Papers

An overview of 11 proposals for building safe advanced AI

2020-12-04 · Evan Hubinger

This paper analyzes and compares 11 different proposals for building safe advanced AI under the current machine learning paradigm, including major contenders such as iterated amplification, AI safety via debate, and recursive reward modeling. Each proposal is evaluated on the four components of outer alignment, inner alignment, training competitiveness, and performance competitiveness, of which the distinction between the latter two is introduced in this paper. While prior literature has primarily focused on analyzing individual proposals, or primarily focused on outer alignment at the expense of inner alignment, this analysis seeks to take a comparative look at a wide range of proposals including a comparative analysis across all four previously mentioned components.

📄 PDF Abstract BibTeX arXiv:2012.07532

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

International Agreements on AI Safety: Review and Recommendations for a Conditional AI Safety Treaty

2025-03-18 · Rebecca Scholefield, Samuel Martin, Otto Barten

The malicious use or malfunction of advanced general-purpose AI (GPAI) poses risks that, according to leading experts, could lead to the 'marginalisation or extinction of humanity.' To address these risks, there are an i…

Taking control: Policies to address extinction risks from AI

2023-10-31 · Andrea Miotti, Akash Wasil

This paper provides policy recommendations to reduce extinction risks from advanced artificial intelligence (AI). First, we briefly provide background information about extinction risks from AI. Second, we argue that vol…

A Systematic Literature Review on Multi-label Data Stream Classification

2025-08-24 · H. Freire-Oliveira, E. R. F. Paiva, J. Gama, L. Khan 외 arxiv

Classification in the context of multi-label data streams represents a challenge that has attracted significant attention due to its high real-world applicability. However, this task faces problems inherent to dynamic en…

Key Safety Design Overview in AI-driven Autonomous Vehicles

2024-12-12 · Vikas Vyas, Zheyuan Xu

With the increasing presence of autonomous SAE level 3 and level 4, which incorporate artificial intelligence software, along with the complex technical challenges they present, it is essential to maintain a high level o…

Autonomous Vehicles

Stovepiping and Malicious Software: A Critical Review of AGI Containment

2018-11-08 · Jason M. Pittman, Jesus P. Espinoza, Courtney Crosby

Awareness of the possible impacts associated with artificial intelligence has risen in proportion to progress in the field. While there are tremendous benefits to society, many argue that there are just as many, if not m…