arrow
Return

Imitation learning is probably existentially safe

delete2025-11-21
delete0
delete
OA
AI
M
Michael K. Cohen *
M
Marcus Hütter
DOI:10.1002/aaai.70040delete
deleteOriginal
deleteShare
deleteSave
View PDF
Abstract

Abstract

En 中文
Concerns about extinction risk from AI vary among experts in the field. However, AI encompasses a very broad category of algorithms. Perhaps some algorithms would pose an extinction risk, and others would not. Such an observation might be of great interest to both regulators and innovators. This paper argues that advanced imitation learners would likely not cause human extinction. We first present a simple argument to that effect, and then we rebut six different arguments that have been made to the contrary. A common theme of most of these arguments is a story for how a subroutine within an advanced imitation learner could hijack the imitation learner's behavior toward its own ends. However, we argue that each argument is flawed and each story implausible.
Keywords:
imitation learning
existential risk
artificial intelligence
expert opinion
algorithm safety
AI Summary

AI Summary

Key information extracted from the uploaded paper, including a brief overview, abstract, background, key highlights, visual analysis, and future outlook.

Journal

A
AI Magazine
IF:
3.2
Papers:
35
Citations:
3.2K

Organization

University of California System cover
University of California System
Scholars:
37.5W
Papers: 33.7W
Citations: 6.6K
U
university of california berkeley
Scholars:
1.4K
Papers: 799
Citations: 0