Human Compatible

Artificial Intelligence and the Problem of Control

Stuart Russell

17 min read
1m 27s intro

Brief summary

In Human Compatible, AI pioneer Stuart Russell argues that the standard approach to building artificial intelligence is flawed and dangerous. He proposes a new model for creating powerful systems that remain uncertain about human preferences, defer to human control, and learn from human behavior.

Who it's for

This book is for anyone interested in the future of AI who wants a clear explanation of the risks and a concrete proposal for building safe, beneficial systems.

Human Compatible

Audio & text in the Readsome app

The Promise and Peril of Artificial Intelligence

Stuart Russell observes that the field of artificial intelligence has always aimed to create machines that match or exceed human capabilities. However, there has been surprisingly little discussion about what happens if that goal is actually met. If researchers succeed in creating superintelligent systems, it would represent a massive shift in human civilization. This transition could help solve major global catastrophes or extend human life, but it could also be humanity's final event if not managed with care.

The formal quest to build intelligent machines began in 1956 when scientists proposed that every feature of learning could be simulated. Early milestones included programs that could play checkers, but these successes were followed by periods of disappointment when the technology failed to handle complex real-world tasks. Over the decades, the field shifted from simple rule-based systems to complex mathematical models involving probability and statistics. By 2011, deep learning techniques began solving long-standing problems in speech and image recognition, attracting billions of dollars in investment.

The path toward superhuman intelligence requires several major breakthroughs that are difficult to forecast. History shows that scientific progress often happens much faster than experts anticipate, such as when physicists dismissed the possibility of atomic energy just one day before the nuclear chain reaction was invented. This sudden shift from impossible to solved serves as a warning against assuming that advanced artificial intelligence is too far in the future to worry about. Even if success is not immediate, the potential impact is so great that preparing for it now is an absolute necessity.

For decades, the field of artificial intelligence has operated under a specific definition of success called the standard model. In this framework, humans provide a fixed objective, and the machine optimizes its behavior to reach that specific goal. While this works for simple tasks, it becomes dangerous as machines become more capable because humans are often unable to specify their goals perfectly. If a machine is smarter than a human and pursues a flawed objective, it will pursue that goal relentlessly, leading to unintended and potentially disastrous consequences.

Full summary available in the Readsome app

Get it on Google PlayDownload on the App Store

About the author

Stuart Russell

Stuart Russell is a British computer scientist and a leading figure in the field of artificial intelligence. He is a Professor of Computer Science at the University of California, Berkeley, where he founded and leads the Center for Human-Compatible Artificial Intelligence (CHAI). His research encompasses machine learning, probabilistic reasoning, and the long-term future of AI, and he is the co-author of the standard textbook in the field, *Artificial Intelligence: A Modern Approach*.

Similar book summaries