Making Sense with Sam Harris - #494 — A Coin Toss for the Future
Unsupported item type.
UUID: 5rgsDBhI3TbdTLjcEkRwod
Class: podcastepisode
Category: podcast
Overview
Sam Harris speaks with Ryan Greenblatt about AI misalignment and the risk of losing control of increasingly capable systems. They discuss the Hugging Face incident, the spectrum of concern about AI risk, why companies are racing ahead despite high odds of a, reward hacking, the distinction between alignment and control, the danger of AIs reasoning in "neuralese," alignment faking, how an AI takeover might unfold, and other topics. If the Making Sense podcast logo in your player is BLACK, you can SUBSCRIBE to gain access to all full-length episodes at samharris.org/subscribe.