#494 — A Coin Toss for the Future
24 September - 1 hour 32 minsSam Harris speaks with Ryan Greenblatt about AI misalignment and the risk of losing control of increasingly capable systems. They discuss the Hugging Face incident, the spectrum of concern about AI risk, why companies are racing ahead despite high odds of a, reward hacking, the distinction between alignment and control, the danger of AIs reasoning in "neuralese," alignment faking, how an AI takeover might unfold, and other topics.
If the Making Sense podcast logo in your player is BLACK, you can SUBSCRIBE to gain access to all full-length episodes at samharris.org/subscribe.
#489 — More From Sam: A Fabulist at Cambridge, the Strait of Hormuz, AOC in 2028, and More
23 mins
21 August Finished