Anthropic CEO Dario Amodei, plus Amanda Askell and Chris Olah, on Claude, scaling laws, AGI, AI safety, and interpretability.

Dario Amodei (with Amanda Askell and Chris Olah): Dario Amodei is co-founder and CEO of Anthropic, the company behind Claude. He is joined by Amanda Askell, a researcher who designs Claude's character and alignment, and Chris Olah, a pioneer of mechanistic interpretability.
Dario Amodei traces the scaling hypothesis from his early speech-recognition work to today's frontier models, arguing that bigger networks, more data, and more compute reliably yield more intelligence and could reach human-level 'powerful AI' by 2026-2027. He details Anthropic's safety framework (the Responsible Scaling Policy and ASL levels), the 'race to the top' theory of change, his views on regulation, and his optimistic essay 'Machines of Loving Grace.' Amanda Askell explains how Claude's character is crafted as an alignment problem, covering sycophancy, prompting, constitutional AI, and the ethics of AI consciousness. Chris Olah closes with a deep dive into mechanistic interpretability: features, circuits, superposition, sparse autoencoders, and the goal of understanding neural networks for both safety and beauty.