
The episode discusses highlights from Week 2 following Anthropic's Fable launch, covering various AI applications and theories.
Week 2 highlights follows Anthropic’s Fable launch in real workflows, from safety gates and API refusals to autonomous coding, 3D world-building, and a Claude-run Twitter experiment. Geoffrey Irving and Daniel Murfet argue for alignment theory and guarantees before recursive self-improvement, while prinz tests Fable on legal reasoning and monitoring. Rahul Sonwalkar, Shlok Khemani, Tom McGrath, and Andrew Moore add field reports on data agents, hybrid authorship, interpretability, context systems, token economics, and power concentration. Mercury: Run your finances with virtual cards, spending limits, merchant/category locks, and AI-friendly tools like API keys, MCP, and CLI. Check out Mercury at https://mercury.com LINKS: Claude Fable 5 announcement Julius AI platform Rahul Sonwalkar homepage Nate Jones homepage Shlok Khemani homepage FrontierCode benchmark blog Lovelace AI company Andrew Moore Wikipedia profile Geoffrey Irving homepage Daniel Murfet LessWrong profile Sequent Research announcement Timaeus research organization Automated Alignment paper Goodfire AI company Tom McGrath homepage Predictive data debugging tool prinzbench legal benchmark Unit distance conjecture…
Claude
Explore listener stats, chart rankings, contacts and more on the "The Cognitive Revolution" | AI Builders, Researchers, and Live Player Analysis podcast page.