Episode Description
Today Marius Hobbhahn of Apollo Research joins The Cognitive Revolution to discuss their collaboration with OpenAI using "deliberative alignment" to reduce AI scheming behavior by 30x, exploring the safety challenges and concerning findings about models' growing situational awareness and increasingly cryptic reasoning patterns that emerge when frontier models like o3 and o4-mini operate with hidden chains of thought.
Check out our sponsors: Fin, Linear, Oracle Cloud Infrastructure.
Shownotes