Untangling Neural Network Mechanisms: Goodfire's Lee Sharkey on Parameter-based Interpretability

"The Cognitive Revolution" | AI Builders, Researchers, and Live Player Analysis
27 August 2025 2h 2m
0:00 --:--
Episode Description
Today Lee Sharkey of Goodfire joins The Cognitive Revolution to discuss his research on parameter decomposition methods that break down neural networks into interpretable computational components, exploring how his team's "stochastic parameter decomposition" approach addresses the limitations of sparse autoencoders and offers new pathways for understanding, monitoring, and potentially steering AI systems at the mechanistic level. Check out our sponsors: Oracle Cloud Infrastructure, Shopify. S

Shared via Hopper