Episode

Thinking in Silico: Goodfire CTO Dan Balsam on Concept Manifolds & a $1000/Month ML Research Agent

Podcast
"The Cognitive Revolution"
Published
Aug 8, 2026
Duration seconds
7041
Processing state
not_requested
Canonical source
https://www.cognitiverevolution.ai/thinking-in-silico-goodfire-cto-dan-balsam-on-concept-manifolds-a-1000-month-ml-research-agent/
Audio
https://pdst.fm/e/mgln.ai/e/1113/pscrb.fm/rss/p/traffic.megaphone.fm/RINTP8369180865.mp3
JSON
/v1/public/podcasts/the-cognitive-revolution/episodes/thinking-in-silico-goodfire-cto-dan-balsam-on-concept-manifolds-a-1000-month-ml-research-agent
Markdown
/podcast/the-cognitive-revolution/thinking-in-silico-goodfire-cto-dan-balsam-on-concept-manifolds-a-1000-month-ml-research-agent.md

Actions

  • POST https://stenobird.com/v1/public/podcasts/the-cognitive-revolution/episodes/thinking-in-silico-goodfire-cto-dan-balsam-on-concept-manifolds-a-1000-month-ml-research-agent/transcription-requests
    Idempotently request low-priority transcript generation for this episode.
  • GET https://stenobird.com/podcast/the-cognitive-revolution/thinking-in-silico-goodfire-cto-dan-balsam-on-concept-manifolds-a-1000-month-ml-research-agent.md
    Read the agent-friendly Markdown representation of this episode resource.

Summary

Goodfire co-founder and CTO Dan Balsam returns to discuss where interpretability research now stands and to introduce Silico, the $1,000-per-month research platform Goodfire built for itself. He and Nathan explore Predictive Data Debugging, including the idea that fine-tuning and RL often amplify behaviors already latent in pre-training, and that interpretability can identify the data and features driving unwanted updates. The conversation centers on concept manifolds: Dan argues that models do not store concepts as simple one-hot features, but as sparse mixtures of meaningful subspaces whose geometry determines what kinds of steering and control work. The stakes are practical as well as conceptual, from debugging training data and RL to understanding why steering can fail off-manifold and why modern interpretability may be moving beyond its reputation as a toy-model science. Silico: https://www.goodfire.com/silico Predictive data debugging: https://www.goodfire.com/research/predictive-data-debugging# Neural Geometry: https://www.goodfire.com/research/the-world-inside-neural-networks# For full show notes, links, and references, read the episode page:https://www.cognitiverevolution.ai/thinking-in-silico-goodfire-cto-dan-balsam-on-concept-manifolds-a-1000-month-ml-research-agent/ Sponsor: Claude: Claude by Anthropic is an AI collaborator that understands your workflow and helps you tackle research, writing, coding, and organization with deep context. Get started with Claude and explore Claude Pro at https://claude.ai/tcr CHAPTERS: (00:00) About the Episode (03:22) Predictive data debugging (12:36) Concept manifold geometry (21:22) Finding concept manifolds (Part 1) (21:28) Sponsor: Claude (22:57) Finding concept manifolds (Part 2) (33:24) Factoring model internals (49:32…