Episode

Ep 80: CEO of Surge AI Edwin Chen on Why Frontier Labs Are Diverging, RL Environments & Developing Model Taste

Podcast
Unsupervised Learning with Jacob Effron
Published
Dec 15, 2025
Duration seconds
2881
Processing state
not_requested
Canonical source
https://unsupervised-learning.simplecast.com/episodes/ep-80-ceo-of-surge-ai-edwin-chen-on-why-frontier-labs-are-diverging-rl-environments-developing-model-taste-_RAZbWeP
Audio
https://cdn.simplecast.com/audio/2c08ad29-5b79-42c0-a40a-6c1af4327f2f/episodes/90a74f7d-fd57-41fa-8f5d-f56c0ae9c860/audio/11d3a2e1-6726-4865-9c4f-245894b4f3b9/default_tc.mp3?aid=rss_feed&feed=dOSE_bdP
JSON
/v1/public/podcasts/unsupervised-learning-with-jacob-effron-6041643/episodes/ep-80-ceo-of-surge-ai-edwin-chen-on-why-frontier-labs-are-diverging-rl-environments-developing-model-taste
Markdown
/podcast/unsupervised-learning-with-jacob-effron-6041643/ep-80-ceo-of-surge-ai-edwin-chen-on-why-frontier-labs-are-diverging-rl-environments-developing-model-taste.md

Actions

  • POST https://stenobird.com/v1/public/podcasts/unsupervised-learning-with-jacob-effron-6041643/episodes/ep-80-ceo-of-surge-ai-edwin-chen-on-why-frontier-labs-are-diverging-rl-environments-developing-model-taste/transcription-requests
    Idempotently request low-priority transcript generation for this episode.
  • GET https://stenobird.com/podcast/unsupervised-learning-with-jacob-effron-6041643/ep-80-ceo-of-surge-ai-edwin-chen-on-why-frontier-labs-are-diverging-rl-environments-developing-model-taste.md
    Read the agent-friendly Markdown representation of this episode resource.

Summary

Edwin Chen is the founder and CEO of Surge AI, the data infrastructure company behind nearly every major frontier model. Surge works with OpenAI, Anthropic, Meta, and Google, providing the high-quality data and evaluation infrastructure that powers their models. Edwin reveals why optimizing for popular benchmarks like LMArena is "basically optimizing for clickbait," how one frontier lab's models regressed for 6-12 months without anyone knowing, and why the industry's approach to measurement is fundamentally broken. Jacob and Edwin discuss what actually makes elite AI evaluators, why "there's never going to be a one size fits all solution" for AI models, and how frontier labs are taking surprisingly divergent paths to AGI.