Episode

The Professor of Outputmaxxing — Anjney Midha, AMP

Podcast
Latent Space: The AI Engineer Podcast
Published
Jun 18, 2026
Duration seconds
3565
Processing state
not_requested
Canonical source
https://www.latent.space/p/anj
Audio
https://api.substack.com/feed/podcast/202359797/7d6863592b786561d5ce8ed820585ddb.mp3
JSON
/v1/public/podcasts/latent-space-ai-engineer/episodes/the-professor-of-outputmaxxing-anjney-midha-amp
Markdown
/podcast/latent-space-ai-engineer/the-professor-of-outputmaxxing-anjney-midha-amp.md

Actions

  • POST https://stenobird.com/v1/public/podcasts/latent-space-ai-engineer/episodes/the-professor-of-outputmaxxing-anjney-midha-amp/transcription-requests
    Idempotently request low-priority transcript generation for this episode.
  • GET https://stenobird.com/podcast/latent-space-ai-engineer/the-professor-of-outputmaxxing-anjney-midha-amp.md
    Read the agent-friendly Markdown representation of this episode resource.

Summary

Last 4 days before regular tickets sell out at AI Engineer World’s Fair - this is the single biggest gathering of AI Engineers, Founders, Leaders, and Researchers in the world. Attendees get >$5000 worth of sponsor credits and talk tracks are looking FANTASTIC. Join us! The AI scaling debate always focuses on the question of “how do we get more GPUs?” but the better question may be: how do we make the most of ones we already have. The fact that a frontier lab like xAI could be running at sub-10% MFU (Model FLOPs Utilization) is just a hint at what the real problem may be. For context, older frontier-scale training runs were already much higher than 10%. GPT-3 was around 21% MFU . Gopher was around 32% . Megatron-Turing NLG was around 30% . PaLM reached around 46% . And our guest Anjney says best-in-class MFU today is closer to 60–70% . It’s not necessarily that xAI is uniquely incompetent (it’s clear they have talented folks) but rather the priorities may be flipped in the GPU arms race. While GPU access is a bottleneck, simply increasing CapEx won’t automatically translate to better models as frontier AI is increasingly a systems problem : scheduling, utilization, networking, kernels, frameworks, data pipelines, parallelism, cluster reliability, and the thousand small decisions that determine whether your theoretical FLOPs become real training progress. From building Discord’s developer platform and backing frontier AI companies like Anthropic, Mistral, Black Forest Labs, and Periodic Labs to now building AMP’s independent compute grid, Anjney Midha has spent years close to the real bottlenecks of AI scaling. In this episode, Anjney joins swyx at Periodic Labs to unpack why the AI race is not just about buying more GPUs , why 95% utilization would have been considered…