{"podcast":{"title":"Latent Space: The AI Engineer Podcast","slug":"latent-space-ai-engineer","podcast_index_feed_id":6058902,"rss_url":"https://api.substack.com/feed/podcast/1084089.rss","website_url":"https://www.latent.space/podcast","image_url":"https://substackcdn.com/feed/podcast/1084089/ca7468da5614a246d2906ee8926f6de7.jpg","author":"Latent.Space","episode_count":217,"summary":"The AI Engineer newsletter + Top technical AI podcast. How leading labs build Agents, Models, Infra, & AI for Science. See https://latent.space/about for highlights from Greg Brockman, Andrej Karpathy, George Hotz, Simon Willison, Soumith Chintala et al!","last_synced_at":"2026-08-02T15:06:40.969607+00:00","page_url":"https://stenobird.com/podcast/latent-space-ai-engineer"},"episode":{"title":"The Professor of Outputmaxxing — Anjney Midha, AMP","slug":"the-professor-of-outputmaxxing-anjney-midha-amp","published_at":"2026-06-18T17:30:00+00:00","page_url":"https://stenobird.com/podcast/latent-space-ai-engineer/the-professor-of-outputmaxxing-anjney-midha-amp","show_page_url":"https://stenobird.com/podcast/latent-space-ai-engineer","url":"https://www.latent.space/p/anj","audio_url":"https://api.substack.com/feed/podcast/202359797/7d6863592b786561d5ce8ed820585ddb.mp3","summary":"The AI scaling race is often framed as a pursuit of more GPUs, but the real frontier lies in maximizing Model FLOPs Utilization (MFU). Anjney Midha argues that inefficient cluster management and misaligned incentives are causing massive computational waste.","meta_description":"Anjney Midha discusses why the next era of AI infrastructure requires maximizing MFU, building decentralized compute grids, and solving the alignment prob…","key_points":["Main idea: Scaling AI is increasingly a systems engineering problem involving scheduling, networking, and kernels rather than just raw CapEx","Failure mode: Low Model FLOPs Utilization (MFU) in frontier labs suggests that simply adding more GPUs won't yield proportional progress without better orchestration","Practical takeaway: The future of compute lies in a decentralized, protocol-based grid where supply and demand can flow like a utility","Main idea: High-performance clusters should aim for much higher utilization rates, noting that 95% node utilization is a standard for reliability at scale","Strategic insight: To enable new hardware, developers should adopt the NVIDIA reference architecture to ensure compatibility with existing software stacks"],"chapters":[{"start_ms":60000,"title":"The Alignment of Compute and Capital","summary":"An analysis of how the gap between funding and deployment leads to massive inefficiencies and wasted compute in large-scale clusters."},{"start_ms":600000,"title":"Building In-House Infrastructure","summary":"Lessons from Discord's approach to building proprietary communication infrastructure to avoid third-party bottlenecks."},{"start_ms":1380000,"title":"AI for End-of-Life Prediction","summary":"Exploring how AI can be applied to clinical decision-making and reducing the societal burden of healthcare through science."},{"start_ms":1920000,"title":"The Importance of Co-design","summary":"Why hardware and software developers need visibility into future model generations to prevent architectural mismatches."},{"start_ms":2220000,"title":"The AMP Compute Grid Vision","summary":"How AMP uses excess compute to support non-profits and universities while building a scalable infrastructure model."},{"start_ms":3000000,"title":"Mission Alignment and Culture","summary":"A discussion on maintaining organizational integrity and mission-driven values in the face of rapid scaling and mercenary incentives."}],"topics":["AI Infrastructure","Model FLOPs Utilization","GPU Scaling","Compute Grids","Systems Engineering","Decentralized Computing","Machine Learning Operations","Hardware-Software Co-design"],"duration_seconds":3565,"processing_state":"processed","actions":[{"name":"request_transcript","method":"POST","url":"https://stenobird.com/v1/public/podcasts/latent-space-ai-engineer/episodes/the-professor-of-outputmaxxing-anjney-midha-amp/transcription-requests","description":"Idempotently request low-priority transcript generation for this episode."},{"name":"read_markdown","method":"GET","url":"https://stenobird.com/podcast/latent-space-ai-engineer/the-professor-of-outputmaxxing-anjney-midha-amp.md","description":"Read the agent-friendly Markdown representation of this episode resource."}]}}