Episode
Lowering the Cost of Intelligence With NVIDIA's Ian Buck - Ep. 284
- Podcast
- NVIDIA AI Podcast
- Published
- Dec 29, 2025
- Duration seconds
- 2295
- Processing state
not_requested
Actions
POST https://stenobird.com/v1/public/podcasts/nvidia-ai-podcast/episodes/lowering-the-cost-of-intelligence-with-nvidia-s-ian-buck-ep-284/transcription-requests
Idempotently request low-priority transcript generation for this episode.GET https://stenobird.com/podcast/nvidia-ai-podcast/lowering-the-cost-of-intelligence-with-nvidia-s-ian-buck-ep-284.md
Read the agent-friendly Markdown representation of this episode resource.
Summary
Discover how mixture‑of‑experts (MoE) architecture is enabling smarter AI models without a proportional increase in the required compute and cost. Using vivid analogies and real-world examples, NVIDIA’s Ian Buck breaks down MoE models, their hidden complexities, and why extreme co-design across compute, networking, and software is essential to realizing their full potential. Learn more: https://blogs.nvidia.com/blog/mixture-of-experts-frontier-models/