Episode

Lowering the Cost of Intelligence With NVIDIA's Ian Buck - Ep. 284

Podcast
NVIDIA AI Podcast
Published
Dec 29, 2025
Duration seconds
2295
Processing state
not_requested
Canonical source
https://cohst.app/pdcst/9V4R8T/traffic.megaphone.fm/NVC3854335892.mp3
Audio
https://cohst.app/pdcst/9V4R8T/traffic.megaphone.fm/NVC3854335892.mp3
JSON
/v1/public/podcasts/nvidia-ai-podcast/episodes/lowering-the-cost-of-intelligence-with-nvidia-s-ian-buck-ep-284
Markdown
/podcast/nvidia-ai-podcast/lowering-the-cost-of-intelligence-with-nvidia-s-ian-buck-ep-284.md

Actions

  • POST https://stenobird.com/v1/public/podcasts/nvidia-ai-podcast/episodes/lowering-the-cost-of-intelligence-with-nvidia-s-ian-buck-ep-284/transcription-requests
    Idempotently request low-priority transcript generation for this episode.
  • GET https://stenobird.com/podcast/nvidia-ai-podcast/lowering-the-cost-of-intelligence-with-nvidia-s-ian-buck-ep-284.md
    Read the agent-friendly Markdown representation of this episode resource.

Summary

Discover how mixture‑of‑experts (MoE) architecture is enabling smarter AI models without a proportional increase in the required compute and cost. Using vivid analogies and real-world examples, NVIDIA’s Ian Buck breaks down MoE models, their hidden complexities, and why extreme co-design across compute, networking, and software is essential to realizing their full potential. Learn more: https://blogs.nvidia.com/blog/mixture-of-experts-frontier-models/