# Lowering the Cost of Intelligence With NVIDIA's Ian Buck - Ep. 284 Page: https://stenobird.com/podcast/nvidia-ai-podcast/lowering-the-cost-of-intelligence-with-nvidia-s-ian-buck-ep-284 Text version: https://stenobird.com/podcast/nvidia-ai-podcast/lowering-the-cost-of-intelligence-with-nvidia-s-ian-buck-ep-284.md Podcast: [NVIDIA AI Podcast](https://stenobird.com/podcast/nvidia-ai-podcast) Published: 2025-12-29T16:57:00+00:00 Episode link: https://cohst.app/pdcst/9V4R8T/traffic.megaphone.fm/NVC3854335892.mp3 Audio file: https://cohst.app/pdcst/9V4R8T/traffic.megaphone.fm/NVC3854335892.mp3 Processing state: not_requested JSON: https://stenobird.com/v1/public/podcasts/nvidia-ai-podcast/episodes/lowering-the-cost-of-intelligence-with-nvidia-s-ian-buck-ep-284 Duration seconds: 2295 ## Resource Discover how mixture‑of‑experts (MoE) architecture is enabling smarter AI models without a proportional increase in the required compute and cost. Using vivid analogies and real-world examples, NVIDIA’s Ian Buck breaks down MoE models, their hidden complexities, and why extreme co-design across compute, networking, and software is essential to realizing their full potential. Learn more: https://blogs.nvidia.com/blog/mixture-of-experts-frontier-models/ ## Actions - request_transcript: `POST https://stenobird.com/v1/public/podcasts/nvidia-ai-podcast/episodes/lowering-the-cost-of-intelligence-with-nvidia-s-ian-buck-ep-284/transcription-requests` — Idempotently request low-priority transcript generation for this episode. - read_markdown: `GET https://stenobird.com/podcast/nvidia-ai-podcast/lowering-the-cost-of-intelligence-with-nvidia-s-ian-buck-ep-284.md` — Read the agent-friendly Markdown representation of this episode resource. A page view does not enqueue transcription. Agents should invoke `request_transcript` explicitly when they need this episode processed. ## Transcript Full transcripts are not published on public pages unless there is a clear rights basis.