Episode
Interview #84 Hagay Lupesko, SVP of AI Inference at Cerebras Systems
- Published
- Apr 2, 2026
- Duration seconds
- 2854
- Processing state
not_requested
Actions
POST https://stenobird.com/v1/public/podcasts/the-artificial-intelligence-podcast-3635848/episodes/interview-84-hagay-lupesko-svp-of-ai-inference-at-cerebras-systems/transcription-requests
Idempotently request low-priority transcript generation for this episode.GET https://stenobird.com/podcast/the-artificial-intelligence-podcast-3635848/interview-84-hagay-lupesko-svp-of-ai-inference-at-cerebras-systems.md
Read the agent-friendly Markdown representation of this episode resource.
Summary
Join Hagay Lupesko, SVP of AI Inference at Cerebras Systems, for a deep dive into the rapidly evolving world of AI inference. Hagay breaks down why inference has overtaken training as the dominant AI workload, how Cerebras' wafer-scale chip architecture delivers 10-20x faster performance than NVIDIA GPUs, and why CUDA is no longer the moat many think it is. He also covers how DeepSeek wiping $600 billion off NVIDIA's market cap in a single day was both a foundational and deeply misunderstood moment for the industry, the growing energy crisis in AI infrastructure, and what it will take to support the explosive rise of AI agents in the enterprise.