# Interview #84 Hagay Lupesko, SVP of AI Inference at Cerebras Systems Page: https://stenobird.com/podcast/the-artificial-intelligence-podcast-3635848/interview-84-hagay-lupesko-svp-of-ai-inference-at-cerebras-systems Text version: https://stenobird.com/podcast/the-artificial-intelligence-podcast-3635848/interview-84-hagay-lupesko-svp-of-ai-inference-at-cerebras-systems.md Podcast: [The Artificial Intelligence Podcast](https://stenobird.com/podcast/the-artificial-intelligence-podcast-3635848) Published: 2026-04-02T13:02:32+00:00 Episode link: https://podcasters.spotify.com/pod/show/tonyphoang/episodes/Interview-84-Hagay-Lupesko--SVP-of-AI-Inference-at-Cerebras-Systems-e3hb03d Audio file: https://traffic.megaphone.fm/APO3670023582.mp3 Processing state: not_requested JSON: https://stenobird.com/v1/public/podcasts/the-artificial-intelligence-podcast-3635848/episodes/interview-84-hagay-lupesko-svp-of-ai-inference-at-cerebras-systems Duration seconds: 2854 ## Resource Join Hagay Lupesko, SVP of AI Inference at Cerebras Systems, for a deep dive into the rapidly evolving world of AI inference. Hagay breaks down why inference has overtaken training as the dominant AI workload, how Cerebras' wafer-scale chip architecture delivers 10-20x faster performance than NVIDIA GPUs, and why CUDA is no longer the moat many think it is. He also covers how DeepSeek wiping $600 billion off NVIDIA's market cap in a single day was both a foundational and deeply misunderstood moment for the industry, the growing energy crisis in AI infrastructure, and what it will take to support the explosive rise of AI agents in the enterprise. ## Actions - request_transcript: `POST https://stenobird.com/v1/public/podcasts/the-artificial-intelligence-podcast-3635848/episodes/interview-84-hagay-lupesko-svp-of-ai-inference-at-cerebras-systems/transcription-requests` — Idempotently request low-priority transcript generation for this episode. - read_markdown: `GET https://stenobird.com/podcast/the-artificial-intelligence-podcast-3635848/interview-84-hagay-lupesko-svp-of-ai-inference-at-cerebras-systems.md` — Read the agent-friendly Markdown representation of this episode resource. A page view does not enqueue transcription. Agents should invoke `request_transcript` explicitly when they need this episode processed. ## Transcript Full transcripts are not published on public pages unless there is a clear rights basis.