# GPUs, Kubernetes & AI Infrastructure Realities Page: https://stenobird.com/podcast/virtually-speaking-podcast-174853/gpus-kubernetes-ai-infrastructure-realities Text version: https://stenobird.com/podcast/virtually-speaking-podcast-174853/gpus-kubernetes-ai-infrastructure-realities.md Podcast: [Virtually Speaking Podcast](https://stenobird.com/podcast/virtually-speaking-podcast-174853) Published: 2026-05-22T19:47:08+00:00 Episode link: https://www.vspeakingpodcast.com/e/gpus-kubernetes-ai-infrastructure-realities/ Audio file: https://mcdn.podbean.com/mf/web/2k2xptntz7uz6ztg/VSP-KUBECON-FRANK-V2.mp3 Processing state: not_requested JSON: https://stenobird.com/v1/public/podcasts/virtually-speaking-podcast-174853/episodes/gpus-kubernetes-ai-infrastructure-realities Duration seconds: 1110 ## Resource At KubeCon 2026, Pete Flecha and John Nicholson sit down with VMware by Broadcom’s Frank Denneman to explore one of the biggest infrastructure conversations happening in AI today: should Kubernetes workloads run on bare metal or virtualized infrastructure? The discussion dives deep into how AI workloads are changing infrastructure design, why Kubernetes and virtualization are becoming increasingly connected, and how technologies like DRS and Dynamic Resource Allocation (DRA) are evolving to support modern GPU-intensive environments. Frank explains the operational, security, and resource management challenges organizations face as AI adoption accelerates — especially when dealing with expensive GPU clusters, multi-tenant AI workloads, and the rise of AI agents. Topics include: Why virtualization still matters for Kubernetes and AI GPU scheduling, topology awareness, and resource isolation DRA (Dynamic Resource Allocation) in Kubernetes AI infrastructure efficiency and GPU utilization Security and isolation for AI agents and workloads Token governance and AI operational guardrails Lessons learned from decades of virtualization applied to AI infrastructure If you’re trying to understand where Kubernetes, virtualization, and AI infrastructure are headed next, this is a conversation you won’t want to miss. ## Actions - request_transcript: `POST https://stenobird.com/v1/public/podcasts/virtually-speaking-podcast-174853/episodes/gpus-kubernetes-ai-infrastructure-realities/transcription-requests` — Idempotently request low-priority transcript generation for this episode. - read_markdown: `GET https://stenobird.com/podcast/virtually-speaking-podcast-174853/gpus-kubernetes-ai-infrastructure-realities.md` — Read the agent-friendly Markdown representation of this episode resource. A page view does not enqueue transcription. Agents should invoke `request_transcript` explicitly when they need this episode processed. ## Transcript Full transcripts are not published on public pages unless there is a clear rights basis.