Episode

The Future of AI Infrastructure with CoreWeave

Podcast
Practical AI
Published
Jul 17, 2026
Duration seconds
3004
Processing state
processed
Canonical source
https://share.transistor.fm/s/01c30767
Audio
https://pscrb.fm/rss/p/dts.podtrac.com/redirect.mp3/media.transistor.fm/01c30767/eed885fc.mp3
JSON
/v1/public/podcasts/practical-ai/episodes/the-future-of-ai-infrastructure-with-coreweave
Markdown
/podcast/practical-ai/the-future-of-ai-infrastructure-with-coreweave.md

Actions

  • POST https://stenobird.com/v1/public/podcasts/practical-ai/episodes/the-future-of-ai-infrastructure-with-coreweave/transcription-requests
    Idempotently request low-priority transcript generation for this episode.
  • GET https://stenobird.com/podcast/practical-ai/the-future-of-ai-infrastructure-with-coreweave.md
    Read the agent-friendly Markdown representation of this episode resource.

Summary

Traditional cloud computing is insufficient for the specialized demands of modern AI workloads. Corey Sanders explains why the industry is shifting toward AI-native infrastructure to support complex, multi-model agentic applications.

Topics

  • AI Infrastructure
  • GPU Computing
  • Cloud Computing
  • Agentic AI
  • CoreWeave
  • Machine Learning Operations
  • Inference Optimization
  • Edge Computing

Highlights

  • Main idea: AI-native infrastructure is required to handle the unique compute and storage demands of training and inference workloads
  • Practical takeaway: Developers should move toward 'infrastructure-aware' applications that optimize for caching and specialized model routing
  • Failure mode: Using massive frontier models for simple tasks like spell-checking is inefficient and creates unnecessary cost overhead
  • Main idea: The future of software lies in agentic workflows where multiple specialized models interact rather than single-purpose web interfaces
  • Practical takeaway: Democratizing access to high-performance GPU clusters is essential for enterprises lacking frontier-lab research talent

Chapters

  1. 1:00 The Evolution of Cloud to AI: Corey Sanders compares the current AI boom to the early days of Azure, noting the parallel growth patterns in infrastructure needs.
  2. 5:00 The Gap in AI Applications: An exploration of what is currently missing in the stack to support the transition from simple prompts to complex, business-critical AI applications.
  3. 16:00 The Art of GPU Orchestration: Discussing the technical difficulty of managing tens of thousands of GPUs and the importance of identifying performance bottlenecks in large clusters.
  4. 23:00 The Rise of Agentic Workflows: How backend infrastructure will evolve to support a web of interacting agents and specialized models performing specific tasks.
  5. 34:00 Portability vs. Performance: The tension between maintaining cloud-agnostic portability and the need for deep hardware optimization for maximum efficiency.
  6. 45:00 The Death of the Traditional Website: Predicting a future where AI-first experiences replace button-based web interfaces with fluid, human-interaction models.