Episode
The Future of AI Infrastructure with CoreWeave
- Podcast
- Practical AI
- Published
- Jul 17, 2026
- Duration seconds
- 3004
- Processing state
processed- Canonical source
- https://share.transistor.fm/s/01c30767
Actions
POST https://stenobird.com/v1/public/podcasts/practical-ai/episodes/the-future-of-ai-infrastructure-with-coreweave/transcription-requests
Idempotently request low-priority transcript generation for this episode.GET https://stenobird.com/podcast/practical-ai/the-future-of-ai-infrastructure-with-coreweave.md
Read the agent-friendly Markdown representation of this episode resource.
Summary
Traditional cloud computing is insufficient for the specialized demands of modern AI workloads. Corey Sanders explains why the industry is shifting toward AI-native infrastructure to support complex, multi-model agentic applications.
Topics
- AI Infrastructure
- GPU Computing
- Cloud Computing
- Agentic AI
- CoreWeave
- Machine Learning Operations
- Inference Optimization
- Edge Computing
Highlights
- Main idea: AI-native infrastructure is required to handle the unique compute and storage demands of training and inference workloads
- Practical takeaway: Developers should move toward 'infrastructure-aware' applications that optimize for caching and specialized model routing
- Failure mode: Using massive frontier models for simple tasks like spell-checking is inefficient and creates unnecessary cost overhead
- Main idea: The future of software lies in agentic workflows where multiple specialized models interact rather than single-purpose web interfaces
- Practical takeaway: Democratizing access to high-performance GPU clusters is essential for enterprises lacking frontier-lab research talent
Chapters
1:00The Evolution of Cloud to AI: Corey Sanders compares the current AI boom to the early days of Azure, noting the parallel growth patterns in infrastructure needs.5:00The Gap in AI Applications: An exploration of what is currently missing in the stack to support the transition from simple prompts to complex, business-critical AI applications.16:00The Art of GPU Orchestration: Discussing the technical difficulty of managing tens of thousands of GPUs and the importance of identifying performance bottlenecks in large clusters.23:00The Rise of Agentic Workflows: How backend infrastructure will evolve to support a web of interacting agents and specialized models performing specific tasks.34:00Portability vs. Performance: The tension between maintaining cloud-agnostic portability and the need for deep hardware optimization for maximum efficiency.45:00The Death of the Traditional Website: Predicting a future where AI-first experiences replace button-based web interfaces with fluid, human-interaction models.