Episode
How ZML Is Unlocking Free AI Inference Across Chips
- Published
- Jul 8, 2026
- Duration seconds
- 604
- Processing state
not_requested
Actions
POST https://stenobird.com/v1/public/podcasts/the-ai-podcast-with-fexingo-artificial-intelligence-machine-learning-and-modern-ai-models-7872011/episodes/how-zml-is-unlocking-free-ai-inference-across-chips/transcription-requests
Idempotently request low-priority transcript generation for this episode.GET https://stenobird.com/podcast/the-ai-podcast-with-fexingo-artificial-intelligence-machine-learning-and-modern-ai-models-7872011/how-zml-is-unlocking-free-ai-inference-across-chips.md
Read the agent-friendly Markdown representation of this episode resource.
Summary
A tiny French startup called ZML just released a free tool that promises to speed up AI inference across a wide range of chips — including NVIDIA's H100, AMD's MI300X, and even Intel's Gaudi 3. Lucas and Luna unpack why this matters right now, as the AI hardware market splits into winners and losers. They discuss how ZML's approach differs from NVIDIA's proprietary CUDA ecosystem, what the stock moves this week tell us about the market's anxiety over chip dependence, and whether a free software layer could actually reshape the $50 billion AI chip industry. Along the way, they touch on SambaNova's latest $1 billion raise, Meta's Muse Image controversy, and the broader question of whether open-source inference tools are a threat or an opportunity for the big chip makers. If you've been following the AI chip narrative and wondering why AMD and Intel are down while Meta is up, this episode connects the dots. #ZML #AIInference #OpenSourceAI #NVIDIA #AMD #Intel #SambaNova #Meta #CUDAMoat #AIChipMarket #InferenceOptimization #TechPodcast #Technology #FexingoBusiness #BusinessPodcast #AIPodcast #MachineLearning #ChipDiversification Keep every episode free: buymeacoffee.com/fexingo