{"podcast":{"title":"The AI Podcast with Fexingo: Artificial Intelligence, Machine Learning, and Modern AI Models","slug":"the-ai-podcast-with-fexingo-artificial-intelligence-machine-learning-and-modern-ai-models-7872011","podcast_index_feed_id":7872011,"rss_url":"https://feeds.fexingo.com/business/the-ai-podcast.xml","website_url":"https://www.fexingo.com/","image_url":"https://audio.fexingo.com/business/the-ai-podcast/cover.png","author":"Fexingo","episode_count":104,"summary":"Lucas and Luna dissect the week in artificial intelligence — not the hype, but the actual models, benchmarks, and deployment decisions shaping the industry. Each episode anchors on a specific paper, product launch, or policy move: from Mixture-of-Experts architecture changes to EU AI Act enforcement, from OpenAI's governance restructuring to open-weight model licensing battles. They compare LLM benchmark scores across reasoning, coding, and multilingual tasks, examine inference cost curves per million tokens, and trace how foundation model competition affects downstream startups. Lucas, a journalist covering tech policy, brings the regulatory and competitive landscape; Luna, an ML engineer turned product lead, presses on technical tradeoffs and real-world performance. Together they avoid speculation and focus on data: what the latest Nvidia GPU cluster means for training efficiency, why a particular transformer variant reduced latency by 40%, or how retrieval-augmented generation changes enterprise search ROI. The show serves engineers, product managers, and investors who need to separate signal from noise in AI. If you want to understand why one billion-parameter model beats anot…","last_synced_at":"2026-07-11T14:18:53.849941+00:00","page_url":"https://stenobird.com/podcast/the-ai-podcast-with-fexingo-artificial-intelligence-machine-learning-and-modern-ai-models-7872011"},"episode":{"title":"How ZML Is Unlocking Free AI Inference Across Chips","slug":"how-zml-is-unlocking-free-ai-inference-across-chips","published_at":"2026-07-08T09:59:44+00:00","page_url":"https://stenobird.com/podcast/the-ai-podcast-with-fexingo-artificial-intelligence-machine-learning-and-modern-ai-models-7872011/how-zml-is-unlocking-free-ai-inference-across-chips","show_page_url":"https://stenobird.com/podcast/the-ai-podcast-with-fexingo-artificial-intelligence-machine-learning-and-modern-ai-models-7872011","url":"https://audio.fexingo.com/business/the-ai-podcast/episode-0098.mp3","audio_url":"https://audio.fexingo.com/business/the-ai-podcast/episode-0098.mp3","summary":"A tiny French startup called ZML just released a free tool that promises to speed up AI inference across a wide range of chips — including NVIDIA's H100, AMD's MI300X, and even Intel's Gaudi 3. Lucas and Luna unpack why this matters right now, as the AI hardware market splits into winners and losers. They discuss how ZML's approach differs from NVIDIA's proprietary CUDA ecosystem, what the stock moves this week tell us about the market's anxiety over chip dependence, and whether a free software layer could actually reshape the $50 billion AI chip industry. Along the way, they touch on SambaNova's latest $1 billion raise, Meta's Muse Image controversy, and the broader question of whether open-source inference tools are a threat or an opportunity for the big chip makers. If you've been following the AI chip narrative and wondering why AMD and Intel are down while Meta is up, this episode connects the dots. #ZML #AIInference #OpenSourceAI #NVIDIA #AMD #Intel #SambaNova #Meta #CUDAMoat #AIChipMarket #InferenceOptimization #TechPodcast #Technology #FexingoBusiness #BusinessPodcast #AIPodcast #MachineLearning #ChipDiversification Keep every episode free: buymeacoffee.com/fexingo","meta_description":"A tiny French startup called ZML just released a free tool that promises to speed up AI inference across a wide range of chips — including NVIDIA's H100,…","key_points":[],"chapters":[],"topics":[],"duration_seconds":604,"processing_state":"not_requested","actions":[{"name":"request_transcript","method":"POST","url":"https://stenobird.com/v1/public/podcasts/the-ai-podcast-with-fexingo-artificial-intelligence-machine-learning-and-modern-ai-models-7872011/episodes/how-zml-is-unlocking-free-ai-inference-across-chips/transcription-requests","description":"Idempotently request low-priority transcript generation for this episode."},{"name":"read_markdown","method":"GET","url":"https://stenobird.com/podcast/the-ai-podcast-with-fexingo-artificial-intelligence-machine-learning-and-modern-ai-models-7872011/how-zml-is-unlocking-free-ai-inference-across-chips.md","description":"Read the agent-friendly Markdown representation of this episode resource."}]}}