{"podcast":{"title":"The AI Podcast with Fexingo: Artificial Intelligence, Machine Learning, and Modern AI Models","slug":"the-ai-podcast-with-fexingo-artificial-intelligence-machine-learning-and-modern-ai-models-7872011","podcast_index_feed_id":7872011,"rss_url":"https://feeds.fexingo.com/business/the-ai-podcast.xml","website_url":"https://www.fexingo.com/","image_url":"https://audio.fexingo.com/business/the-ai-podcast/cover.png","author":"Fexingo","episode_count":104,"summary":"Lucas and Luna dissect the week in artificial intelligence — not the hype, but the actual models, benchmarks, and deployment decisions shaping the industry. Each episode anchors on a specific paper, product launch, or policy move: from Mixture-of-Experts architecture changes to EU AI Act enforcement, from OpenAI's governance restructuring to open-weight model licensing battles. They compare LLM benchmark scores across reasoning, coding, and multilingual tasks, examine inference cost curves per million tokens, and trace how foundation model competition affects downstream startups. Lucas, a journalist covering tech policy, brings the regulatory and competitive landscape; Luna, an ML engineer turned product lead, presses on technical tradeoffs and real-world performance. Together they avoid speculation and focus on data: what the latest Nvidia GPU cluster means for training efficiency, why a particular transformer variant reduced latency by 40%, or how retrieval-augmented generation changes enterprise search ROI. The show serves engineers, product managers, and investors who need to separate signal from noise in AI. If you want to understand why one billion-parameter model beats anot…","last_synced_at":"2026-07-11T14:18:53.849941+00:00","page_url":"https://stenobird.com/podcast/the-ai-podcast-with-fexingo-artificial-intelligence-machine-learning-and-modern-ai-models-7872011"},"episode":{"title":"Why AI Model Inference Is Splitting Into Two Markets","slug":"why-ai-model-inference-is-splitting-into-two-markets","published_at":"2026-07-06T20:28:22+00:00","page_url":"https://stenobird.com/podcast/the-ai-podcast-with-fexingo-artificial-intelligence-machine-learning-and-modern-ai-models-7872011/why-ai-model-inference-is-splitting-into-two-markets","show_page_url":"https://stenobird.com/podcast/the-ai-podcast-with-fexingo-artificial-intelligence-machine-learning-and-modern-ai-models-7872011","url":"https://audio.fexingo.com/business/the-ai-podcast/episode-0095.mp3","audio_url":"https://audio.fexingo.com/business/the-ai-podcast/episode-0095.mp3","summary":"Lucas and Luna explore the surprising divergence in AI infrastructure as model inference splits into two distinct markets: high-throughput batch processing for enterprise and low-latency real-time inference for consumer apps. They discuss how companies like Together AI and Anthropic are betting on custom silicon, why NVIDIA's data center revenue mix is shifting, and what the recent sell-off in equipment makers like ASML and Applied Materials signals for the next wave of AI deployment. The hosts also unpack a recent interview with Vercel's CEO on the fight to separate models from agents, and what it means for startups building on top of foundation models. #AIInference #ModelSplitting #NVIDIA #TogetherAI #Anthropic #Vercel #ASML #AppliedMaterials #CustomSilicon #AIInfrastructure #LatencyVsThroughput #RealTimeAI #BatchProcessing #AIStartups #Technology #FexingoBusiness #BusinessPodcast #AIPodcast Keep every episode free: buymeacoffee.com/fexingo","meta_description":"Lucas and Luna explore the surprising divergence in AI infrastructure as model inference splits into two distinct markets: high-throughput batch processin…","key_points":[],"chapters":[],"topics":[],"duration_seconds":551,"processing_state":"not_requested","actions":[{"name":"request_transcript","method":"POST","url":"https://stenobird.com/v1/public/podcasts/the-ai-podcast-with-fexingo-artificial-intelligence-machine-learning-and-modern-ai-models-7872011/episodes/why-ai-model-inference-is-splitting-into-two-markets/transcription-requests","description":"Idempotently request low-priority transcript generation for this episode."},{"name":"read_markdown","method":"GET","url":"https://stenobird.com/podcast/the-ai-podcast-with-fexingo-artificial-intelligence-machine-learning-and-modern-ai-models-7872011/why-ai-model-inference-is-splitting-into-two-markets.md","description":"Read the agent-friendly Markdown representation of this episode resource."}]}}