{"podcast":{"title":"The AI Podcast with Fexingo: Artificial Intelligence, Machine Learning, and Modern AI Models","slug":"the-ai-podcast-with-fexingo-artificial-intelligence-machine-learning-and-modern-ai-models-7872011","podcast_index_feed_id":7872011,"rss_url":"https://feeds.fexingo.com/business/the-ai-podcast.xml","website_url":"https://www.fexingo.com/","image_url":"https://audio.fexingo.com/business/the-ai-podcast/cover.png","author":"Fexingo","episode_count":104,"summary":"Lucas and Luna dissect the week in artificial intelligence — not the hype, but the actual models, benchmarks, and deployment decisions shaping the industry. Each episode anchors on a specific paper, product launch, or policy move: from Mixture-of-Experts architecture changes to EU AI Act enforcement, from OpenAI's governance restructuring to open-weight model licensing battles. They compare LLM benchmark scores across reasoning, coding, and multilingual tasks, examine inference cost curves per million tokens, and trace how foundation model competition affects downstream startups. Lucas, a journalist covering tech policy, brings the regulatory and competitive landscape; Luna, an ML engineer turned product lead, presses on technical tradeoffs and real-world performance. Together they avoid speculation and focus on data: what the latest Nvidia GPU cluster means for training efficiency, why a particular transformer variant reduced latency by 40%, or how retrieval-augmented generation changes enterprise search ROI. The show serves engineers, product managers, and investors who need to separate signal from noise in AI. If you want to understand why one billion-parameter model beats anot…","last_synced_at":"2026-07-11T14:18:53.849941+00:00","page_url":"https://stenobird.com/podcast/the-ai-podcast-with-fexingo-artificial-intelligence-machine-learning-and-modern-ai-models-7872011"},"episode":{"title":"How Meta Is Using Its Own Chip to Cut AI Inference Costs","slug":"how-meta-is-using-its-own-chip-to-cut-ai-inference-costs","published_at":"2026-07-09T20:32:16+00:00","page_url":"https://stenobird.com/podcast/the-ai-podcast-with-fexingo-artificial-intelligence-machine-learning-and-modern-ai-models-7872011/how-meta-is-using-its-own-chip-to-cut-ai-inference-costs","show_page_url":"https://stenobird.com/podcast/the-ai-podcast-with-fexingo-artificial-intelligence-machine-learning-and-modern-ai-models-7872011","url":"https://audio.fexingo.com/business/the-ai-podcast/episode-0101.mp3","audio_url":"https://audio.fexingo.com/business/the-ai-podcast/episode-0101.mp3","summary":"Meta just released Muse Spark 1.1, its latest AI coding model, but the bigger story is how the company is running inference on its custom Meta Training and Inference Accelerator chip. In this episode, Lucas and Luna break down the cost savings of owning your own inference silicon, why Broadcom's 11 percent weekly surge is connected to custom chip deals, and how the math of AI inference is reshaping the hardware landscape. They explain the difference between general-purpose GPUs from NVIDIA and AMD versus custom ASICs like Meta's MTIA, and what that means for the next wave of AI deployment at scale. If you follow AI hardware, this is the angle the headlines are missing. #Meta #MuseSpark #AICoding #CustomChip #MTIA #AIInference #Broadcom #NVIDIA #ASIC #Hardware #Technology #TechPodcast #ArtificialIntelligence #ChipDesign #InferenceCosts #FexingoBusiness #BusinessPodcast #AIPodcast Keep every episode free: buymeacoffee.com/fexingo","meta_description":"Meta just released Muse Spark 1.1, its latest AI coding model, but the bigger story is how the company is running inference on its custom Meta Training an…","key_points":[],"chapters":[],"topics":[],"duration_seconds":618,"processing_state":"not_requested","actions":[{"name":"request_transcript","method":"POST","url":"https://stenobird.com/v1/public/podcasts/the-ai-podcast-with-fexingo-artificial-intelligence-machine-learning-and-modern-ai-models-7872011/episodes/how-meta-is-using-its-own-chip-to-cut-ai-inference-costs/transcription-requests","description":"Idempotently request low-priority transcript generation for this episode."},{"name":"read_markdown","method":"GET","url":"https://stenobird.com/podcast/the-ai-podcast-with-fexingo-artificial-intelligence-machine-learning-and-modern-ai-models-7872011/how-meta-is-using-its-own-chip-to-cut-ai-inference-costs.md","description":"Read the agent-friendly Markdown representation of this episode resource."}]}}