# How Meta Is Using Its Own Chip to Cut AI Inference Costs Page: https://stenobird.com/podcast/the-ai-podcast-with-fexingo-artificial-intelligence-machine-learning-and-modern-ai-models-7872011/how-meta-is-using-its-own-chip-to-cut-ai-inference-costs Text version: https://stenobird.com/podcast/the-ai-podcast-with-fexingo-artificial-intelligence-machine-learning-and-modern-ai-models-7872011/how-meta-is-using-its-own-chip-to-cut-ai-inference-costs.md Podcast: [The AI Podcast with Fexingo: Artificial Intelligence, Machine Learning, and Modern AI Models](https://stenobird.com/podcast/the-ai-podcast-with-fexingo-artificial-intelligence-machine-learning-and-modern-ai-models-7872011) Published: 2026-07-09T20:32:16+00:00 Episode link: https://audio.fexingo.com/business/the-ai-podcast/episode-0101.mp3 Audio file: https://audio.fexingo.com/business/the-ai-podcast/episode-0101.mp3 Processing state: not_requested JSON: https://stenobird.com/v1/public/podcasts/the-ai-podcast-with-fexingo-artificial-intelligence-machine-learning-and-modern-ai-models-7872011/episodes/how-meta-is-using-its-own-chip-to-cut-ai-inference-costs Duration seconds: 618 ## Resource Meta just released Muse Spark 1.1, its latest AI coding model, but the bigger story is how the company is running inference on its custom Meta Training and Inference Accelerator chip. In this episode, Lucas and Luna break down the cost savings of owning your own inference silicon, why Broadcom's 11 percent weekly surge is connected to custom chip deals, and how the math of AI inference is reshaping the hardware landscape. They explain the difference between general-purpose GPUs from NVIDIA and AMD versus custom ASICs like Meta's MTIA, and what that means for the next wave of AI deployment at scale. If you follow AI hardware, this is the angle the headlines are missing. #Meta #MuseSpark #AICoding #CustomChip #MTIA #AIInference #Broadcom #NVIDIA #ASIC #Hardware #Technology #TechPodcast #ArtificialIntelligence #ChipDesign #InferenceCosts #FexingoBusiness #BusinessPodcast #AIPodcast Keep every episode free: buymeacoffee.com/fexingo ## Actions - request_transcript: `POST https://stenobird.com/v1/public/podcasts/the-ai-podcast-with-fexingo-artificial-intelligence-machine-learning-and-modern-ai-models-7872011/episodes/how-meta-is-using-its-own-chip-to-cut-ai-inference-costs/transcription-requests` — Idempotently request low-priority transcript generation for this episode. - read_markdown: `GET https://stenobird.com/podcast/the-ai-podcast-with-fexingo-artificial-intelligence-machine-learning-and-modern-ai-models-7872011/how-meta-is-using-its-own-chip-to-cut-ai-inference-costs.md` — Read the agent-friendly Markdown representation of this episode resource. A page view does not enqueue transcription. Agents should invoke `request_transcript` explicitly when they need this episode processed. ## Transcript Full transcripts are not published on public pages unless there is a clear rights basis.