Episode

Inference Just Got Cheaper. The Market Panicked.

Podcast
YPO Technology Network AI Brief
Published
Jun 25, 2026
Duration seconds
592
Processing state
not_requested
Canonical source
https://rss.com/podcasts/ypo-technology-network-ai-brief/2940872
Audio
https://content.rss.com/episodes/382927/2940872/ypo-technology-network-ai-brief/2026_06_24_14_47_33_f7ecae30-9ef8-4151-bbc0-af46ccbdf693.mp3
JSON
/v1/public/podcasts/ypo-technology-network-ai-brief-7728971/episodes/inference-just-got-cheaper-the-market-panicked
Markdown
/podcast/ypo-technology-network-ai-brief-7728971/inference-just-got-cheaper-the-market-panicked.md

Actions

  • POST https://stenobird.com/v1/public/podcasts/ypo-technology-network-ai-brief-7728971/episodes/inference-just-got-cheaper-the-market-panicked/transcription-requests
    Idempotently request low-priority transcript generation for this episode.
  • GET https://stenobird.com/podcast/ypo-technology-network-ai-brief-7728971/inference-just-got-cheaper-the-market-panicked.md
    Read the agent-friendly Markdown representation of this episode resource.

Summary

OpenAI unveiled its first custom chip the same week the market sold off on fears the AI buildout has gone too far. Stephen Forte argues those are the same story told from opposite ends — and that what looks like a bubble is closer to a re-pricing. In this episode: OpenAI's "Jalapeno" chip — built with Broadcom, purpose-made for inference, roughly 50% more cost-efficient than standard AI GPUs in early tests, designed in nine months, deploying at gigawatt scale by year-end. The selloff — Nasdaq off about 2.2%, Nvidia down roughly 4%, Alphabet's worst day in over a year, on AI-buildout cost fears, rate jitters, and a memory-chip wobble. Why it is a re-pricing, not a bubble — the cost of inference has fallen about 10x a year for three years; software efficiencies like Mixture-of-Experts compound on hardware gains, so the buildout grows but not in a straight line. What it means for operators — roughly 80% of workflows will run on small, local models inside your own network; only the highest-reasoning work needs the frontier cloud. Anthropic's Claude Tag — an always-on Claude teammate in Slack, and a live example of the new workloads that cheaper inference unlocks. The YPO Technology Network AI Brief is a daily briefing on the AI news that matters to CEOs and senior operators, hosted by Stephen Forte.