# Inference Just Got Cheaper. The Market Panicked. Page: https://stenobird.com/podcast/ypo-technology-network-ai-brief-7728971/inference-just-got-cheaper-the-market-panicked Text version: https://stenobird.com/podcast/ypo-technology-network-ai-brief-7728971/inference-just-got-cheaper-the-market-panicked.md Podcast: [YPO Technology Network AI Brief](https://stenobird.com/podcast/ypo-technology-network-ai-brief-7728971) Published: 2026-06-25T10:00:00+00:00 Episode link: https://rss.com/podcasts/ypo-technology-network-ai-brief/2940872 Audio file: https://content.rss.com/episodes/382927/2940872/ypo-technology-network-ai-brief/2026_06_24_14_47_33_f7ecae30-9ef8-4151-bbc0-af46ccbdf693.mp3 Processing state: not_requested JSON: https://stenobird.com/v1/public/podcasts/ypo-technology-network-ai-brief-7728971/episodes/inference-just-got-cheaper-the-market-panicked Duration seconds: 592 ## Resource OpenAI unveiled its first custom chip the same week the market sold off on fears the AI buildout has gone too far. Stephen Forte argues those are the same story told from opposite ends — and that what looks like a bubble is closer to a re-pricing. In this episode: OpenAI's "Jalapeno" chip — built with Broadcom, purpose-made for inference, roughly 50% more cost-efficient than standard AI GPUs in early tests, designed in nine months, deploying at gigawatt scale by year-end. The selloff — Nasdaq off about 2.2%, Nvidia down roughly 4%, Alphabet's worst day in over a year, on AI-buildout cost fears, rate jitters, and a memory-chip wobble. Why it is a re-pricing, not a bubble — the cost of inference has fallen about 10x a year for three years; software efficiencies like Mixture-of-Experts compound on hardware gains, so the buildout grows but not in a straight line. What it means for operators — roughly 80% of workflows will run on small, local models inside your own network; only the highest-reasoning work needs the frontier cloud. Anthropic's Claude Tag — an always-on Claude teammate in Slack, and a live example of the new workloads that cheaper inference unlocks. The YPO Technology Network AI Brief is a daily briefing on the AI news that matters to CEOs and senior operators, hosted by Stephen Forte. ## Actions - request_transcript: `POST https://stenobird.com/v1/public/podcasts/ypo-technology-network-ai-brief-7728971/episodes/inference-just-got-cheaper-the-market-panicked/transcription-requests` — Idempotently request low-priority transcript generation for this episode. - read_markdown: `GET https://stenobird.com/podcast/ypo-technology-network-ai-brief-7728971/inference-just-got-cheaper-the-market-panicked.md` — Read the agent-friendly Markdown representation of this episode resource. A page view does not enqueue transcription. Agents should invoke `request_transcript` explicitly when they need this episode processed. ## Transcript Full transcripts are not published on public pages unless there is a clear rights basis.