Episode

GPT 5.5 just did what no other model could

Podcast
How I AI
Published
Apr 23, 2026
Duration seconds
1416
Processing state
not_requested
Canonical source
https://podcasters.spotify.com/pod/show/pen-name/episodes/GPT-5-5-just-did-what-no-other-model-could-e3ic88s
Audio
https://anchor.fm/s/1035b1568/podcast/play/118939356/https%3A%2F%2Fd3ctxlq1ktw2nl.cloudfront.net%2Fstaging%2F2026-3-23%2F422726716-44100-2-dfde17f6eb81.mp3
JSON
/v1/public/podcasts/how-i-ai-7304222/episodes/gpt-5-5-just-did-what-no-other-model-could
Markdown
/podcast/how-i-ai-7304222/gpt-5-5-just-did-what-no-other-model-could.md

Actions

  • POST https://stenobird.com/v1/public/podcasts/how-i-ai-7304222/episodes/gpt-5-5-just-did-what-no-other-model-could/transcription-requests
    Idempotently request low-priority transcript generation for this episode.
  • GET https://stenobird.com/podcast/how-i-ai-7304222/gpt-5-5-just-did-what-no-other-model-could.md
    Read the agent-friendly Markdown representation of this episode resource.

Summary

In this mini episode, I break down OpenAI’s new GPT 5.5 and GPT 5.5 Pro after weeks of early testing. I walk through three real jobs I threw at the model: building an app for me to teach my second grader more advanced subtraction concepts, tackling a tech debt problem in the ChatPRD codebase, and hacking into a proprietary Bluetooth pixel display that every other model had failed me on. My verdict: higher intelligence, better efficiency, and genuinely autonomous long-running loops that change what I think is worth tackling. What you’ll learn: How I think about GPT 5.5 Pro’s pricing vs engineering time, and when I believe the “intelligence tax” is worth paying Why I treat GPT 5.5 as a developer model first, and why I couldn’t find a consumer use case that justified its intelligence The exact prompt pattern I use to unlock a long-running autonomous subagent loop How I got a near-six-hour autonomous run to one-shot 98% of edge cases in a migration over millions of chat threads and drop my Sentry error rate to the floor Why I’m now throwing GPT 5.5 at tech debt, flaky tests, and security backlogs first How I combined a Bluetooth packet sniffer and GPT 5.5 to reverse-engineer a proprietary pixel speaker after Claude Code and GPT 5.4 both gave up How I use the /personality command inside Codex to swap the default “baked potato” tone for something I actually enjoy working with — In this episode, I cover: (00:00) Introduction to GPT 5.5 testing (00:40) What is GPT 5.5 and how much does it cost? (03:23) Testing GPT 5.5 in ChatGPT: the intelligence overhang problem (07:12) Moving to Codex: where GPT 5.5 really shines (16:01) Hacking a Chinese Bluetooth speaker (21:47) Final thoughts on GPT 5.5’s intelligence and efficiency — Tools referenced: • GPT 5.5 and GPT 5.5 Pro: https://o…