# GPT-5.6 Sol vs. Claude Fable: Why OpenAI’s new model crushes my benchmark Page: https://stenobird.com/podcast/how-i-ai-7304222/gpt-5-6-sol-vs-claude-fable-why-openai-s-new-model-crushes-my-benchmark Text version: https://stenobird.com/podcast/how-i-ai-7304222/gpt-5-6-sol-vs-claude-fable-why-openai-s-new-model-crushes-my-benchmark.md Podcast: [How I AI](https://stenobird.com/podcast/how-i-ai-7304222) Published: 2026-07-09T17:33:20+00:00 Episode link: https://podcasters.spotify.com/pod/show/pen-name/episodes/GPT-5-6-Sol-vs--Claude-Fable-Why-OpenAIs-new-model-crushes-my-benchmark-e3lrcne Audio file: https://anchor.fm/s/1035b1568/podcast/play/122581166/https%3A%2F%2Fd3ctxlq1ktw2nl.cloudfront.net%2Fstaging%2F2026-6-9%2F427622015-44100-2-39b3fe43deb86.mp3 Processing state: not_requested JSON: https://stenobird.com/v1/public/podcasts/how-i-ai-7304222/episodes/gpt-5-6-sol-vs-claude-fable-why-openai-s-new-model-crushes-my-benchmark Duration seconds: 2200 ## Resource GPT-5.6 Sol is back, and I ran it through my full How I AI vibe benchmark against GPT-5.6 Terra, Luna, Claude Fable 5, and Sonnet 5 across five categories: PRDs, prototypes, wireframes, debugging, and agentic voice. Sol won by a meaningful margin on my Claire Weighted Index (70% my taste, 30% Terminal Bench 2.1), and I also tested two use cases I can't stop thinking about: building a gamified homework tracking app for my kids in one shot with Codex, and browser automation with Chrome that burned through 500 LinkedIn replies while I did literally nothing. What you’ll learn: How I scored five AI models (including GPT 5.6 Sol, Fable 5, and Sonnet 5) using my “Claire Weighted Index” benchmark across PRDs, prototypes, code, and agentic voice The difference between GPT-5.6 Sol (Terra) and Sol for PRD writing How Fable’s precision and pedantry made it harder to collaborate with, and the exact moment Sol broke through where Fable got stuck Why Sonnet 5 is still my go-to for agentic voice in OpenClaw, even after this whole benchmark How I used GPT-5.6 Sol in Codex to build a fully gamified homework tracking app for my kids in one shot The video editing use case that saved me hours clipping a talk I gave at Cursor’s event How to use Codex plus GPT-5.6 and Chrome for browser automation, and why this is my single most-loved use case right now — In this episode, I cover: (00:00) Intro (01:10) The three GPT-5.6 models: Sol, Terra, Luna (02:17) Pricing: Sol vs. Fable API costs (03:24) The How I AI benchmark (05:03) Claire-weighted Index results (07:00) Per-task winners: prototypes, PRDs, agentic voice (11:59) What Claire actually rewards (13:20) Full-fidelity prototype side-by-sides (Sol vs. Fable) (17:45) Wireframes (18:19) Agentic voice (19:15) Where Sol is better than other models… ## Actions - request_transcript: `POST https://stenobird.com/v1/public/podcasts/how-i-ai-7304222/episodes/gpt-5-6-sol-vs-claude-fable-why-openai-s-new-model-crushes-my-benchmark/transcription-requests` — Idempotently request low-priority transcript generation for this episode. - read_markdown: `GET https://stenobird.com/podcast/how-i-ai-7304222/gpt-5-6-sol-vs-claude-fable-why-openai-s-new-model-crushes-my-benchmark.md` — Read the agent-friendly Markdown representation of this episode resource. A page view does not enqueue transcription. Agents should invoke `request_transcript` explicitly when they need this episode processed. ## Transcript Full transcripts are not published on public pages unless there is a clear rights basis.