# The Two-Billion-Token Weekend That Cost $125 (For Now) Page: https://stenobird.com/podcast/ai-at-work-7563679/the-two-billion-token-weekend-that-cost-125-for-now Text version: https://stenobird.com/podcast/ai-at-work-7563679/the-two-billion-token-weekend-that-cost-125-for-now.md Podcast: [AI at Work](https://stenobird.com/podcast/ai-at-work-7563679) Published: 2026-09-03T19:59:15+00:00 Episode link: https://api.riverside.com/hosting-analytics/media/3ed7f23025199c476eebd41e15fcb688b714ff262b78c96877ea9d76766f91a0/eyJlcGlzb2RlSWQiOiIxZTIwZGM1OS0xZTI4LTQxMzAtYjVkYS1mYjFkNjExNWU4ZDUiLCJwb2RjYXN0SWQiOiJlM2QxODk0ZC05OGZjLTQwMDItOTgwYi04NzU1MWU5YjYzYjkiLCJhY2NvdW50SWQiOiI2MDdjNWM5MmU4YjZhMjQ3ZDYzZTNjZTUiLCJwYXRoIjoibWVkaWEvY2xpcHMvNmE5OWI5MGY0MjM0N2IwODFlMWVkMTg2L2FpLWF0LXdvcmstMjAyNi05LTNfXzE4LTE0LTM5Lm1wMyJ9.mp3 Audio file: https://api.riverside.com/hosting-analytics/media/3ed7f23025199c476eebd41e15fcb688b714ff262b78c96877ea9d76766f91a0/eyJlcGlzb2RlSWQiOiIxZTIwZGM1OS0xZTI4LTQxMzAtYjVkYS1mYjFkNjExNWU4ZDUiLCJwb2RjYXN0SWQiOiJlM2QxODk0ZC05OGZjLTQwMDItOTgwYi04NzU1MWU5YjYzYjkiLCJhY2NvdW50SWQiOiI2MDdjNWM5MmU4YjZhMjQ3ZDYzZTNjZTUiLCJwYXRoIjoibWVkaWEvY2xpcHMvNmE5OWI5MGY0MjM0N2IwODFlMWVkMTg2L2FpLWF0LXdvcmstMjAyNi05LTNfXzE4LTE0LTM5Lm1wMyJ9.mp3 Processing state: not_requested JSON: https://stenobird.com/v1/public/podcasts/ai-at-work-7563679/episodes/the-two-billion-token-weekend-that-cost-125-for-now Duration seconds: 2808 ## Resource Kevin spent a rainy weekend running up to 10 simultaneous Codex agents across three internal builds and burned two billion tokens. On a consumer plan, propped up by repeated usage resets, that cost him about $125. At retail API rates on GPT 5.6 high, the same weekend runs north of $8,000 to $10,000 and enterprise teams building in Codex fall outside the consumer plan entirely. Matt and Kevin dig into what that gap means for anyone standing up internal AI infrastructure today. Matt's team built an internal system that captures every call, Slack message, and email on a client project and flags when something drifts. It runs affordably now. His concern isn't the build cost it's what happens if the run cost triples. They also cover why switching between Claude Code and Codex costs you almost nothing if your code lives in GitHub, why that undercuts the case for price increases, and what happened when Kevin turned a swarm loose on 45 repositories to fix a favicon. Topics: - Why AI avatars still fail for self-cloning after three years of trying - The 82/18 split between deterministic and agentic workflow steps - LinkedIn bans, bot-detectable IPs, and 200 profile lookups in five minutes - Codex sidebar chats, work trees, and browser control versus Claude Code - Two billion tokens, $125, and the $8,000 enterprise equivalent - The runaway agent that logged into an abandoned Lovable account Ready to find out where your organization actually stands on AI readiness? https://assessment.ascendlabs.ai/ Want to talk through your AI tool decisions before you commit? tidycal.com/kevinwilliams CHECK OUT KEVIN’S STUFF: Website: https://ascendlabs.ai/ LinkedIn: https://www.google.com/search?q=https ... – Free Tool: Take the https://assessment.ascendlabs.ai/ Deep Dive: Read the https://resou… ## Actions - request_transcript: `POST https://stenobird.com/v1/public/podcasts/ai-at-work-7563679/episodes/the-two-billion-token-weekend-that-cost-125-for-now/transcription-requests` — Idempotently request low-priority transcript generation for this episode. - read_markdown: `GET https://stenobird.com/podcast/ai-at-work-7563679/the-two-billion-token-weekend-that-cost-125-for-now.md` — Read the agent-friendly Markdown representation of this episode resource. A page view does not enqueue transcription. Agents should invoke `request_transcript` explicitly when they need this episode processed. ## Transcript Full transcripts are not published on public pages unless there is a clear rights basis.