# Tiny model, huge benchmarks & Million-token open-source coding model - AI News (Jun 18, 2026) Page: https://stenobird.com/podcast/the-automated-daily-ai-news-edition-6657064/tiny-model-huge-benchmarks-million-token-open-source-coding-model-ai-news-jun-18-2026 Text version: https://stenobird.com/podcast/the-automated-daily-ai-news-edition-6657064/tiny-model-huge-benchmarks-million-token-open-source-coding-model-ai-news-jun-18-2026.md Podcast: [The Automated Daily - AI News Edition](https://stenobird.com/podcast/the-automated-daily-ai-news-edition-6657064) Published: 2026-06-18T12:53:18+00:00 Episode link: https://theautomateddaily.com/episodes/2026-06-18-tiny-model-huge-benchmarks-million-token-open-source-coding-model Audio file: https://dts.podtrac.com/redirect.mp3/cdn.theautomateddaily.com/audio/hn-ai/2026-06-18/en/episode.mp3 Processing state: not_requested JSON: https://stenobird.com/v1/public/podcasts/the-automated-daily-ai-news-edition-6657064/episodes/tiny-model-huge-benchmarks-million-token-open-source-coding-model-ai-news-jun-18-2026 Duration seconds: 657 ## Resource Please support this podcast by checking out our sponsors: - Prezi: Create AI presentations fast - https://try.prezi.com/automated_daily - KrispCall: Agentic Cloud Telephony - https://try.krispcall.com/tad - Lindy is your ultimate AI assistant that proactively manages your inbox - https://try.lindy.ai/tad Support The Automated Daily directly: Buy me a coffee: https://buymeacoffee.com/theautomateddaily Today's topics: Tiny model, huge benchmarks - Sina Weibo’s VibeThinker-3B posts standout reasoning scores (AIME 2026) despite only 3B parameters, fueling debate about benchmark validity, post-training, and real-world reliability. Million-token open-source coding model - Z.ai releases GLM-5.2 under an MIT license, targeting stable 1M-token context for long-horizon coding agents, with new training focused on messy, hours-long engineering workflows. Agent tooling inside the browser - OpenAI adds Chrome DevTools Protocol support to Codex browser-use, letting agents read console logs, network traffic, and page state—key for debugging web apps with AI assistance. Voice AI gets truly interactive - OpenAI is reportedly preparing a new bidirectional voice model (GPT-Bidi-1) designed for natural interruptions and real-time conversation, pushing voice toward a primary AI interface. Anthropic pauses agent billing shift - Anthropic pauses its planned token-based billing shift for the Claude Agent SDK after developer backlash, highlighting rising sensitivity around agent usage costs and pricing models. Windows local AI on RTX - Microsoft experiments with running Phi Silica locally on Windows using Nvidia RTX GPUs, expanding on-device AI development beyond NPUs while exposing uneven feature tiers across hardware. NVIDIA Blackwell tops MLPerf - NVIDIA’s Blackwell platform leads MLPerf Tra… ## Actions - request_transcript: `POST https://stenobird.com/v1/public/podcasts/the-automated-daily-ai-news-edition-6657064/episodes/tiny-model-huge-benchmarks-million-token-open-source-coding-model-ai-news-jun-18-2026/transcription-requests` — Idempotently request low-priority transcript generation for this episode. - read_markdown: `GET https://stenobird.com/podcast/the-automated-daily-ai-news-edition-6657064/tiny-model-huge-benchmarks-million-token-open-source-coding-model-ai-news-jun-18-2026.md` — Read the agent-friendly Markdown representation of this episode resource. A page view does not enqueue transcription. Agents should invoke `request_transcript` explicitly when they need this episode processed. ## Transcript Full transcripts are not published on public pages unless there is a clear rights basis.