# Episode 63: Why Gemini 3 Will Change How You Build AI Agents with Ravin Kumar (Google DeepMind) Page: https://stenobird.com/podcast/vanishing-gradients-4989163/episode-63-why-gemini-3-will-change-how-you-build-ai-agents-with-ravin-kumar-google-deepmind Text version: https://stenobird.com/podcast/vanishing-gradients-4989163/episode-63-why-gemini-3-will-change-how-you-build-ai-agents-with-ravin-kumar-google-deepmind.md Podcast: [Vanishing Gradients](https://stenobird.com/podcast/vanishing-gradients-4989163) Published: 2025-11-22T07:30:00+00:00 Episode link: https://hugobowne.substack.com/p/episode-63-why-gemini-3-will-change-2fd Audio file: https://api.substack.com/feed/podcast/181324505/66a93c30730a3672690bdd89ce4aefe2.mp3 Processing state: not_requested JSON: https://stenobird.com/v1/public/podcasts/vanishing-gradients-4989163/episodes/episode-63-why-gemini-3-will-change-how-you-build-ai-agents-with-ravin-kumar-google-deepmind Duration seconds: 3613 ## Resource Gemini 3 is a few days old and the massive leap in performance and model reasoning has big implications for builders: as models begin to self-heal, builders are literally tearing out the functionality they built just months ago... ripping out the defensive coding and reshipping their agent harnesses entirely. Ravin Kumar (Google DeepMind) joins Hugo to breaks down exactly why the rapid evolution of models like Gemini 3 is changing how we build software. They detail the shift from simple tool calling to building reliable "Agent Harnesses", explore the architectural tradeoffs between deterministic workflows and high-agency systems, the nuance of preventing context rot in massive windows, and why proper evaluation infrastructure is the only way to manage the chaos of autonomous loops. They talk through: - The implications of models that can "self-heal" and fix their own code - The two cultures of agents: LLM workflows with a few tools versus when you should unleash high-agency, autonomous systems. - Inside NotebookLM: moving from prototypes to viral production features like Audio Overviews - Why Needle in a Haystack benchmarks often fail to predict real-world performance - How to build agent harnesses that turn model capabilities into product velocity - The shift from measuring latency to managing time-to-compute for reasoning tasks LINKS From Context Engineering to AI Agent Harnesses: The New Software Discipline, a podcast Hugo did with Lance Martin, LangChain ( https://high-signal.delphina.ai/episode/context-engineering-to-ai-agent-harnesses-the-new-software-discipline ) Context Rot: How Increasing Input Tokens Impacts LLM Performance ( https://research.trychroma.com/context-rot ) Effective context engineering for AI agents by Anthropic ( https://www.anthropic.com/engin… ## Actions - request_transcript: `POST https://stenobird.com/v1/public/podcasts/vanishing-gradients-4989163/episodes/episode-63-why-gemini-3-will-change-how-you-build-ai-agents-with-ravin-kumar-google-deepmind/transcription-requests` — Idempotently request low-priority transcript generation for this episode. - read_markdown: `GET https://stenobird.com/podcast/vanishing-gradients-4989163/episode-63-why-gemini-3-will-change-how-you-build-ai-agents-with-ravin-kumar-google-deepmind.md` — Read the agent-friendly Markdown representation of this episode resource. A page view does not enqueue transcription. Agents should invoke `request_transcript` explicitly when they need this episode processed. ## Transcript Full transcripts are not published on public pages unless there is a clear rights basis.