# The misaligned incentives behind AI coding agents | Russell Kaplan, Cognition Page: https://stenobird.com/podcast/max-agency-7807113/the-misaligned-incentives-behind-ai-coding-agents-russell-kaplan-cognition Text version: https://stenobird.com/podcast/max-agency-7807113/the-misaligned-incentives-behind-ai-coding-agents-russell-kaplan-cognition.md Podcast: [Max Agency](https://stenobird.com/podcast/max-agency-7807113) Published: 2026-07-30T14:00:00+00:00 Episode link: https://podcasters.spotify.com/pod/show/supermix2/episodes/The-misaligned-incentives-behind-AI-coding-agents--Russell-Kaplan--Cognition-e3mnkq5 Audio file: https://anchor.fm/s/1112ba400/podcast/play/123506949/https%3A%2F%2Fd3ctxlq1ktw2nl.cloudfront.net%2Fstaging%2F2026-6-29%2F428869237-44100-2-5ff31799bb19c.mp3 Processing state: not_requested JSON: https://stenobird.com/v1/public/podcasts/max-agency-7807113/episodes/the-misaligned-incentives-behind-ai-coding-agents-russell-kaplan-cognition Duration seconds: 3016 ## Resource Russell Kaplan is the President of Cognition, the $26 billion company behind Devin, the AI software engineer. He started his career in machine learning at Tesla's Autopilot team, then led the ML org at Scale AI, which had acquired his computer vision startup Helia. In this conversation, Russell unpacks why more and more engineering work is becoming "intelligence saturated," how the sidekick architecture inside Devin Fusion cuts cost without cutting quality, and why Cognition is underwriting a $10 million productivity guarantee. – We also discuss: Why mergeability is the next eval bar When to use Fable vs GPT models How smarter routing unlocked 35% better price performance Why letting users pick their own model is a UX bug Wiring agents to be proactive instead of reactive How Cognition scores a session productive or unproductive – Timestamps: (00:00) Introduction (01:21) Launching Devin at 13% on SWE-bench (02:47) When Devin became the number one committer to Devin (05:56) Intelligence saturation: when speed and cost start to matter more (07:05) When to use Fable vs GPT models (08:41) Why mergeability is the next eval bar (10:51) How open source maintainers shaped FrontierCode (14:56) The telltale sign of agent-written code (17:56) When per-person token spend starts to eclipse salaries (20:26) Why letting users pick their own model is a UX bug (21:46) How smarter routing unlocked 35% better price performance (23:50) Why Cognition still trains its own models (27:13) Deploying Cerebras at scale at 950 tokens per second (28:21) The Tesla chip debate: does an inference chip need division? (30:06) Wiring agents to be proactive instead of reactive (32:19) Why "software factory" is the wrong term (35:16) Why everyone is now a new grad (41:10) Forward deploy… ## Actions - request_transcript: `POST https://stenobird.com/v1/public/podcasts/max-agency-7807113/episodes/the-misaligned-incentives-behind-ai-coding-agents-russell-kaplan-cognition/transcription-requests` — Idempotently request low-priority transcript generation for this episode. - read_markdown: `GET https://stenobird.com/podcast/max-agency-7807113/the-misaligned-incentives-behind-ai-coding-agents-russell-kaplan-cognition.md` — Read the agent-friendly Markdown representation of this episode resource. A page view does not enqueue transcription. Agents should invoke `request_transcript` explicitly when they need this episode processed. ## Transcript Full transcripts are not published on public pages unless there is a clear rights basis.