Episode
Red-Teaming after Mythos — Zico Kolter & Matt Fredrikson, Gray Swan
- Published
- Jun 22, 2026
- Duration seconds
- 3983
- Processing state
not_requested- Canonical source
- https://www.latent.space/p/gray-swan
Actions
POST https://stenobird.com/v1/public/podcasts/latent-space-ai-engineer/episodes/red-teaming-after-mythos-zico-kolter-matt-fredrikson-gray-swan/transcription-requests
Idempotently request low-priority transcript generation for this episode.GET https://stenobird.com/podcast/latent-space-ai-engineer/red-teaming-after-mythos-zico-kolter-matt-fredrikson-gray-swan.md
Read the agent-friendly Markdown representation of this episode resource.
Summary
AI Engineer World’s Fair regular bird tix will sell out ~today! Join us next week ahead of the Late Bird price hike and get >$40,000 in sponsor credits for attending ! Thanks to the US Government issuing an export control directive on Mythos and Fable , the risks of jailbreaks and (industry term) indirect prompt injection are suddenly the talk of the town, though we have been covering AI security for a few years now, from Hackaprompt to the enigmatic Pliny the Elder . Zico Kolter, member of OpenAI’s board of directors on the Safety & Security Committee , and Matt Fredrikson, CMU professor and CEO of Gray Swan , co-authored the definitive paper on Indirect Prompt Injections , and Gray Swan were cited authorities on the Mythos model card , directly investigating the exact capabilities that are under scrutiny right now: We seized the opportunity to ask them the state of AI Red Teaming, and Shade , the adversarial red teaming tool that Anthropic used to evaluate the robustness of their models against prompt injection attacks in coding environments. Shade is part of their overall toolkit covering Simon Willison’s Lethal Trifecta , including Cygnal , an AI guardrails product, and the world’s largest AI Red Teaming Arena , including AIRT celebrity Wyatt Walls . All of this security tooling, and yet, we’re only staving off the inevitable. The risks of extremely smart AI increasingly feel like gray swan events: an event that everyone can see coming. In this episode, Gray Swan cofounders Zico Kolter and Matt Fredrikson join swyx to explain why AI security is not just “cybersecurity with AI,” why agents introduce a new class of vulnerabilities, and why the next major AI incident may be a gray swan: unlikely, but clearly visible before it happens. We go deep on prompt injection ,…