Episode
"OpenAI Shares Some Alignment Problems" by Zvi
- Published
- Jul 22, 2026
- Duration seconds
- 1082
- Processing state
not_requested
Actions
POST https://stenobird.com/v1/public/podcasts/lesswrong-curated-popular-5643401/episodes/openai-shares-some-alignment-problems-by-zvi/transcription-requests
Idempotently request low-priority transcript generation for this episode.GET https://stenobird.com/podcast/lesswrong-curated-popular-5643401/openai-shares-some-alignment-problems-by-zvi.md
Read the agent-friendly Markdown representation of this episode resource.
Summary
Kudos to OpenAI for sharing their recent experiences with a misaligned internal model, where they encountered problems sufficiently severe they were forced to take the model offline to work on new mitigations and defense-to-depth. And also further kudos for actually taking the model offline for a time to build new safeguards. They gave us one hell of a candid report. The tone is professional throughout, whereas my reaction reading it was less professional and more this: Wit...