# 📅 May 28 - Opus 4.8 ships mid-show, the Pope writes 42K words on AI, 11labs dubs the world and DeepSwe breaks coding evals Page: https://stenobird.com/podcast/thursdai-the-top-ai-news-from-the-past-week-6519604/may-28-opus-4-8-ships-mid-show-the-pope-writes-42k-words-on-ai-11labs-dubs-the-world-and-deepswe-breaks-coding-evals Text version: https://stenobird.com/podcast/thursdai-the-top-ai-news-from-the-past-week-6519604/may-28-opus-4-8-ships-mid-show-the-pope-writes-42k-words-on-ai-11labs-dubs-the-world-and-deepswe-breaks-coding-evals.md Podcast: [ThursdAI - The top AI news from the past week](https://stenobird.com/podcast/thursdai-the-top-ai-news-from-the-past-week-6519604) Published: 2026-05-29T00:23:57+00:00 Episode link: https://sub.thursdai.news/p/may-28-opus-48-ships-mid-show-the Audio file: https://prfx.byspotify.com/e/api.substack.com/feed/podcast/199665761/cc143aafd8642b56d08ec4c83e74a4f7.mp3 Processing state: not_requested JSON: https://stenobird.com/v1/public/podcasts/thursdai-the-top-ai-news-from-the-past-week-6519604/episodes/may-28-opus-4-8-ships-mid-show-the-pope-writes-42k-words-on-ai-11labs-dubs-the-world-and-deepswe-breaks-coding-evals Duration seconds: 5951 ## Resource Hey folks, this is Alex, let me catch you up! First, Opus 4.8 dropped during the show, we immediately tested it, read on for our initial reviews. Also, we dedicated a heavy chunk of the show today to cover Pope Leo XIV’s encyclical letter on AI called “Magnifica Humanitas” and talked about a new bench called DeepSWE . And then, just after the show, both ElevenLabs and Cartesia dropped released that honestly blew my mind, and I don’t get my mind blown often. I got so excited that I had to record a video on it (instead of writing the newsletter, so sorry if it’s a bit later today). Plus, a few open source models and Microsoft surprises as #3 on Image Arena with MAI Image 2.5! Crazy week, let’s get into it! ThursdAI - Highest signal weekly AI news show is a reader-supported publication. To receive new posts and support my work, consider becoming a free or paid subscriber. Big CO LLMs + APIs Anthropic ships Claude Opus 4.8, live during the show ( blog , system card ) Let me get into the big one. Halfway through the episode, Opus 4.8 went live, so we read the blog and the system card in real time (and I got to press the big “breaking news” button!) Anthropic frames it as their most capable model for ambitious work. It does not claim to beat their unreleased Mythos preview, but the numbers are strong anyway. SWE-bench Pro is at 69.2% , up from 64.3% on Opus 4.7 and ahead of GPT-5.5 at 58.6%. Humanity’s Last Exam is the new best score at 49.8% without tools and 57.9% with tools. OSWorld-Verified (computer use) lands at 83.4%. The one place it loses is Terminal-Bench 2.1, where GPT-5.5 still wins 78.2 to 74.6. Wolfram made a good point here: Terminal-Bench is time-limited, so cranking the thinking level can actually hurt the score, because you burn the clock thinking instead o… ## Actions - request_transcript: `POST https://stenobird.com/v1/public/podcasts/thursdai-the-top-ai-news-from-the-past-week-6519604/episodes/may-28-opus-4-8-ships-mid-show-the-pope-writes-42k-words-on-ai-11labs-dubs-the-world-and-deepswe-breaks-coding-evals/transcription-requests` — Idempotently request low-priority transcript generation for this episode. - read_markdown: `GET https://stenobird.com/podcast/thursdai-the-top-ai-news-from-the-past-week-6519604/may-28-opus-4-8-ships-mid-show-the-pope-writes-42k-words-on-ai-11labs-dubs-the-world-and-deepswe-breaks-coding-evals.md` — Read the agent-friendly Markdown representation of this episode resource. A page view does not enqueue transcription. Agents should invoke `request_transcript` explicitly when they need this episode processed. ## Transcript Full transcripts are not published on public pages unless there is a clear rights basis.