Episode

NEW GLM 5.2 BEATS Claude?

Podcast
AI News Today | Julian Goldie Podcast
Published
Jun 16, 2026
Duration seconds
542
Processing state
not_requested
Canonical source
https://podcasters.spotify.com/pod/show/julian-goldie9/episodes/NEW-GLM-5-2-BEATS-Claude-e3ksj4v
Audio
https://anchor.fm/s/10b0edd94/podcast/play/121571935/https%3A%2F%2Fd3ctxlq1ktw2nl.cloudfront.net%2Fstaging%2F2026-5-16%2F426266228-44100-2-c73a247d902c5.mp3
JSON
/v1/public/podcasts/ai-news-today-julian-goldie-podcast-7573784/episodes/new-glm-5-2-beats-claude
Markdown
/podcast/ai-news-today-julian-goldie-podcast-7573784/new-glm-5-2-beats-claude.md

Actions

  • POST https://stenobird.com/v1/public/podcasts/ai-news-today-julian-goldie-podcast-7573784/episodes/new-glm-5-2-beats-claude/transcription-requests
    Idempotently request low-priority transcript generation for this episode.
  • GET https://stenobird.com/podcast/ai-news-today-julian-goldie-podcast-7573784/new-glm-5-2-beats-claude.md
    Read the agent-friendly Markdown representation of this episode resource.

Summary

GLM 52 vs Qwen 37 Max vs Claude Opus 48: Real-World Tests vs Benchmarks (No Second Chances)The episode compares GLM 52 (ZAI), Qwen 37 Max (Alibaba), and Claude Opus 48 (Anthropic) head-to-head on five one-shot tasks, arguing that benchmark rankings didn’t match real usability. In coding-focused tests like a voxel runner game, a liquid-in-a-bowl animation, a business landing page, and an arcade game, GLM 52 produced the most fun, polished, and feature-rich results, while Claude’s outputs were often basic and Qwen’s were sometimes buggy or incomplete; Claude clearly won the solar-system orbit map task. The script also notes Qwen’s strong reported benchmarks and faster replies, GLM’s slower responses in agents but strong CLI coding, and highlights limitations integrating Claude into agent workflows compared to Qwen/GLM in Hermes and the creator’s agent operating system. 00:00 Head To Head Setup 01:27 Coding Tests Results 04:09 Arcade Game Showdown 04:50 Benchmarks Versus Reality 06:01 Agents Workflow Tradeoffs 07:59 Final Recommendations