# Cheaper, Faster & Smarter: Qwen 3.8 27B, GLM 5.3, Grok 4.6, Cerebras Page: https://stenobird.com/podcast/generative-ai-meetup/cheaper-faster-smarter-qwen-3-8-27b-glm-5-3-grok-4-6-cerebras Text version: https://stenobird.com/podcast/generative-ai-meetup/cheaper-faster-smarter-qwen-3-8-27b-glm-5-3-grok-4-6-cerebras.md Podcast: [The Generative AI Meetup Podcast](https://stenobird.com/podcast/generative-ai-meetup) Published: 2026-08-18T05:40:52+00:00 Episode link: https://podcast.genaimeetup.com/e/cheaper-faster-smarter-qwen-38-27b-glm-53-grok-46-cerebras/ Audio file: https://mcdn.podbean.com/mf/web/i7trpajygvv62e3f/podcast-enhanced.mp3 Processing state: not_requested JSON: https://stenobird.com/v1/public/podcasts/generative-ai-meetup/episodes/cheaper-faster-smarter-qwen-3-8-27b-glm-5-3-grok-4-6-cerebras Duration seconds: 6577 ## Resource https://novacut.ai/ https://genaimeetup.com/ Jeff returns to the Gen AI Meetup Podcast for a wide-ranging discussion on where AI is heading—and why powerful models running on consumer hardware could change the economics of the entire industry. We dive into Qwen 3.8 27B and the growing viability of running capable LLMs locally, GLM 5.3 and the latest Chinese open-source models, DeepSeek, Gemini 3.7, Grok 4.6, Meta’s latest models, and OpenAI’s partnership with Cerebras for dramatically faster inference. We also discuss whether foundation models are becoming commodities, what that means for companies like OpenAI and Anthropic, and why more value may ultimately move to the application layer. Jeff shares how his team approaches AI in healthcare, including self-hosting, data sovereignty, classifiers, fine-tuning, and spec-driven development for building reliable AI-assisted software without accumulating a mountain of vibe-coded technical debt. Plus: Jeff Dean’s departure from Google, Discovery Loop, Stripe’s OpenRouter acquisition, Anthropic’s controversial AI-text watermarking experiments, and whether watermarking could affect model quality. Topics include: Qwen 3.8 27B, GLM 5.3, DeepSeek V4, Grok 4.6, Gemini 3.7, Cerebras, OpenAI, Anthropic, Meta, local LLMs, open-source AI, model commoditization, spec-driven development, AI healthcare, data sovereignty, AI coding agents, and model watermarking. ## Actions - request_transcript: `POST https://stenobird.com/v1/public/podcasts/generative-ai-meetup/episodes/cheaper-faster-smarter-qwen-3-8-27b-glm-5-3-grok-4-6-cerebras/transcription-requests` — Idempotently request low-priority transcript generation for this episode. - read_markdown: `GET https://stenobird.com/podcast/generative-ai-meetup/cheaper-faster-smarter-qwen-3-8-27b-glm-5-3-grok-4-6-cerebras.md` — Read the agent-friendly Markdown representation of this episode resource. A page view does not enqueue transcription. Agents should invoke `request_transcript` explicitly when they need this episode processed. ## Transcript Full transcripts are not published on public pages unless there is a clear rights basis.