Episode

DeepSeek: 2 Months Out

Podcast
muckrAIkers
Published
Apr 9, 2025
Duration seconds
5491
Processing state
not_requested
Canonical source
https://kairos.fm/e012/
Audio
https://op3.dev/e/media.transistor.fm/b3bf778f/0f502957.mp3
JSON
/v1/public/podcasts/muckraikers-7026051/episodes/deepseek-2-months-out
Markdown
/podcast/muckraikers-7026051/deepseek-2-months-out.md

Actions

  • POST https://stenobird.com/v1/public/podcasts/muckraikers-7026051/episodes/deepseek-2-months-out/transcription-requests
    Idempotently request low-priority transcript generation for this episode.
  • GET https://stenobird.com/podcast/muckraikers-7026051/deepseek-2-months-out.md
    Read the agent-friendly Markdown representation of this episode resource.

Summary

DeepSeek has been out for over 2 months now, and things have begun to settle down. We take this opportunity to contextualize the developments that have occurred in its wake, both within the AI industry and the world economy. As systems get more "agentic" and users are willing to spend increasing amounts of time waiting for their outputs, the value of supposed "reasoning" models continues to be peddled by AI system developers, but does the data really back these claims? Check out our DeepSeek minisode for a snappier overview! EPISODE RECORDED 2025.03.30 (00:40) - DeepSeek R1 recap (02:46) - What makes it new? (08:53) - What is reasoning? (14:51) - Limitations of reasoning models (why we hate reasoning) (31:16) - Claims about R1 training on Open AI (37:30) - “Deep Research” (49:13) - Developments and drama in the AI industry (56:26) - Proposed economic value (01:14:20) - US government involvement (01:23:28) - OpenAI uses MCP (01:28:15) - Outro Links DeepSeek website DeepSeek paper DeepSeek docs - Models and Pricing DeepSeek repo - 3FS Understanding DeepSeek/DeepResearch Explainers Language Models & Co. article - The Illustrated DeepSeek-R1 Towards Data Science article - DeepSeek-V3 Explained 1: Multi-head Latent Attention Jina.ai article - A Practical Guide to Implementing DeepSearch/DeepResearch Han, Not Solo blogpost - The Differences between Deep Research, Deep Research, and Deep Research Analysis and Research Preprint - Understanding R1-Zero-Like Training: A Critical Perspective Blogpost - There May Not be Aha Moment in R1-Zero-like Training — A Pilot Study Preprint - Large Language Monkeys: Scaling Inference Compute with Repeated Sampling Preprint - Chain-of-Thought Reasoning In The Wild Is Not Always Faithful Fallout coverage TechCrunch article - OpenAI calls D…