Episode

Ep 84: OpenAI’s Chief Scientist on Continual Learning Hype, RL Beyond Code, & Future Alignment Directions

Podcast
Unsupervised Learning with Jacob Effron
Published
Apr 9, 2026
Duration seconds
3526
Processing state
not_requested
Canonical source
https://unsupervised-learning.simplecast.com/episodes/ep-84-openais-chief-scientist-on-continual-learning-hype-rl-beyond-code-future-alignment-directions-l43Yra90
Audio
https://cdn.simplecast.com/media/audio/transcoded/b3414ac6-61c8-4752-8722-491e1457c3bf/2c08ad29-5b79-42c0-a40a-6c1af4327f2f/episodes/audio/group/e877d39c-130c-4301-a631-aa7ef0a93964/group-item/8e0caf31-dcb2-4cdd-9f8a-78f2a86c7a87/128_default_tc.mp3?aid=rss_feed&feed=dOSE_bdP
JSON
/v1/public/podcasts/unsupervised-learning-with-jacob-effron-6041643/episodes/ep-84-openai-s-chief-scientist-on-continual-learning-hype-rl-beyond-code-future-alignment-directions
Markdown
/podcast/unsupervised-learning-with-jacob-effron-6041643/ep-84-openai-s-chief-scientist-on-continual-learning-hype-rl-beyond-code-future-alignment-directions.md

Actions

  • POST https://stenobird.com/v1/public/podcasts/unsupervised-learning-with-jacob-effron-6041643/episodes/ep-84-openai-s-chief-scientist-on-continual-learning-hype-rl-beyond-code-future-alignment-directions/transcription-requests
    Idempotently request low-priority transcript generation for this episode.
  • GET https://stenobird.com/podcast/unsupervised-learning-with-jacob-effron-6041643/ep-84-openai-s-chief-scientist-on-continual-learning-hype-rl-beyond-code-future-alignment-directions.md
    Read the agent-friendly Markdown representation of this episode resource.

Summary

Jakub Pachocki, OpenAI's Chief Scientist, sits down with Jacob to cover the full arc of where AI research stands today and where it's headed. The conversation spans the explosive growth of coding agents and what it signals about near-term AI capability, the use of math and physics benchmarks as proxies for general intelligence, how reinforcement learning is being extended beyond easily-verified domains toward longer-horizon tasks, and what it means to run a research organization at the precise moment the models themselves are starting to accelerate the research. Jakub shares a candid take on the competitive landscape, why chain-of-thought monitoring is one of the most promising tools in the alignment toolkit, and — with unusual directness — why the concentration of power enabled by highly automated AI organizations is a societal problem that doesn't yet have an obvious solution.