{"podcast":{"title":"The Robot Brains Podcast","slug":"robot-brains-podcast","podcast_index_feed_id":2136067,"rss_url":"https://feeds.acast.com/public/shows/the-robot-brains","website_url":"https://shows.acast.com/the-robot-brains","image_url":"https://assets.pippa.io/shows/6053a29a0d11b0148adcfc96/1617091773845-808cfbe200b5e99e0a9deac35e5f0dab.jpeg","author":"The Robot Brains Podcast","episode_count":67,"summary":"In each episode of The Robot Brains podcast, renowned artificial intelligence researcher, professor and entrepreneur Pieter Abbeel meets the brilliant minds attempting to build robots with brains. Pieter is joined by leading experts in AI Robotics from all over the world as he explores how far humanity has come in its mission to create conscious computers, mindful machines and rational robots. Host: Pieter Abbeel | Executive Producers: Alice Patel & Henry Tobias Jones | Audio Production: Kieron Matthew Banerji | Title Music: Alejandro Del Pozo Hosted on Acast. See acast.com/privacy for more information.","last_synced_at":null,"page_url":"https://stenobird.com/podcast/robot-brains-podcast"},"episode":{"title":"Noam Brown: from Open AI on solving Poker and Diplomacy with AI","slug":"noam-brown-from-open-ai-on-solving-poker-and-diplomacy-with-ai","published_at":"2023-06-28T20:14:45+00:00","page_url":"https://stenobird.com/podcast/robot-brains-podcast/noam-brown-from-open-ai-on-solving-poker-and-diplomacy-with-ai","show_page_url":"https://stenobird.com/podcast/robot-brains-podcast","url":"https://www.therobotbrains.ai/who-is-noam-brown","audio_url":"https://sphinx.acast.com/p/open/s/6053a29a0d11b0148adcfc96/e/649c94b6ab8b53001141f804/media.mp3","summary":"Noam Brown explains how AI can master imperfect information games like Poker and Diplomacy through strategic planning and self-play. He explores the transition from solving zero-sum games to navigating complex human-like negotiations.","meta_description":"AI researcher Noam Brown discusses the breakthroughs of Pluribus and Cicero, mastering poker, and the future of reasoning in large language models.","key_points":["Main idea: Games serve as natural, un-overfittable benchmarks for evaluating AI progress because they are easy to score and compare against human experts","Technical insight: Using entropy regularization helps ensure that subgame equilibria remain consistent with the original game's equilibrium","Failure mode: The success of reinforcement learning in Chess and Go does not automatically translate to games with hidden information or high variance","Practical takeaway: Future breakthroughs in AI reasoning may come from integrating planning and strategic anticipation into large language models","Research insight: Effective AI in Diplomacy requires managing complex, human-like dialogue and negotiating trust among multiple players"],"chapters":[{"start_ms":400000,"title":"The Challenge of Imperfect Information","summary":"Comparing the asymmetry of information in poker to the perfect information found in chess and Go."},{"start_ms":1115000,"title":"Anticipation and Future States","summary":"How AI uses planning to anticipate opponent moves and evaluate possible future game states."},{"start_ms":1800000,"title":"Achieving Equilibrium via Regularization","summary":"A technical look at using entropy regularization to solve equilibrium problems in complex subgames."},{"start_ms":2130000,"title":"Beating Human Professionals","summary":"The story behind Libratus and the difficulty of scaling poker AI from heads-up to multi-player formats."},{"start_ms":3160000,"title":"Limitations of Reinforcement Learning","summary":"Why the paradigms used for Chess and Go fail when applied to games with high uncertainty."},{"start_ms":3480000,"title":"AI in Diplomacy and Dialogue","summary":"Using language to facilitate human-like negotiation and strategic communication in complex games."},{"start_ms":3840000,"title":"The Future of LLM Reasoning","summary":"Exploring the intersection of large language models and strategic planning capabilities."}],"topics":["Artificial Intelligence","Game Theory","Poker AI","Large Language Models","Reinforcement Learning","Diplomacy AI","Strategic Planning","Imperfect Information Games"],"duration_seconds":4478,"processing_state":"processed","actions":[{"name":"request_transcript","method":"POST","url":"https://stenobird.com/v1/public/podcasts/robot-brains-podcast/episodes/noam-brown-from-open-ai-on-solving-poker-and-diplomacy-with-ai/transcription-requests","description":"Idempotently request low-priority transcript generation for this episode."},{"name":"read_markdown","method":"GET","url":"https://stenobird.com/podcast/robot-brains-podcast/noam-brown-from-open-ai-on-solving-poker-and-diplomacy-with-ai.md","description":"Read the agent-friendly Markdown representation of this episode resource."}]}}