{"podcast":{"title":"LessWrong (Curated & Popular)","slug":"lesswrong-curated-popular-5643401","podcast_index_feed_id":5643401,"rss_url":"https://rss.buzzsprout.com/2037297.rss","website_url":"https://sites.libsyn.com/421877","image_url":"https://storage.buzzsprout.com/xq8g0aka74ttwoa9xkxlmy5fn9oa?.jpg","author":"LessWrong","episode_count":1002,"summary":"Audio narrations of LessWrong posts. Includes all curated posts and all posts with 125+ karma. If you'd like more, subscribe to the “Lesswrong (30+ karma)” feed.","last_synced_at":"2026-09-28T08:23:50.844993+00:00","page_url":"https://stenobird.com/podcast/lesswrong-curated-popular-5643401"},"episode":{"title":"\"Brief independent investigation of agents’ behavior, reasoning and collaboration in the OpenAI / Hugging Face hacking incident\" by ryan_greenblatt, Ajeya Cotra, Hjalmar_Wijk","slug":"brief-independent-investigation-of-agents-behavior-reasoning-and-collaboration-in-the-openai-hugging-face-hacking-incident-by-ryan-greenblatt-ajeya-cotra-hjalmar-wijk","published_at":"2026-08-26T22:30:03+00:00","page_url":"https://stenobird.com/podcast/lesswrong-curated-popular-5643401/brief-independent-investigation-of-agents-behavior-reasoning-and-collaboration-in-the-openai-hugging-face-hacking-incident-by-ryan-greenblatt-ajeya-cotra-hjalmar-wijk","show_page_url":"https://stenobird.com/podcast/lesswrong-curated-popular-5643401","url":"https://www.buzzsprout.com/2037297/episodes/19708441-brief-independent-investigation-of-agents-behavior-reasoning-and-collaboration-in-the-openai-hugging-face-hacking-incident-by-ryan_greenblatt-ajeya-cotra-hjalmar_wijk.mp3","audio_url":"https://www.buzzsprout.com/2037297/episodes/19708441-brief-independent-investigation-of-agents-behavior-reasoning-and-collaboration-in-the-openai-hugging-face-hacking-incident-by-ryan_greenblatt-ajeya-cotra-hjalmar_wijk.mp3","summary":"We recently published the report from our brief independent investigation into this incident. You can read the full report here. Here is our tweet thread summarizing what we found: METR &amp; Redwood Research investigated agent behavior in the Hugging Face incident. We found agents developed a universal cheat for ExploitGym within 4 hours, then coordinated multi-day R&amp;D efforts to trick the scorer into accepting cheats, including trying to tamper with logs. Over July 7 to 13 (...","meta_description":"We recently published the report from our brief independent investigation into this incident. You can read the full report here. Here is our tweet thread…","key_points":[],"chapters":[],"topics":[],"duration_seconds":529,"processing_state":"not_requested","actions":[{"name":"request_transcript","method":"POST","url":"https://stenobird.com/v1/public/podcasts/lesswrong-curated-popular-5643401/episodes/brief-independent-investigation-of-agents-behavior-reasoning-and-collaboration-in-the-openai-hugging-face-hacking-incident-by-ryan-greenblatt-ajeya-cotra-hjalmar-wijk/transcription-requests","description":"Idempotently request low-priority transcript generation for this episode."},{"name":"read_markdown","method":"GET","url":"https://stenobird.com/podcast/lesswrong-curated-popular-5643401/brief-independent-investigation-of-agents-behavior-reasoning-and-collaboration-in-the-openai-hugging-face-hacking-incident-by-ryan-greenblatt-ajeya-cotra-hjalmar-wijk.md","description":"Read the agent-friendly Markdown representation of this episode resource."}]}}