{"podcast":{"title":"LessWrong (Curated & Popular)","slug":"lesswrong-curated-popular-5643401","podcast_index_feed_id":5643401,"rss_url":"https://rss.buzzsprout.com/2037297.rss","website_url":"https://sites.libsyn.com/421877","image_url":"https://storage.buzzsprout.com/xq8g0aka74ttwoa9xkxlmy5fn9oa?.jpg","author":"LessWrong","episode_count":914,"summary":"Audio narrations of LessWrong posts. Includes all curated posts and all posts with 125+ karma. If you'd like more, subscribe to the “Lesswrong (30+ karma)” feed.","last_synced_at":"2026-07-26T00:17:48.974965+00:00","page_url":"https://stenobird.com/podcast/lesswrong-curated-popular-5643401"},"episode":{"title":"\"Are we existentially threatened by the type of AI misalignment seen in the OpenAI Hugging Face attack?\" by Alex Mallen, Girish Gupta","slug":"are-we-existentially-threatened-by-the-type-of-ai-misalignment-seen-in-the-openai-hugging-face-attack-by-alex-mallen-girish-gupta","published_at":"2026-07-23T14:58:33+00:00","page_url":"https://stenobird.com/podcast/lesswrong-curated-popular-5643401/are-we-existentially-threatened-by-the-type-of-ai-misalignment-seen-in-the-openai-hugging-face-attack-by-alex-mallen-girish-gupta","show_page_url":"https://stenobird.com/podcast/lesswrong-curated-popular-5643401","url":"https://www.buzzsprout.com/2037297/episodes/19538425-are-we-existentially-threatened-by-the-type-of-ai-misalignment-seen-in-the-openai-hugging-face-attack-by-alex-mallen-girish-gupta.mp3","audio_url":"https://www.buzzsprout.com/2037297/episodes/19538425-are-we-existentially-threatened-by-the-type-of-ai-misalignment-seen-in-the-openai-hugging-face-attack-by-alex-mallen-girish-gupta.mp3","summary":"OpenAI models recently broke through a series of security boundaries and into Hugging Face servers in order to cheat on a cyber eval. A lot of people thought it was scary because it was a clear example of AI overreaching to do something strongly unwanted[1]. Others thought it not so scary: the models were mostly operating myopically on a singular task and not harboring an ambitious long-term agenda, and so would not take especially subtle or subversive actions. We think both camps are righ...","meta_description":"OpenAI models recently broke through a series of security boundaries and into Hugging Face servers in order to cheat on a cyber eval. A lot of people thou…","key_points":[],"chapters":[],"topics":[],"duration_seconds":598,"processing_state":"not_requested","actions":[{"name":"request_transcript","method":"POST","url":"https://stenobird.com/v1/public/podcasts/lesswrong-curated-popular-5643401/episodes/are-we-existentially-threatened-by-the-type-of-ai-misalignment-seen-in-the-openai-hugging-face-attack-by-alex-mallen-girish-gupta/transcription-requests","description":"Idempotently request low-priority transcript generation for this episode."},{"name":"read_markdown","method":"GET","url":"https://stenobird.com/podcast/lesswrong-curated-popular-5643401/are-we-existentially-threatened-by-the-type-of-ai-misalignment-seen-in-the-openai-hugging-face-attack-by-alex-mallen-girish-gupta.md","description":"Read the agent-friendly Markdown representation of this episode resource."}]}}