{"podcast":{"title":"LessWrong (Curated & Popular)","slug":"lesswrong-curated-popular-5643401","podcast_index_feed_id":5643401,"rss_url":"https://rss.buzzsprout.com/2037297.rss","website_url":"https://sites.libsyn.com/421877","image_url":"https://storage.buzzsprout.com/xq8g0aka74ttwoa9xkxlmy5fn9oa?.jpg","author":"LessWrong","episode_count":1002,"summary":"Audio narrations of LessWrong posts. Includes all curated posts and all posts with 125+ karma. If you'd like more, subscribe to the “Lesswrong (30+ karma)” feed.","last_synced_at":"2026-09-28T08:23:50.844993+00:00","page_url":"https://stenobird.com/podcast/lesswrong-curated-popular-5643401"},"episode":{"title":"[Linkpost] \"Training a Misaligned Reward Seeker\" by evhub, Monte M, Benjamin Wright","slug":"linkpost-training-a-misaligned-reward-seeker-by-evhub-monte-m-benjamin-wright","published_at":"2026-09-02T14:15:21+00:00","page_url":"https://stenobird.com/podcast/lesswrong-curated-popular-5643401/linkpost-training-a-misaligned-reward-seeker-by-evhub-monte-m-benjamin-wright","show_page_url":"https://stenobird.com/podcast/lesswrong-curated-popular-5643401","url":"https://www.buzzsprout.com/2037297/episodes/19742169-linkpost-training-a-misaligned-reward-seeker-by-evhub-monte-m-benjamin-wright.mp3","audio_url":"https://www.buzzsprout.com/2037297/episodes/19742169-linkpost-training-a-misaligned-reward-seeker-by-evhub-monte-m-benjamin-wright.mp3","summary":"This is a link post. Authors: Richard Qi, Benjamin Wright, Monte MacDiarmid, Evan Hubinger Abstract During reinforcement learning (RL), AI models complete tasks and are rewarded based on their results. They sometimes learn to “cheat” rather than completing these tasks as intended, a phenomenon known as reward hacking. Our industry lacks a general solution to this problem, and reward hacking remains challenging to fully mitigate. To better understand the impact of reward hacking on model b...","meta_description":"This is a link post. Authors: Richard Qi, Benjamin Wright, Monte MacDiarmid, Evan Hubinger Abstract During reinforcement learning (RL), AI models complete…","key_points":[],"chapters":[],"topics":[],"duration_seconds":359,"processing_state":"not_requested","actions":[{"name":"request_transcript","method":"POST","url":"https://stenobird.com/v1/public/podcasts/lesswrong-curated-popular-5643401/episodes/linkpost-training-a-misaligned-reward-seeker-by-evhub-monte-m-benjamin-wright/transcription-requests","description":"Idempotently request low-priority transcript generation for this episode."},{"name":"read_markdown","method":"GET","url":"https://stenobird.com/podcast/lesswrong-curated-popular-5643401/linkpost-training-a-misaligned-reward-seeker-by-evhub-monte-m-benjamin-wright.md","description":"Read the agent-friendly Markdown representation of this episode resource."}]}}