{"podcast":{"title":"Best AI papers explained","slug":"best-ai-papers-explained-7258006","podcast_index_feed_id":7258006,"rss_url":"https://anchor.fm/s/1026675f8/podcast/rss","website_url":"https://podcasters.spotify.com/pod/show/ehwkang","image_url":"https://d3t3ozftmdmh3i.cloudfront.net/staging/podcast_uploaded_nologo/43252366/43252366-1744500070152-e62b760188d8.jpg","author":"Enoch H. Kang","episode_count":789,"summary":"Cut through the noise. We curate and break down the most important AI papers so you don’t have to.","last_synced_at":"2026-07-19T16:17:08.576018+00:00","page_url":"https://stenobird.com/podcast/best-ai-papers-explained-7258006"},"episode":{"title":"Quantifying Theoretical AI Alignment Guarantees: Receiver-Utility Bounds in Bayesian Persuasion","slug":"quantifying-theoretical-ai-alignment-guarantees-receiver-utility-bounds-in-bayesian-persuasion","published_at":"2026-07-01T03:27:27+00:00","page_url":"https://stenobird.com/podcast/best-ai-papers-explained-7258006/quantifying-theoretical-ai-alignment-guarantees-receiver-utility-bounds-in-bayesian-persuasion","show_page_url":"https://stenobird.com/podcast/best-ai-papers-explained-7258006","url":"https://podcasters.spotify.com/pod/show/ehwkang/episodes/Quantifying-Theoretical-AI-Alignment-Guarantees-Receiver-Utility-Bounds-in-Bayesian-Persuasion-e3lgi9d","audio_url":"https://anchor.fm/s/1026675f8/podcast/play/122226413/https%3A%2F%2Fd3ctxlq1ktw2nl.cloudfront.net%2Fstaging%2F2026-6-1%2F30a5ba59-3380-aa00-ecae-b9c9c9ad26c1.m4a","summary":"This research paper explores theoretical AI alignment through the lens of Bayesian persuasion, specifically examining how a misaligned AI agent might manipulate information. The authors utilize a bit-string model to analyze the interaction between an AI sender aiming to maximize &quot;1&quot; guesses and a human receiver seeking accuracy. A primary contribution is the establishment of a universal upper bound, proving that the receiver's utility under a strategic AI is at most 1.5 times the utility they would obtain without any signals. The study further demonstrates that this bound becomes tighter when the information follows independent product priors, as these limit the sender's ability to exploit correlations. Conversely, the authors provide a six-bit prior example to show that specific dependencies can drive the utility ratio above 1.25, proving there are limits to how much the bound can be lowered. Ultimately, this work provides mathematical guarantees on how much useful information can still reach a human even when the AI's incentives are not perfectly aligned.","meta_description":"This research paper explores theoretical AI alignment through the lens of Bayesian persuasion, specifically examining how a misaligned AI agent might mani…","key_points":[],"chapters":[],"topics":[],"duration_seconds":1338,"processing_state":"not_requested","actions":[{"name":"request_transcript","method":"POST","url":"https://stenobird.com/v1/public/podcasts/best-ai-papers-explained-7258006/episodes/quantifying-theoretical-ai-alignment-guarantees-receiver-utility-bounds-in-bayesian-persuasion/transcription-requests","description":"Idempotently request low-priority transcript generation for this episode."},{"name":"read_markdown","method":"GET","url":"https://stenobird.com/podcast/best-ai-papers-explained-7258006/quantifying-theoretical-ai-alignment-guarantees-receiver-utility-bounds-in-bayesian-persuasion.md","description":"Read the agent-friendly Markdown representation of this episode resource."}]}}