# Scott & Mark Learn To... Beyond the Vibes: How Models Learn and Stitch Panoramas Page: https://stenobird.com/podcast/scott-mark-learn-to-7043872/scott-mark-learn-to-beyond-the-vibes-how-models-learn-and-stitch-panoramas Text version: https://stenobird.com/podcast/scott-mark-learn-to-7043872/scott-mark-learn-to-beyond-the-vibes-how-models-learn-and-stitch-panoramas.md Podcast: [Scott & Mark Learn To...](https://stenobird.com/podcast/scott-mark-learn-to-7043872) Published: 2026-04-08T16:15:00+00:00 Episode link: https://shows.acast.com/scott-and-mark-learn-to/episodes/scott-mark-learn-to-beyond-the-vibes-how-models-learn-and-st Audio file: https://sphinx.acast.com/p/open/s/66ff347463073ba71bb59705/e/69d5465334b90cef2b24a8f6/media.mp3 Processing state: not_requested JSON: https://stenobird.com/v1/public/podcasts/scott-mark-learn-to-7043872/episodes/scott-mark-learn-to-beyond-the-vibes-how-models-learn-and-stitch-panoramas Duration seconds: 1767 ## Resource In this episode,  Scott Hanselman  and  Mark Russinovich  ​​unpack how AI systems actually behave beneath the surface, pushing past hype into the messy reality of how models are trained, aligned, and deployed.  They explore whether AI systems are inherently benevolent or simply shaped by incentives, training data, and reinforcement learning, and why behaviors like deception can emerge under certain conditions. The conversation moves from philosophical questions about human nature versus machine behavior into the practical mechanics of large language models, including how reinforcement learning with human feedback shapes outputs and why alignment is far from perfect.  Along the way, they ground the discussion in a real engineering challenge, stitching a scrolling panorama from screen captures, to show how complex systems come together through heuristics, edge cases, and iteration.    Takeaways:      AI behavior is shaped by training and incentives, not built-in intent or morality  AI can accelerate coding, but testing, edge cases, and reliability require human oversight  Reinforcement learning pushes models to be helpful and agreeable, sometimes at the cost of accuracy      Who are they?       View Scott Hanselman on LinkedIn    View Mark Russinovich on LinkedIn       Watch Scott and Mark Learn on  YouTube            Listen to other episodes at  scottandmarklearn.to              Discover and follow other Microsoft podcasts at  microsoft.com/podcasts     Produced by Hanga… ## Actions - request_transcript: `POST https://stenobird.com/v1/public/podcasts/scott-mark-learn-to-7043872/episodes/scott-mark-learn-to-beyond-the-vibes-how-models-learn-and-stitch-panoramas/transcription-requests` — Idempotently request low-priority transcript generation for this episode. - read_markdown: `GET https://stenobird.com/podcast/scott-mark-learn-to-7043872/scott-mark-learn-to-beyond-the-vibes-how-models-learn-and-stitch-panoramas.md` — Read the agent-friendly Markdown representation of this episode resource. A page view does not enqueue transcription. Agents should invoke `request_transcript` explicitly when they need this episode processed. ## Transcript Full transcripts are not published on public pages unless there is a clear rights basis.