Episode

How Microsoft Scales Testing and Safety for Generative AI with Sarah Bird - #691

Podcast
The TWIML AI Podcast (formerly This Week in Machine Learning & Artificial Intelligence)
Published
Jul 1, 2024
Duration seconds
3432
Processing state
failed
Canonical source
https://twimlai.com/podcast/twimlai/how-microsoft-scales-testing-and-safety-for-generative-ai/
Audio
https://pscrb.fm/rss/p/traffic.megaphone.fm/MLN3542757284.mp3?updated=1719851928
JSON
/v1/public/podcasts/twiml-ai-podcast/episodes/how-microsoft-scales-testing-and-safety-for-generative-ai-with-sarah-bird-691
Markdown
/podcast/twiml-ai-podcast/how-microsoft-scales-testing-and-safety-for-generative-ai-with-sarah-bird-691.md

Actions

  • POST https://stenobird.com/v1/public/podcasts/twiml-ai-podcast/episodes/how-microsoft-scales-testing-and-safety-for-generative-ai-with-sarah-bird-691/transcription-requests
    Idempotently request low-priority transcript generation for this episode.
  • GET https://stenobird.com/podcast/twiml-ai-podcast/how-microsoft-scales-testing-and-safety-for-generative-ai-with-sarah-bird-691.md
    Read the agent-friendly Markdown representation of this episode resource.

Summary

Today, we're joined by Sarah Bird, chief product officer of responsible AI at Microsoft. We discuss the testing and evaluation techniques Microsoft applies to ensure safe deployment and use of generative AI, large language models, and image generation. In our conversation, we explore the unique risks and challenges presented by generative AI, the balance between fairness and security concerns, the application of adaptive and layered defense strategies for rapid response to unforeseen AI behaviors, the importance of automated AI safety testing and evaluation alongside human judgment, and the implementation of red teaming and governance. Sarah also shares learnings from Microsoft's ‘Tay’ and ‘Bing Chat’ incidents along with her thoughts on the rapidly evolving GenAI landscape. The complete show notes for this episode can be found at https://twimlai.com/go/691.