{"podcast":{"title":"Generative AI 101","slug":"generative-ai-101-6932184","podcast_index_feed_id":6932184,"rss_url":"https://feed.podbean.com/generativeai101/feed.xml","website_url":"https://generativeai101.podbean.com","image_url":"https://pbcdn1.podbean.com/imglogo/image-logo/18828923/2026-Podcast-Cover-GenAI-101-002.jpg","author":"Emily Laird","episode_count":338,"summary":"Welcome to Generative AI 101, your go-to podcast for learning the basics of generative artificial intelligence in easy-to-understand, bite-sized episodes. Join host Emily Laird, AI Integration Technologist and AI lecturer, to explore key concepts, applications, and ethical considerations, making AI accessible for everyone.","last_synced_at":"2026-09-17T00:18:23.945962+00:00","page_url":"https://stenobird.com/podcast/generative-ai-101-6932184"},"episode":{"title":"Why AI Evaluation Still Needs Human Experts","slug":"why-ai-evaluation-still-needs-human-experts","published_at":"2026-09-10T12:58:30+00:00","page_url":"https://stenobird.com/podcast/generative-ai-101-6932184/why-ai-evaluation-still-needs-human-experts","show_page_url":"https://stenobird.com/podcast/generative-ai-101-6932184","url":"https://generativeai101.podbean.com/e/why-ai-evaluation-still-needs-human-experts/","audio_url":"https://mcdn.podbean.com/mf/web/tpbyh7dg96te6ev9/Evaluation.mp3","summary":"The most dangerous AI output isn't the ridiculous one; it's the polished answer with one critical error hiding in plain sight. In this episode, host Emily Laird puts generative AI evaluation on trial, from OpenAI's GDPval and Anthropic's TASTE study to the uncomfortable fact that automated AI judges still can't match experienced human reviewers. She breaks down metamorphic testing (a terrible name for a very useful idea) and explains how every caught mistake can become a test your systems have to survive. If you can no longer evaluate your own work, you haven't bought a productivity tool; you've built a dependency. 🎯 JOIN THE AI WEEKLY MEETUPS https://www.uwstout.edu/ai-weekly-meetup 📩 EMAIL REMINDERS FOR THE MEETUPS https://app.e2ma.net/app2/audience/signup/2101263/1779703/ 💬 CONNECT WITH EMILY LAIRD ON LINKEDIN http://www.linkedin.com/in/meet-emily-laird","meta_description":"The most dangerous AI output isn't the ridiculous one; it's the polished answer with one critical error hiding in plain sight. In this episode, host Emily…","key_points":[],"chapters":[],"topics":[],"duration_seconds":820,"processing_state":"not_requested","actions":[{"name":"request_transcript","method":"POST","url":"https://stenobird.com/v1/public/podcasts/generative-ai-101-6932184/episodes/why-ai-evaluation-still-needs-human-experts/transcription-requests","description":"Idempotently request low-priority transcript generation for this episode."},{"name":"read_markdown","method":"GET","url":"https://stenobird.com/podcast/generative-ai-101-6932184/why-ai-evaluation-still-needs-human-experts.md","description":"Read the agent-friendly Markdown representation of this episode resource."}]}}