{"podcast":{"title":"The Site Reliability Podcast with Fexingo: SRE, Uptime, and Production Engineering","slug":"the-site-reliability-podcast-with-fexingo-sre-uptime-and-production-engineering-7871923","podcast_index_feed_id":7871923,"rss_url":"https://feeds.fexingo.com/business/the-site-reliability-podcast.xml","website_url":"https://www.fexingo.com/","image_url":"https://audio.fexingo.com/business/the-site-reliability-podcast/cover.png","author":"Fexingo","episode_count":119,"summary":"Lucas and Luna cut through the noise around site reliability engineering to examine how real-world SRE teams balance uptime, incident response, and production change. Each episode takes a single concept — error budgets, toil automation, postmortem culture, capacity planning — and grounds it in a specific case: how a major streaming service reduced paging noise, how a payments platform rebuilt its incident command structure, or how a cloud provider manages multi-region failover. Lucas brings the numbers — latency percentiles, MTTR trends, SLO burn rates — while Luna pushes on the human and organizational trade-offs: What does a junior SRE need to know about on-call? How do you measure reliability without crushing innovation? Why do some blameless postmortems actually work? Together they treat SRE not as a certification topic but as a living practice, citing real outages, open-source tools, and engineering blogs. This show is for engineers, ops leads, and platform teams who already know the basics and want to debate the hard edges: Is 99.999% uptime always worth the cost? When should you deliberately degrade service to improve reliability? How do you design for resilience when your…","last_synced_at":"2026-07-18T22:18:41.635042+00:00","page_url":"https://stenobird.com/podcast/the-site-reliability-podcast-with-fexingo-sre-uptime-and-production-engineering-7871923"},"episode":{"title":"How SRE Teams Use Canary Deployments to Reduce Release Risk","slug":"how-sre-teams-use-canary-deployments-to-reduce-release-risk","published_at":"2026-07-09T10:08:43+00:00","page_url":"https://stenobird.com/podcast/the-site-reliability-podcast-with-fexingo-sre-uptime-and-production-engineering-7871923/how-sre-teams-use-canary-deployments-to-reduce-release-risk","show_page_url":"https://stenobird.com/podcast/the-site-reliability-podcast-with-fexingo-sre-uptime-and-production-engineering-7871923","url":"https://audio.fexingo.com/business/the-site-reliability-podcast/episode-0100.mp3","audio_url":"https://audio.fexingo.com/business/the-site-reliability-podcast/episode-0100.mp3","summary":"In this milestone 100th episode, Lucas and Luna zero in on canary deployments – a key SRE strategy for rolling out software changes safely to a small subset of users before full release. They walk through a concrete example from a major online retailer that cut its mean time to recover from four hours to under seven minutes by integrating canary analysis into its continuous delivery pipeline. The conversation covers progressive traffic shifting, automated rollback triggers, monitoring integration, and the common pitfalls teams face when first adopting canaries. Lucas explains the difference between blue-green deployments and canaries, and Luna pushes on how to choose the right canary size and duration. The hosts also discuss how canaries fit into an error budget policy and why some organizations struggle to get developers to actually watch the canary dashboards. A practical, example-driven episode for any engineer or SRE looking to ship changes faster without catching fire. #CanaryDeployments #SiteReliabilityEngineering #ReleaseRisk #ProgressiveDelivery #IncidentResponse #ContinuousDelivery #BlueGreenDeployment #AutomatedRollback #ErrorBudgets #MeanTimeToRecover #Observability #SRE #Uptime #Technology #FexingoBusiness #BusinessPodcast #Ep100 #ReleaseEngineering Keep every episode free: buymeacoffee.com/fexingo","meta_description":"In this milestone 100th episode, Lucas and Luna zero in on canary deployments – a key SRE strategy for rolling out software changes safely to a small subs…","key_points":[],"chapters":[],"topics":[],"duration_seconds":456,"processing_state":"not_requested","actions":[{"name":"request_transcript","method":"POST","url":"https://stenobird.com/v1/public/podcasts/the-site-reliability-podcast-with-fexingo-sre-uptime-and-production-engineering-7871923/episodes/how-sre-teams-use-canary-deployments-to-reduce-release-risk/transcription-requests","description":"Idempotently request low-priority transcript generation for this episode."},{"name":"read_markdown","method":"GET","url":"https://stenobird.com/podcast/the-site-reliability-podcast-with-fexingo-sre-uptime-and-production-engineering-7871923/how-sre-teams-use-canary-deployments-to-reduce-release-risk.md","description":"Read the agent-friendly Markdown representation of this episode resource."}]}}