Episode

How Slack Cut Mean Time to Acknowledge by 60 Percent With On-Call Orchestration

Podcast
The Site Reliability Podcast with Fexingo: SRE, Uptime, and Production Engineering
Published
Jul 17, 2026
Duration seconds
614
Processing state
not_requested
Canonical source
https://audio.fexingo.com/business/the-site-reliability-podcast/episode-0117.mp3
Audio
https://audio.fexingo.com/business/the-site-reliability-podcast/episode-0117.mp3
JSON
/v1/public/podcasts/the-site-reliability-podcast-with-fexingo-sre-uptime-and-production-engineering-7871923/episodes/how-slack-cut-mean-time-to-acknowledge-by-60-percent-with-on-call-orchestration
Markdown
/podcast/the-site-reliability-podcast-with-fexingo-sre-uptime-and-production-engineering-7871923/how-slack-cut-mean-time-to-acknowledge-by-60-percent-with-on-call-orchestration.md

Actions

  • POST https://stenobird.com/v1/public/podcasts/the-site-reliability-podcast-with-fexingo-sre-uptime-and-production-engineering-7871923/episodes/how-slack-cut-mean-time-to-acknowledge-by-60-percent-with-on-call-orchestration/transcription-requests
    Idempotently request low-priority transcript generation for this episode.
  • GET https://stenobird.com/podcast/the-site-reliability-podcast-with-fexingo-sre-uptime-and-production-engineering-7871923/how-slack-cut-mean-time-to-acknowledge-by-60-percent-with-on-call-orchestration.md
    Read the agent-friendly Markdown representation of this episode resource.

Summary

In this episode of The Site Reliability Podcast with Fexingo, Lucas and Luna dive into how Slack's SRE team overhauled their on-call escalation system to reduce mean time to acknowledge (MTTA) by 60 percent. They walk through the specific mechanics: tiered alerting based on service criticality, automated handoffs when a primary responder misses a page, and a real-time dashboard that shows who is actually available. The conversation touches on the trade-off between speed and context, why Slack chose a 'noisy' paging model over a quiet one, and how the team used post-incident data to tune escalation policies. A concrete case study in turning alert fatigue into reliable response. #Slack #SiteReliabilityEngineering #OnCall #IncidentResponse #MTTA #Alerting #EscalationPolicy #PagerDuty #Reliability #SRE #Uptime #ProductionEngineering #IncidentManagement #Observability #Automation #Technology #FexingoBusiness #BusinessPodcast Keep every episode free: buymeacoffee.com/fexingo