Episode
How Stripe Migrated Payment Routing to 99.999% Uptime
- Published
- Jun 16, 2026
- Duration seconds
- 529
- Processing state
not_requested
Actions
POST https://stenobird.com/v1/public/podcasts/the-cto-podcast-with-fexingo-technical-leadership-architecture-and-engineering-org-7871807/episodes/how-stripe-migrated-payment-routing-to-99-999-uptime/transcription-requests
Idempotently request low-priority transcript generation for this episode.GET https://stenobird.com/podcast/the-cto-podcast-with-fexingo-technical-leadership-architecture-and-engineering-org-7871807/how-stripe-migrated-payment-routing-to-99-999-uptime.md
Read the agent-friendly Markdown representation of this episode resource.
Summary
Episode 55 of The CTO Podcast dives into how Stripe rebuilt its payment routing engine to achieve 99.999% uptime. Lucas and Luna break down the architectural shift from a monolithic routing layer to a distributed, deterministic system that handles millions of transactions per second. They explore the team's decision to move away from traditional load balancers, the role of formal verification in routing logic, and how Stripe's engineers stress-tested the system with simulated global outages. Along the way, they discuss the trade-offs between latency and consistency, and why a gradual canary deployment was critical. This episode offers concrete lessons for engineering leaders designing fault-tolerant systems at scale. #Stripe #PaymentRouting #99.999PercentUptime #DistributedSystems #Architecture #FaultTolerance #FormalVerification #CanaryDeployment #LatencyConsistencyTradeoff #PaymentProcessing #EngineeringLeadership #SystemDesign #HighAvailability #BusinessAndTechnology #FexingoBusiness #BusinessPodcast #CTOPodcast #TechLeadership Keep every episode free: buymeacoffee.com/fexingo