{"podcast":{"title":"Cloud Computing with Fexingo: AWS, Azure, GCP, and Modern Infrastructure Conversations","slug":"cloud-computing-with-fexingo-aws-azure-gcp-and-modern-infrastructure-conversations-7871918","podcast_index_feed_id":7871918,"rss_url":"https://feeds.fexingo.com/business/cloud-computing.xml","website_url":"https://www.fexingo.com/","image_url":"https://audio.fexingo.com/business/cloud-computing/cover.png","author":"Fexingo","episode_count":105,"summary":"Cloud computing is the backbone of modern business, but the landscape is shifting fast. Lucas and Luna cut through the vendor noise to examine the real-world strategies behind AWS, Azure, and GCP — from multi-cloud architectures to edge computing and container orchestration. Each episode takes a single infrastructure decision, like choosing a database service or designing for disaster recovery, and traces its implications for cost, latency, and developer productivity. Lucas brings deep technical fluency and a journalist's skepticism toward marketing claims; Luna tests each argument against case studies from companies like Netflix, Capital One, and Adobe. They don't just compare prices — they explore trade-offs in lock-in, compliance, and operational complexity. Whether dissecting a Kubernetes outage or the economics of serverless, the conversation is always grounded in concrete specs and real bills. This is the podcast for engineering leaders and cloud architects who want to make informed bets, not follow trends. But what happens when the cloud giants change their pricing mid-contract? That's exactly the kind of tension Lucas and Luna live inside — and the conversation that will r…","last_synced_at":"2026-07-12T08:17:01.746696+00:00","page_url":"https://stenobird.com/podcast/cloud-computing-with-fexingo-aws-azure-gcp-and-modern-infrastructure-conversations-7871918"},"episode":{"title":"How Cloud Providers Are Tiering AI Model Access in 2026","slug":"how-cloud-providers-are-tiering-ai-model-access-in-2026","published_at":"2026-06-19T20:27:01+00:00","page_url":"https://stenobird.com/podcast/cloud-computing-with-fexingo-aws-azure-gcp-and-modern-infrastructure-conversations-7871918/how-cloud-providers-are-tiering-ai-model-access-in-2026","show_page_url":"https://stenobird.com/podcast/cloud-computing-with-fexingo-aws-azure-gcp-and-modern-infrastructure-conversations-7871918","url":"https://audio.fexingo.com/business/cloud-computing/episode-0061.mp3","audio_url":"https://audio.fexingo.com/business/cloud-computing/episode-0061.mp3","summary":"Cloud providers are introducing tiered access to large language models based on compute priority and latency requirements. This episode examines how AWS, Azure, and Google Cloud are segmenting AI inference into high-priority, standard, and best-effort tiers, and what that means for pricing, performance, and application design. We discuss a specific example: AWS's new AI Inference Priority Tier, which guarantees sub-100ms latency for a 30% premium over standard inference. We also explore how startups are responding by batching non-urgent requests and using speculative execution to stay on lower-cost tiers. Lucas and Luna debate whether this creates a two-speed AI economy or simply reflects the reality of scarce GPU capacity. A must-listen for anyone building AI applications on cloud infrastructure in 2026. #CloudComputing #AWS #Azure #GoogleCloud #AIModelAccess #InferenceTiers #MachineLearning #Infrastructure #TechNews #FexingoBusiness #BusinessPodcast #CloudInfrastructure #AIInference #GPUCapacity #Latency #StartupStrategy #PricingTiers #CloudEconomics Keep every episode free: buymeacoffee.com/fexingo","meta_description":"Cloud providers are introducing tiered access to large language models based on compute priority and latency requirements. This episode examines how AWS,…","key_points":[],"chapters":[],"topics":[],"duration_seconds":626,"processing_state":"not_requested","actions":[{"name":"request_transcript","method":"POST","url":"https://stenobird.com/v1/public/podcasts/cloud-computing-with-fexingo-aws-azure-gcp-and-modern-infrastructure-conversations-7871918/episodes/how-cloud-providers-are-tiering-ai-model-access-in-2026/transcription-requests","description":"Idempotently request low-priority transcript generation for this episode."},{"name":"read_markdown","method":"GET","url":"https://stenobird.com/podcast/cloud-computing-with-fexingo-aws-azure-gcp-and-modern-infrastructure-conversations-7871918/how-cloud-providers-are-tiering-ai-model-access-in-2026.md","description":"Read the agent-friendly Markdown representation of this episode resource."}]}}