{"podcast":{"title":"Cloud Computing with Fexingo: AWS, Azure, GCP, and Modern Infrastructure Conversations","slug":"cloud-computing-with-fexingo-aws-azure-gcp-and-modern-infrastructure-conversations-7871918","podcast_index_feed_id":7871918,"rss_url":"https://feeds.fexingo.com/business/cloud-computing.xml","website_url":"https://www.fexingo.com/","image_url":"https://audio.fexingo.com/business/cloud-computing/cover.png","author":"Fexingo","episode_count":105,"summary":"Cloud computing is the backbone of modern business, but the landscape is shifting fast. Lucas and Luna cut through the vendor noise to examine the real-world strategies behind AWS, Azure, and GCP — from multi-cloud architectures to edge computing and container orchestration. Each episode takes a single infrastructure decision, like choosing a database service or designing for disaster recovery, and traces its implications for cost, latency, and developer productivity. Lucas brings deep technical fluency and a journalist's skepticism toward marketing claims; Luna tests each argument against case studies from companies like Netflix, Capital One, and Adobe. They don't just compare prices — they explore trade-offs in lock-in, compliance, and operational complexity. Whether dissecting a Kubernetes outage or the economics of serverless, the conversation is always grounded in concrete specs and real bills. This is the podcast for engineering leaders and cloud architects who want to make informed bets, not follow trends. But what happens when the cloud giants change their pricing mid-contract? That's exactly the kind of tension Lucas and Luna live inside — and the conversation that will r…","last_synced_at":"2026-07-12T08:17:01.746696+00:00","page_url":"https://stenobird.com/podcast/cloud-computing-with-fexingo-aws-azure-gcp-and-modern-infrastructure-conversations-7871918"},"episode":{"title":"How Cloud Bills Are Adding AI Inference Surcharges","slug":"how-cloud-bills-are-adding-ai-inference-surcharges","published_at":"2026-06-26T08:20:13+00:00","page_url":"https://stenobird.com/podcast/cloud-computing-with-fexingo-aws-azure-gcp-and-modern-infrastructure-conversations-7871918/how-cloud-bills-are-adding-ai-inference-surcharges","show_page_url":"https://stenobird.com/podcast/cloud-computing-with-fexingo-aws-azure-gcp-and-modern-infrastructure-conversations-7871918","url":"https://audio.fexingo.com/business/cloud-computing/episode-0074.mp3","audio_url":"https://audio.fexingo.com/business/cloud-computing/episode-0074.mp3","summary":"In this episode, Lucas and Luna dig into a new line item showing up on enterprise cloud invoices: the AI inference surcharge. Amazon, Microsoft, and Google are now charging a premium per million tokens when customers use their managed inference APIs on AWS Bedrock, Azure OpenAI Service, and Vertex AI. The hosts break down why the premium ranges from 30% to 120% above base compute costs, how it's tied to NVIDIA's H100 GPU scarcity and the cost of high-bandwidth memory, and what it means for startups building AI features. They also discuss the fine print: some providers waive the surcharge if you commit to reserved GPU instances, while others levy it even on spot usage. Real numbers from a real bill show a 47% increase in monthly spend for a mid-stage startup. This episode is a practical guide to understanding and negotiating the newest cloud cost. #AIInference #CloudCosts #AWSBedrock #AzureOpenAI #VertexAI #NVIDIAH100 #GPUScarcity #EnterpriseBilling #CloudPricing #Technology #TechPodcast #FexingoBusiness #BusinessPodcast #LucasAndLuna #CloudComputing #InferenceSurcharge #TokenPricing #StartupCosts Keep every episode free: buymeacoffee.com/fexingo","meta_description":"In this episode, Lucas and Luna dig into a new line item showing up on enterprise cloud invoices: the AI inference surcharge. Amazon, Microsoft, and Googl…","key_points":[],"chapters":[],"topics":[],"duration_seconds":412,"processing_state":"not_requested","actions":[{"name":"request_transcript","method":"POST","url":"https://stenobird.com/v1/public/podcasts/cloud-computing-with-fexingo-aws-azure-gcp-and-modern-infrastructure-conversations-7871918/episodes/how-cloud-bills-are-adding-ai-inference-surcharges/transcription-requests","description":"Idempotently request low-priority transcript generation for this episode."},{"name":"read_markdown","method":"GET","url":"https://stenobird.com/podcast/cloud-computing-with-fexingo-aws-azure-gcp-and-modern-infrastructure-conversations-7871918/how-cloud-bills-are-adding-ai-inference-surcharges.md","description":"Read the agent-friendly Markdown representation of this episode resource."}]}}