Episode

Why Cloud Bills Now Tax GPU Memory Bandwidth

Podcast
Cloud Computing with Fexingo: AWS, Azure, GCP, and Modern Infrastructure Conversations
Published
Jul 8, 2026
Duration seconds
633
Processing state
not_requested
Canonical source
https://audio.fexingo.com/business/cloud-computing/episode-0098.mp3
Audio
https://audio.fexingo.com/business/cloud-computing/episode-0098.mp3
JSON
/v1/public/podcasts/cloud-computing-with-fexingo-aws-azure-gcp-and-modern-infrastructure-conversations-7871918/episodes/why-cloud-bills-now-tax-gpu-memory-bandwidth
Markdown
/podcast/cloud-computing-with-fexingo-aws-azure-gcp-and-modern-infrastructure-conversations-7871918/why-cloud-bills-now-tax-gpu-memory-bandwidth.md

Actions

  • POST https://stenobird.com/v1/public/podcasts/cloud-computing-with-fexingo-aws-azure-gcp-and-modern-infrastructure-conversations-7871918/episodes/why-cloud-bills-now-tax-gpu-memory-bandwidth/transcription-requests
    Idempotently request low-priority transcript generation for this episode.
  • GET https://stenobird.com/podcast/cloud-computing-with-fexingo-aws-azure-gcp-and-modern-infrastructure-conversations-7871918/why-cloud-bills-now-tax-gpu-memory-bandwidth.md
    Read the agent-friendly Markdown representation of this episode resource.

Summary

Cloud bills are getting more granular — and more expensive. In this episode, Lucas and Luna dig into the latest line item appearing on enterprise invoices: GPU memory bandwidth charges. AWS, Azure, and Google Cloud have all introduced per-gigabyte-per-second pricing for high-bandwidth memory on their AI training instances. Lucas explains how NVIDIA's H100 and B200 GPUs with HBM3 and HBM4 memory are driving the change, and why a single training run on a p5.48xlarge instance can rack up thousands in bandwidth fees alone. Luna pushes back on whether this is legitimate cost recovery or just another way to extract rent from AI workloads. They walk through a concrete example: fine-tuning a 70-billion-parameter model for 30 days and how the bandwidth charge adds 15-20% to the total bill. If you're running GPU workloads in the cloud, this is the hidden tax you need to watch. #GPU #MemoryBandwidth #HBM #NVIDIA #H100 #B200 #CloudComputing #AWS #Azure #GoogleCloud #AICosts #Training #Inference #CloudBills #Technology #FexingoBusiness #BusinessPodcast #CloudInfrastructure Keep every episode free: buymeacoffee.com/fexingo