# How Cloud Bills Now Charge for Shared Accelerator Memory Page: https://stenobird.com/podcast/cloud-computing-with-fexingo-aws-azure-gcp-and-modern-infrastructure-conversations-7871918/how-cloud-bills-now-charge-for-shared-accelerator-memory Text version: https://stenobird.com/podcast/cloud-computing-with-fexingo-aws-azure-gcp-and-modern-infrastructure-conversations-7871918/how-cloud-bills-now-charge-for-shared-accelerator-memory.md Podcast: [Cloud Computing with Fexingo: AWS, Azure, GCP, and Modern Infrastructure Conversations](https://stenobird.com/podcast/cloud-computing-with-fexingo-aws-azure-gcp-and-modern-infrastructure-conversations-7871918) Published: 2026-07-02T08:22:40+00:00 Episode link: https://audio.fexingo.com/business/cloud-computing/episode-0086.mp3 Audio file: https://audio.fexingo.com/business/cloud-computing/episode-0086.mp3 Processing state: not_requested JSON: https://stenobird.com/v1/public/podcasts/cloud-computing-with-fexingo-aws-azure-gcp-and-modern-infrastructure-conversations-7871918/episodes/how-cloud-bills-now-charge-for-shared-accelerator-memory Duration seconds: 492 ## Resource Episode 86 of Cloud Computing with Fexingo dives into the latest line item on enterprise cloud invoices: shared accelerator memory. Lucas explains how AWS, Azure, and GCP are now charging for GPU and TPU memory that was previously bundled into compute costs. He breaks down the pricing model using NVIDIA H100 GPUs on AWS as a concrete example, showing how a single 80 GB H100 can now incur an extra $0.40 per GB per hour for memory reserved across instances. Luna questions whether this is a hidden price hike or a genuine reflection of supply constraints. The episode explores the infrastructure logic behind disaggregated memory, the impact on AI training budgets, and why this shift may accelerate adoption of memory pooling standards like CXL. A must-listen for any team managing cloud GPU workloads. #CloudComputing #AWS #Azure #GCP #GPU #TPU #NVIDIAH100 #SharedMemory #AcceleratorMemory #CXL #AIWorkloads #CloudBilling #Infrastructure #Technology #FexingoBusiness #BusinessPodcast #CloudEconomics #MemoryPooling Keep every episode free: buymeacoffee.com/fexingo ## Actions - request_transcript: `POST https://stenobird.com/v1/public/podcasts/cloud-computing-with-fexingo-aws-azure-gcp-and-modern-infrastructure-conversations-7871918/episodes/how-cloud-bills-now-charge-for-shared-accelerator-memory/transcription-requests` — Idempotently request low-priority transcript generation for this episode. - read_markdown: `GET https://stenobird.com/podcast/cloud-computing-with-fexingo-aws-azure-gcp-and-modern-infrastructure-conversations-7871918/how-cloud-bills-now-charge-for-shared-accelerator-memory.md` — Read the agent-friendly Markdown representation of this episode resource. A page view does not enqueue transcription. Agents should invoke `request_transcript` explicitly when they need this episode processed. ## Transcript Full transcripts are not published on public pages unless there is a clear rights basis.