GPU Cloud Costs Drop to 18% of On-Demand Price Using Interruptible Instances
A developer reports paying between $0.55 and $0.70 per hour for GPU instances on interruptible capacity, compared to $3.04 per hour at on-demand rates — a saving of over 80%. Unlike CPU instances where discounts typically range from 55–75%, GPU families appear to qualify for discounts near the top of AWS's advertised 90% ceiling. However, interruptible pricing varies significantly by availability zone, and the cheapest zones often carry the highest risk of interruption and low capacity ratings. A harder constraint is the vCPU quota for interruptible GPU instances, which is capped at 64 per region and cannot easily be raised through self-service. This quota ceiling limits scaling flexibility and forces users to plan workloads around fixed instance shapes rather than adjusting capacity incrementally.
This is an AI-generated summary. ShortSingh links to the original source for the complete article.
Discussion (0)
Log in to join the discussion and vote.
Log in