ChatGPT and Gemini Hit 1 Billion Users, but the Infrastructure Cost Is Staggering
ChatGPT surpassed one billion weekly users faster than any consumer product in history, with Google's Gemini reaching similar numbers shortly after. Unlike traditional web platforms where serving more users gets cheaper over time, every AI query requires real GPU compute, meaning the marginal cost per user does not shrink the way it did for companies like Facebook. Providers like OpenAI are simultaneously cutting prices — with some tiers dropping to around $0.20 per million input tokens — while absolute compute demand explodes, forcing them to absorb losses now in hopes that unit economics improve later. For businesses building on these APIs, this creates volatile, hard-to-predict costs that behave less like a standard API call and more like an expensive compute job priced per token. Experts warn that teams must actively route queries to cheaper models for simpler tasks and closely attribute AI spending by feature and team, applying the same financial discipline used for cloud infrastructure.
This is an AI-generated summary. ShortSingh links to the original source for the complete article.
Discussion (0)
Log in to join the discussion and vote.
Log in