https://openrouter.ai/rankings#top-models
And their market share sits at around 16-20%.
At 75.3 trillion tokens for the week ending 10 Aug 2026, that means that up to 450 trillion tokens were plausibly demanded by the whole market for that week.
My take: At max saturation, each person on earth could have their demands satiated by an average of 16 agents running concurrently. Sometimes more, often times less, but the average would likely be at 16.
At 200 tokens/second for each agent, that would mean 15.48288 quintillion tokens per week.
We're currently at about 0.00290643601% of the calculated demand ceiling.
Even if the demand limit per person is just 1 agent at 50 tokens/second, the current demand's still 0.186011905% of the theoretical ceiling.
Unlimited.
What has been the limit to electricity demand globally?
Unlimited.
We can't get enough and never will. Costs have to become pretty severe to turn back the demand as well.
They're typically not built where you want housing, and the buildings are distinctly the wrong shape.
If you can't use the power infrastructure profitably my next thought would be warehousing.
But also... we've seen a pretty continually increasing demand for compute. Even if AI busts a bit (or becomes a bit more efficient) I bet most data centres stay data centres, just less profitable ones.
> built on both the insurmountable trillions of debt, and the assumption that only GPUs are all we need to continue scaling.
Insurmountable according to whom? And who assume that only GPUs are all we need to continue scaling? Google, Amazon, Microsoft, Meta and OpenAI, all have or plan custom non-GPU AI chips. Do they plan to use them not for scaling?