upvote
That number sounds about right, if a little low. According to OpenRouter, the weighted average input cost (which includes cache discounts) of GLM 5.3 is $0.2337 and the output cost is $3.291 per million tokens. If we assume 80% of the tokens are inputs, the cost of 450M tokens should be right around $300, which is the correct order of magnitude. And it depends highly on the provider(s) that the author is using, the ratio of inputs to outputs, etc.

It really makes you see how heavily subsidized the subscriptions are.

Edit: Fixed my math. Edit 2: I was looking at the wrong model on OR. Either way, the math is within the correct ballpark.

reply
EDIT: Proof (can't edit the original comment)

https://imgur.com/a/mSXGzzJ

reply
The high usage was due to omp in vibe mode overnight, probably working way too hard through things. The high cost, yes we’d rather pay extra to work with providers that provide other benefits than just lowest cost possible (open models, no training, EU DC, etc)
reply