Opus 5.5: Found 8/14 issues. Total cost: $15.40
Fable 5.1: Found 7/14 issues. Total cost: $66.34
Opus 5: Found 6/14 issues. Total cost: $15.19
Sonnet 5: Found 2/14 issues. Total cost: $19.15
This is a relatively small sample size, but it was both the best and the cheapest.
ETA: NB this is "Equivalent API" cost as reported by claude's CLI; I was using my subscription.
It does work out to be a similar cost per task though
https://artificialanalysis.ai/models/claude-opus-5-5#intelli...
It is most of the pareto frontier.
https://artificialanalysis.ai/models/claude-opus-5-5?models=...
Fable 5.1 literally was a money grabber. While I liked the results, tokens were burned so hard it was embarrassing, while Astra seemed to not care.
Also Claude makes it very hard to pay for additional token budgets, allowing only credit cards. I don’t use mine anymore since I don’t need it in everyday life I was dumbfounded.
So Anthropic is just copying OpenAI so to say, matching them and essentially with Opus 5.5 being Fable 5.1 in disguise, all they do is reduce costs.
Competition works.
If 5.5 is any better, I might try to do agentic-assisted development instead of just telling fable to delegate
In this case it's measuring something nearly meaningless. You could charge 100 times less per token, but if task completion takes 1,000 times as many tokens, it's not much of a bargain.
have they ever shared anything about their revenue mix between consumer plans vs per-token billing? this is a revenue cut on their API billing, but they're not saying anything about increased limits on the plans. so all the plan revenue just got more profitable.
It doesn't look like that's happening, on the contrary the prices are falling especially when taking into account capabilities.
I'm hardly a fan of China/Xi, but I do appreciate and benefit from this.
They will burn as much money as necessary to make that happen. And they have a virtually infinite amount of liquidity.
There's only 12 countries that do on the planet, the most "relevant" of them being Guatemala and Haiti.
As or the LLM topic: you can download weights of chinese models and remove any censorship and bias. Can you do so with american closed ones?
> Cache hits and refreshes on Claude Opus 5.5 are priced at 0.05x the base input price.
If they do the same for Haiku and Sonnet 5.5 then we should also see 5c/mtok and 10c/mtok cache read for those models, respectively. Still too high for Haiku IMO, Luna is 2c/mtok.
For long running tasks it is. That's what made Deepseek so cheap.