upvote
My token usage on Claude models has dropped by 83% over the last month - I'm pretty much only using it for quick one off questions or reading papers. it feels impossible for me to get Opus models to stop entering into cyclic loops, and my work is too security adjacent for Fable.

Codex has been an excellent workhorse - doesn't feel like I have to dance around the guardrails, doesn't lose _everything_ when it compacts, and doesn't litter the workspace with a million and one planning to plan files.

reply
I think harness/model pairs matter more than your analysis lets on.

I've had great luck with the ds flash v4, paired with prime-agent for the harness--I like the results a lot. And you get to see thinking tokens.

I haven't liked the model as much in opencode.

Sol & luna have been great everywhere. sol plans, luna builds.

reply
Not mentioning Grok 4.6 here is a crime. Fast and accurate.

And it can communicate, unlike the gobbledygook that comes out of Claude.

reply
> Not mentioning Grok 4.6 here is a crime.

Not yet. Don't give the guy ideas.

reply
Been working a lot recently with Grok 4.6 for implementation and gpt 5.6 sol for review. Worked really good so far.
reply
the future is here and one should be thankful for its slightly uneven distribution. otherwise we would hardly have anything left about which to develop strong opinions!
reply