upvote
In my experience, no. There’s no way to know though. The whole conversation and industry are a combo of benchmaxing, faith, and mysticism.

Since like last December I haven’t had any issues getting work done with whatever the latest Anthropic or OpenAI models at the time were. Tooling and models have only gotten better since then.

reply
Opus 5.5 is so good that I don't want it to be replaced anytime soon. Stop training models, Anthropic, and just serve this thing without regressions for a year or three, can you?
reply
They should etch it into an ASIC. The first model worthy of that honor.
reply
Sol 6 definitely feels kind of dumb and worse than 5.6

Astra seems better though.

Showing one potentially saturated benchmark doesn't necessarily fill me with a lot of confidence in the coding results.

reply
When GPT 6 Sol & Luna were released, everything went down. I have been running Sol at max thinking and it is about the same as old Luna with max thinking, give or take. Sometimes feeling even dumber. I can't trust it to do anything big alone anymore without babysitting.
reply
On r/codex the sentiment seems to be quite wide-spread.
reply