upvote
> I've settled on GLM-5.3 (formerly Deepseek v4 pro 0813) for architecting

Dude, GLM-5.3 released _today_.

The phrasing "I've settled on" is incorrect for this context.

reply
hence the "former deepseek v4 pro". I tried it out this morning and have had no complaints. I already liked glm 5.2
reply
> Deepseek v4 pro 0813

Which itself released yesterday? You're writing, reading, and evaluating enough software in a ~36 hour period to form, reject, and form another opinion about which model makes better _architectural_ choices?

reply
The sentence still doesn't make sense, because "settled on" implies a long testing phase with a verdict eventually emerging out of that.

What you're currently doing is "testing out"

reply
Honest question, how do you assess models this quickly? What metrics are you using? Would love to get my suite from multiple days and hundreds of prompts down to minutes. Got a few first pass tasks I run upon release for an initial experience, but those only work because even Fable and Sol fail despite objectively correct solutions existing, so it works because most models fail, but then, those are consciously not enough for coding, tool use, adherence or task specific inference and assessment…
reply
What are you working on? That can dictate which models are best.
reply
You should check out Grok, it's quite a good deal from the Cursor subscription side but it's cheap even by API prices.
reply
No serious person or sane person uses the LLM that's constantly being tweaked by an anti-woke white-genocide-supporting weird little man. Don't feed the totalitarian wannabe's (or the totalitarians in general, for that matter).
reply
Yeah I'm sure everyone on r/cursor or in previous HN threads about Grok 4.5 or 4.6 are all unserious and insane.

No one actually cares about the politics as long as the model codes well.

Edit, quite interesting to see the reception to this comment compared to essentially the same type of comment I made on a Grok 4.6 benchmark HN post: https://news.ycombinator.com/item?id=49275385#49275571

It's true that Cursor gives a lot of usage with Grok, most users of Cursor don't care about Musk.

reply
Voting with your wallet is still very much a valid way to protest that odious man.

Some people might not mind (or even know), but I sleep better at night trying to work as ethically as I can.

reply
There are a lot of people who are apathetic to what musk is, people that don't care are not people who should inspire you. What the hell is so inspiring about apathy anyway?!

And yeah, people that don't care DO make the world worse through their apathy.

reply
I would say it is sad that there are people who use Grok when there are so many other choices available which don't come with the issues of supporting Musk.

It is not all just 'politics'. Take a stand on some issues. It doesn't cost much not to use Grok.

reply
This is the “Mussolini made the trains run on time” of ai hot takes.

(Btw, mussolini didn’t make the trains run on time)

reply
"Sure I'm indirectly funding the erosion of basic human rights in the States, but at least I made my Hello World app cheaper!"
reply
Also: "What do you mean everybody at this party is a Nazi? They seem uninterested in politics, and they have been so welcoming to me!"
reply
i care about not financially supporting a person that is actively trying to disenfranchise me, why is that a difficult concept for some people? that not everyone is motivated exclusively by financial profit? is moral bankruptcy so pervasive that some people assume it is unanimous?
reply
They are all unserious and insane, yes.
reply
i think you will like luna if you haven't tried it yet
reply
Luna is twice the price of Deepseek V4 Flash 0731, and less capable :/
reply
[delayed]
reply
I'm convinced a lot of the anti-open-weight model comments at this point are inorganic traffic - there's trillions in investor money riding on a world where these models aren't cheap commodities. Having actually used things like the recent GLM, Kimi, and Qwen I think any edge the labs have is marginal at most and actually prefer the open weight models in most day to day usage.

Anthropic's recent releases are wordy to the point of exhaustion. Every time I use opus recently I find myself wanting to yell "GET TO THE POINT" at a terminal, which is exacerbated by it being slow.

reply