upvote
What would you use if you wanted to stay (at least) open weight?
reply
qwen3.8, kimi3, kimi2.7, GLM-5.3 are all good families I use in my coding team

I'm mainly using flash varients, at least as the default, bump.up to stronger model as needed (less often these days)

reply
qwen3.8 27B or 2.4T? they are completely different models with completely different pricing
reply
I use all the qwen!

It's my favorite model family to interact with, it's prose is the best imo, it makes me laugh from time-to-time (like when it said it would "crib" some code from another project, lul)

I currently have qwen-flash working on an NES emulator harness so qwen-little can play my first RPG (ff1)

(tho I have used all the others I mentioned, happenstance I'm using qwen this iteration/task)

reply
Deepseek 4.1 Flash and Mimo 2.6 Flash.
reply
Flash is plenty good for planning and reviewing, for my needs. In fact, I use it for that because it's too slow for execution, despite the name.

edit: I subscribe to z.ai, I don't host.

reply
Will agree on this. From z.ai i have found the non flash to have way more consistent performance.
reply
> I use it for that because it's too slow for execution

What kind of hardware and what particular quant?

reply