upvote
I am running GLM 5.3 across 2x DGX Sparks and was doing comparisons and it absolutely can beat Gemini. Yesterday it corrected a poor Fable 5 response even
reply
I should have clarified, you can’t run it on their local hardware, which is a 16gb gpu. Glm-5.3 and k3 are of course near the frontier.
reply
GLM 5.3 Flash? Qwen 3.8 Flash Next? I believe those both are as good as the best Gemini, competitive with Terra.
reply
Yes they are quite good, but are not able to run on a 16GB RX 9070.

Quantized Qwen 3.8 Flash Next could maybe run eventually on that card with a highly optimized inference engine that dynamically caches the hottest layer experts. Even then you run into some hard limits.

reply
You are insanely non-technical then. Yes you can. Skill issue

You probably pay $200/mo for text-to-text!

reply