upvote
Exactly; when I first got my RTX 5070 Ti (16gb, to game with!!!, upgrading from VEGA56), I loaded then-latest Qwen3.6 (~30B, cannot remember exactly). My only prior LLM experience was with models <8gb, primarily llama3.1.

My technical-expert twin played around with these LLMs, for about an hour, and then correctly reasoned "it's able to be WRONG, faster."

This seems apt. My next LLM machine will be closer to 96gb+ vRAM.

reply
Once I get some kind of settlement after getting beaten up by a cop my first purchase will be some RTX Pro 6000s.
reply
Dude, I'm saying this with the best of intent. Get help.
reply
Reddit might be leaking today.
reply