points
In practice, you will be able to run models a bit bigger than 35B.
https://unsloth.ai/docs/models/qwen3.8-next
38GB of vram resident. more tk/s, more prefill.