The thing with 3.8 next is that it uses a variant of ngrams. Part of the network is replaced by a lookup table you can store on a fast ssd.
In practice, you will be able to run models a bit bigger than 35B.
https://unsloth.ai/docs/models/qwen3.8-next