There is no official qwen 3.8 9b
From the model card:
> Clef-Flash is post-trained from Qwen/Qwen3.5-9B. See Clef for the larger variant.
16ms latency. And locally run.
Why go big when you can go small ?
To counter, most of the AI is not open. So is none of Microsoft Products. As long as they work, we keep using them.