upvote
It's not the blog post, but there's some info here:

https://ifm.ai/k2/

375 A23B, 36 A4B, 32B, 7B, 3.7B, 0.9B variants.

> 32B: Ranking among the top models in its class, 32B is our most powerful dense model, balancing capability, adaptability, and local deployability.

> 7B: The industry’s best-performing model under 10B combines strong software engineering and expert knowledge in a package small enough to run on a phone.

reply
https://ifm.ai/k2/ seems to work for me.
reply
But it's missing the all-important charts that the blog had before it started asking for authentication.
reply
You can find some of the charts on huggingface

https://huggingface.co/collections/IFM/k2-horizon

reply
Thanks! Qwen-3.8 27B seems to benchmark better but I'd like to try this some time.
reply
little qwen is my favorite for the homelab, vllm 0.28 now supports the dflash2 to go with it
reply