> 32B: Ranking among the top models in its class, 32B is our most powerful dense model, balancing capability, adaptability, and local deployability.
> 7B: The industry’s best-performing model under 10B combines strong software engineering and expert knowledge in a package small enough to run on a phone.
It's not the blog post, but there's some info here:
https://ifm.ai/k2/
375 A23B, 36 A4B, 32B, 7B, 3.7B, 0.9B variants.
> 32B: Ranking among the top models in its class, 32B is our most powerful dense model, balancing capability, adaptability, and local deployability.
> 7B: The industry’s best-performing model under 10B combines strong software engineering and expert knowledge in a package small enough to run on a phone.
https://ifm.ai/k2/ seems to work for me.
But it's missing the all-important charts that the blog had before it started asking for authentication.
You can find some of the charts on huggingface
https://huggingface.co/collections/IFM/k2-horizon
Thanks! Qwen-3.8 27B seems to benchmark better but I'd like to try this some time.
little qwen is my favorite for the homelab, vllm 0.28 now supports the dflash2 to go with it