quants for K2-Horizon are now available

Reddit r/LocalLLaMA Models

Summary

Quantized versions of the K2-Horizon AI model series are now available for download on Hugging Face, supporting various sizes from 0.9B to 36B parameters. However, the llama.cpp support is still in progress, requiring a fork for use.

You can finally downloads quants from: https://huggingface.co/IFM/K2-Horizon-MoVA-36B-A4B-GGUF https://preview.redd.it/cfhl43pps4rh1.png?width=2800&format=png&auto=webp&s=313eab309fb407a0a0e5da61e2a120c677d3d3cd https://huggingface.co/IFM/K2-Horizon-32B-GGUF https://preview.redd.it/jeh7lg6ss4rh1.png?width=2800&format=png&auto=webp&s=1982d31b019c5b85db68ea4520e380ce5748b03e https://huggingface.co/IFM/K2-Horizon-7B-GGUF https://preview.redd.it/xy699t9us4rh1.png?width=3800&format=png&auto=webp&s=a1420ba20d63aabe10eac7d11d6188da264ca814 https://huggingface.co/IFM/K2-Horizon-3.7B-GGUF https://preview.redd.it/rq5jiqzvs4rh1.png?width=3800&format=png&auto=webp&s=6fbd59b5b419a63185b4c844b87d7240f0932001 https://huggingface.co/IFM/K2-Horizon-0.9B-GGUF https://preview.redd.it/7bydjf4xs4rh1.png?width=3800&format=png&auto=webp&s=535ec81b2ca7b556715b0c1999964b697f9e6b40 however their llama.cpp PR is still in progress so you need to use the fork
Original Article

Similar Articles

The new k2 horizon models seem like an absolute beast

Reddit r/LocalLLaMA

The k2 horizon AI models, particularly the 7B variant, are praised for outperforming muse glimmer despite smaller size, with full open-sourcing that could set a new standard if benchmarks are accurate.

IFM/K2-Horizon-MoVA-36B-A4B-GGUF · Hugging Face

Reddit r/LocalLLaMA

This is the GGUF version of the K2-Horizon-MoVA-36B-A4B model, a Mixture-of-Experts AI model with 36B total parameters and 4B active per token, optimized for use with llama.cpp. It demonstrates frontier-class performance on agentic and reasoning benchmarks, competing with larger and closed models.

We quantized the new Ornith 1.5 9B and 35B-A3B

Reddit r/LocalLLaMA

The post details the quantization of Ornith 1.5 9B and 35B-A3B AI models using Atomic Dynamic methods, providing benchmarks against stock quantizations and sharing Hugging Face collections.