Kimi K2.6 Unsloth GGUF is out

Reddit r/LocalLLaMA Models

Summary

Unsloth has released a GGUF-quantized version of the Kimi K2.6 model, enabling efficient local inference.

[https://huggingface.co/unsloth/Kimi-K2.6-GGUF](https://huggingface.co/unsloth/Kimi-K2.6-GGUF) [https://unsloth.ai/docs/basics/unsloth-dynamic-2.0-ggufs](https://unsloth.ai/docs/basics/unsloth-dynamic-2.0-ggufs)
Original Article

Similar Articles

unsloth/Kimi-K2.6-GGUF

Hugging Face Models Trending

Unsloth releases quantized GGUF versions of the open-source 1T-parameter Kimi K2.6 MoE model, optimized for long-horizon coding, autonomous agent swarms, and production-ready design tasks.

unsloth/Kimi-K2.7-Code-GGUF

Hugging Face Models Trending

Unsloth releases GGUF quantizations of Kimi K2.7 Code, a 1 trillion parameter MoE coding model built on Kimi K2.6 with improved token efficiency and agentic coding capabilities.

Unsloth Minimax M3 GGUF

Reddit r/LocalLLaMA

Unsloth is uploading a GGUF quantized version of the MiniMax M3 model to Hugging Face.

Anyone tried the Q1 Kimi K3 yet? (555GB)

Reddit r/LocalLLaMA

Kimi K3 is a massive 2.9 trillion parameter mixture-of-experts model with 104B active parameters, 1M context length, and native MXFP4 training, now available in GGUF quantizations ranging from 540GB to smaller sizes, though requiring substantial hardware to run.