@antirez: Uploading a new HF imatrix GGUF for 2 bits: same name, different content with fixed down layer of shared experts (there…

X AI KOLs Following Models

Summary

A corrected 2-bit GGUF model file has been uploaded to Hugging Face after fixing a bug in the imatrix computation, leading to improved logits recall and reduced error.

Uploading a new HF imatrix GGUF for 2 bits: same name, different content with fixed down layer of shared experts (there was a bug in the imatrix computation). Improved logits recall, less error, ...
Original Article

Similar Articles

Unsloth Minimax M3 GGUF

Reddit r/LocalLLaMA

Unsloth is uploading a GGUF quantized version of the MiniMax M3 model to Hugging Face.

huihui-ai/Huihui-GLM-5.2-abliterated-GGUF

Hugging Face Models Trending

A quantized GGUF version of the abliterated GLM-5.2 model is released on Hugging Face, enabling local inference with various tools like Transformers, llama.cpp, and vLLM.