@bnjmn_marie: For LFM2.5 8B A1B, the MoQ GGUFs are the best They have the best ratio accuracy/size Again, it's interesting to see the…
Summary
The poster states that the MoQ GGUFs of the LFM2.5 8B A1B model offer the best accuracy-to-size ratio, advising against using versions with less than 95% accuracy recovery.
View Cached Full Text
Cached at: 06/10/26, 09:47 AM
For LFM2.5 8B A1B, the MoQ GGUFs are the best
They have the best ratio accuracy/size
Again, it’s interesting to see the performance difference for low-bit versions, but I wouldn’t use any version that doesn’t achieve a minimum of 95% accuracy recovery. https://t.co/GfXB2re4xQ
Similar Articles
I compared GGUF quants of Qwen3.6 27B to NVFP4, AWQ, AutoRound, and FP8
A detailed benchmark comparing 16 quantizations of Qwen3.6 27B across GGUF, NVFP4, AWQ, AutoRound, and FP8 formats, measuring KL divergence from the unquantized reference. Weight-only GGUF quants generally offer the best quality-size tradeoffs, while vLLM quants vary substantially.
@WaleedAhmad1a10: Check out the Qwen 3.5 27B MoQ GGUFs :
A Hugging Face repository (kaitchup/Qwen3.6-27B-GGUF-MoQ) provides GGUF quantized weights for the Qwen3.6-27B MoQ model, enabling local inference with tools like llama.cpp and Ollama.
Unsloth Dynamic 3.0 GGUFs
Unsloth has released Dynamic v3.0 GGUFs for Qwen3.8 models, offering >10% better accuracy at the same size through improved quantization techniques and calibration methods.
@9hills: Qwen3.8-27B Local Deployment Guide 1. Q4 has basically no quality loss, can even use Q3 2. Use Unsloth's GGUF. 3. Turn on low thinking. 4. Enable dflash2
This article provides a local deployment guide for the Qwen3.8-27B model, recommends using Q4 quantization and the Unsloth GGUF tool, and shares performance test results compared to the FP8 benchmark.
@aisearchio: GLM 5.2 GGUF is already here! 8-bit is ~half the size of the full model. Smaller versions coming soon https://huggingfa…
GLM 5.2 GGUF quantized model is released, with 8-bit version half the size of the full model; smaller versions are coming soon.