@bnjmn_marie: For LFM2.5 8B A1B, the MoQ GGUFs are the best They have the best ratio accuracy/size Again, it's interesting to see the…

X AI KOLs Following Tools

Summary

The poster states that the MoQ GGUFs of the LFM2.5 8B A1B model offer the best accuracy-to-size ratio, advising against using versions with less than 95% accuracy recovery.

For LFM2.5 8B A1B, the MoQ GGUFs are the best They have the best ratio accuracy/size Again, it's interesting to see the performance difference for low-bit versions, but I wouldn't use any version that doesn't achieve a minimum of 95% accuracy recovery. https://t.co/GfXB2re4xQ
Original Article
View Cached Full Text

Cached at: 06/10/26, 09:47 AM

For LFM2.5 8B A1B, the MoQ GGUFs are the best

They have the best ratio accuracy/size

Again, it’s interesting to see the performance difference for low-bit versions, but I wouldn’t use any version that doesn’t achieve a minimum of 95% accuracy recovery. https://t.co/GfXB2re4xQ

Similar Articles

I compared GGUF quants of Qwen3.6 27B to NVFP4, AWQ, AutoRound, and FP8

Reddit r/LocalLLaMA

A detailed benchmark comparing 16 quantizations of Qwen3.6 27B across GGUF, NVFP4, AWQ, AutoRound, and FP8 formats, measuring KL divergence from the unquantized reference. Weight-only GGUF quants generally offer the best quality-size tradeoffs, while vLLM quants vary substantially.

Unsloth Dynamic 3.0 GGUFs

Hacker News Top

Unsloth has released Dynamic v3.0 GGUFs for Qwen3.8 models, offering >10% better accuracy at the same size through improved quantization techniques and calibration methods.