Who is your favourite quant publisher and why?
Summary
A user shares their preference for Unsloth quantized models due to fast releases and low perplexity, compares them with Apex MoE quants, and asks the community for their favorite quant publisher.
Similar Articles
MagicQuant Qwen3.8 27B GGUFs with Unsloth v3 & Imatrix
MagicQuant has been updated with Unsloth dynamic v3 and imatrix, showcasing benchmark-driven quantization results and hybrid discoveries for the Qwen3.8 27B model, with performance comparisons to Unsloth's versions.
Voodoo Quant beats Unsloth Dynamic 2.0 KLD by 95% in Qwen3.5 0.8B and 2B
Voodoo Quant, a quantization method, outperforms Unsloth Dynamic 2.0 KLD by 95% on Qwen3.5 0.8B and 2B models.
Quants impact for agentic use and local LLMs?
The author shares findings from testing quantization impacts on local LLMs for agentic use, revealing that many quants are statistically indistinguishable, MoEs are less affected than dense models, and significant degradation occurs below Q4 quantization.
New Muse-Glimmer-30B SoTA Quants - hopefully a new lineup :)
The author releases new SoTA quantizations of the Muse-Glimmer-30B model, claiming they outperform existing quants across VRAM classes with novel techniques, including a Q8 quant that is smaller and closer to BF16. They share methodology on HuggingFace and discuss future write-ups.
Gemma4 26b a4b Apex quant is quite good
User benchmarks the APEX quantized version of Gemma4 26B A4B model on AMD RX 9060 XT, achieving 38 tps at 90k context with no quality degradation, finding it better than previous quantizations.