Qwen 3.6 27B AutoRound GGUF, need your feedback
Summary
A user shares their GGUF quantized version of Qwen 3.6 27B using AutoRound, claiming it performs better than other quants, and invites feedback.
Similar Articles
I compared GGUF quants of Qwen3.6 27B to NVFP4, AWQ, AutoRound, and FP8
A detailed benchmark comparing 16 quantizations of Qwen3.6 27B across GGUF, NVFP4, AWQ, AutoRound, and FP8 formats, measuring KL divergence from the unquantized reference. Weight-only GGUF quants generally offer the best quality-size tradeoffs, while vLLM quants vary substantially.
Why is AutoRound being slept on so hard?
A user questions why AutoRound, a quantization tool offering superior accuracy retention at low bits and direct GGUF export, is overlooked despite outperforming standard AWQ and RTN, especially on complex models like Qwen3.6 27B.
We quantized Qwen 3.8 27B and compared the quants on an RTX 6000
The team quantized Qwen 3.8 27B into various GGUF formats and benchmarked them on an RTX 6000, finding similar performance across quants with AD-Q6_K recommended for safety.
Qwen 3.8 27B Released! Please Share Your Experience
Qwen 3.8 27B has been released, and the author invites users to share their experience, including which frontier model it resembles and which quantization they used.
DavidAU/Qwen3.6-27B-Heretic-Uncensored-FINETUNE-NEO-CODE-Di-IMatrix-MAX-GGUF
A community-finetuned, uncensored version of the Qwen 3.6 27B model featuring high-precision GGUF quantizations.