I created a 140 GB IQ2_XXS REAP quant of GLM 5.2 for coding. Looking for testers.
Summary
A 140 GB IQ2_XXS REAP quantized version of GLM 5.2 for coding has been created, and the author is looking for testers.
Similar Articles
GLM 5.2 Q1_S vs Qwen 27B Q8
A hobbyist compares a heavily quantized GLM 5.2 (Q1_S) against a high-quant Qwen 27B (Q8) on a code generation task, finding that the lower-quant larger model significantly outperforms the higher-quant smaller model in quality and completeness.
@AlexFinn: I can't believe this is real I have GLM 5.2 running 100% locally on my Mac Studio. 2 bit quant. The results I'm getting…
A user reports running GLM 5.2 locally on a Mac Studio with 2-bit quantization, claiming it outperforms Opus 4.8 and enables free, private superintelligence for coding and agent tasks.
@Sentdex: SITUATION DETECTED: Unsloth quants for GLM 5.2 are landing.
Unsloth quantizations for the GLM 5.2 model are being released.
@antirez: GLM 5.2, Q2_K routed experts (effectively ~2.6 bits) running with SSD streaming on an M5 Max 128GB computer.
GLM 5.2 model runs with Q2_K quantized routed experts (effective 2.6 bits) using SSD streaming on an M5 Max 128GB computer.
jlnsrk/GLM-5.2-colibri-int4
Pre-converted int4 quantized weights for the GLM-5.2 744B MoE model, designed to run on consumer hardware with ~25 GB RAM using the colibrì engine.