Hy3 1Bit 89-93 GB

Reddit r/LocalLLaMA Models

Summary

Announcement of Hy3 1-bit quantized model with 89-93 GB memory footprint.

No content available
Original Article

Similar Articles

An official 1-bit quant for Hy4??? 👀

Reddit r/LocalLLaMA

This article presents official 1-bit quantization builds for the Hy4 preview model, offering GGUF files with reduced sizes and instructions for running on patched llama.cpp.

@Xudong07452910: A flagship large model with 295B parameters can now run on a single 96GB inference GPU, with 50% faster decoding. Tencent Hunyuan team releases quantized versions for Hy3 (295B parameters). The 1-bit version (IQ1_M) compresses weights from 598GB to 85.5GB, a 6…

X AI KOLs Timeline

Tencent Hunyuan team releases quantized versions for the 295B-parameter Hy3 large model. The 1-bit version compresses weights to 85.5GB, enabling deployment on a single 96GB inference GPU with ~50% faster decoding. The open-source GGUF format is compatible with the llama.cpp ecosystem.