1-bit-quantization

Tag

Cards List
#1-bit-quantization

Got a 27B model running locally on a Jetson Orin NX 16GB (1-bit). still kind of amazed it works

Reddit r/LocalLLaMA · 2026-07-25

User reports successfully running a 27B parameter model quantized to 1-bit on a Jetson Orin NX 16GB edge device, expressing amazement at the feasibility.

0 favorites 0 likes
#1-bit-quantization

Our 1-bit quant of Hy3 295B runs 2.2x faster than the cloud API with no quality loss

Reddit r/LocalLLaMA · 2026-07-20

A 1-bit quantized version of the Hy3 295B model achieves 2.2x faster inference speed compared to the cloud API with no quality loss.

0 favorites 0 likes
#1-bit-quantization

@xenovacom: Bonsai 27B just changed the local LLM game forever. 1-bit quantization shrinks it from 54GB to just 3.8GB (-93%), while…

X AI KOLs Timeline · 2026-07-14 Cached

Bonsai 27B achieves 93% size reduction via 1-bit quantization while retaining 90% intelligence, enabling local browser inference with custom WebGPU kernels.

0 favorites 0 likes
← Back to home

Submit Feedback