fp16

Tag

Cards List
#fp16

The second K3's weights drop, I'm downloading the full FP16 and storing them in mattresses

Reddit r/LocalLLaMA · 5h ago

A user tweets about downloading the second K3 model's full FP16 weights and storing them, jokingly asking others to leave their doors unlocked.

0 favorites 0 likes
#fp16

Has anyone tested how quantization hits different capabilities separately? My results are surprising.

Reddit r/LocalLLaMA · 2026-07-09

The author shares surprising results from systematic tests on how different quantization levels (e.g., Q4_K_M, Q5_K_M) affect model capabilities separately, showing that math accuracy degrades more than knowledge tasks, and calls for more rigorous testing on context decay across quant levels.

0 favorites 0 likes
#fp16

@0xkeenz: Today I verified something I've been pondering for a long time. The official Qwen3.6 27B model has BF16 weights, but some quantized versions, like cyankiwi / Unsloth, convert some key weights to FP16 instead of keeping BF16. BF16 and FP16 have the same storage footprint, so why not just keep the original BF16 weights...?

X AI KOLs Timeline · 2026-07-08 Cached

The author verified that converting the Qwen3.6 27B model weights from BF16 to FP16 does not cause numerical overflow, and pointed out that FP16 has higher mantissa precision, explaining why quantized versions use FP16 instead of keeping BF16.

0 favorites 0 likes
← Back to home

Submit Feedback