q8

Tag

Cards List
#q8

Building a local AI server for Qwen3 30B with Q8 is this hardware a good fit?

Reddit r/AI_Agents · 2026-07-20

A discussion about building a local AI server for the Qwen3 30B model with Q8 quantization, questioning whether the chosen hardware is a good fit.

0 favorites 0 likes
#q8

Getting close to 100K context on 32GB VRAM with Qwen3.6-27 at Q8

Reddit r/LocalLLaMA · 2026-07-05

A user shares their attempts and configurations to achieve up to 115K context on a Q8-quantized Qwen3.6-27B model using 32GB VRAM on an RTX 5090, with benchmark results and trade-offs between context length and kv-cache quantization.

0 favorites 0 likes
← Back to home

Submit Feedback