Tag
A user tested the Qwen3.8-Next model on an Apple M5 Air with 3-bit quantization, achieving 150 tokens per second prefill and 3.6 tps decode, outperforming a dense 27b model in some metrics.