qwen3.8-next

Tag

Cards List
#qwen3.8-next

Qwen3.8-Next streaming - 150tps prefill, 3.6 tps decode on M5 Air

Reddit r/LocalLLaMA · 2026-08-29

A user tested the Qwen3.8-Next model on an Apple M5 Air with 3-bit quantization, achieving 150 tokens per second prefill and 3.6 tps decode, outperforming a dense 27b model in some metrics.

0 favorites 0 likes
← Back to home

Submit Feedback