Qwen-Next seems worse to me then 3.8 27b for coding, but I feel like I must be missing something?

Reddit r/LocalLLaMA News

Summary

A user shares their experience comparing Qwen-Next and 3.8 27b models for coding, finding the 3.8 27b stronger on harder tasks, and wonders if they're missing something.

Hi! I run both models on MTPLX on my m5 max, and since I have 128GB of ram I run the q8 27b. I think MTPLX only lets me run "optimized for speed" which it says is a dynamic q4 with 8 bit attention. Both of them honestly are very speedy! For coding (in pi agent in nodejs) I've just noticed that 27B feels stronger with harder tasks. But I've read so many people on here say qwen-next is better so I was wondering if maybe I'm just doing or thinking about it wrong? (and p.s. its sooo amazing that alibaba just made and released this amazing models for free! ❤️)
Original Article

Similar Articles

Qwen 3.8 27b is strong even at Q3_xxs

Reddit r/LocalLLaMA

The user finds Qwen 3.8 27b in Q3 quantization highly effective for coding tasks with fast inference speeds, outperforming previous models, despite minor issues in general conversations.

Are you running Qwen 3.8 27b or Qwen Flash Next?

Reddit r/LocalLLaMA

The user discusses preferences between Qwen 3.8 27b and Qwen Flash Next models on Apple hardware, comparing speeds, and inquires about improving performance with MLX and harnesses without reasoning.