@QuixiAI: My experience with Qwen3.8-Flash-Next is that it can't even keep track of a multi-turn conversation. It answers questio…

X AI KOLs Timeline Models

Summary

Users report issues with Qwen3.8-Flash-Next, such as poor multi-turn conversation tracking, compared to the better-performing Qwen3.8-27B.

My experience with Qwen3.8-Flash-Next is that it can't even keep track of a multi-turn conversation. It answers questions I asked 2 turns ago. This is at FP8 https://t.co/tK8QrGRvnh
Original Article
View Cached Full Text

Cached at: 08/29/26, 06:03 AM

My experience with Qwen3.8-Flash-Next is that it can’t even keep track of a multi-turn conversation. It answers questions I asked 2 turns ago. This is at FP8 https://t.co/tK8QrGRvnh

David Hendrickson (@TeksEdge): Qwen3.8-Flash-Next is really bad in my test. ☹️ Qwen3.8-27B killed my Cosmic Dodge game (in a good way), it one-shotted it, and the second shot gave it polish. Qwen3.8-Flash-Next really struggled. In 5 tries, it didn’t get it completely correct. Bosses don’t die, and neither

Similar Articles

Are you running Qwen 3.8 27b or Qwen Flash Next?

Reddit r/LocalLLaMA

The user discusses preferences between Qwen 3.8 27b and Qwen Flash Next models on Apple hardware, comparing speeds, and inquires about improving performance with MLX and harnesses without reasoning.

Qwen3.8-Flash-Next-NVFP4 vs Qwen3.8-27B-FP Test Results

Reddit r/LocalLLaMA

This article presents detailed test results comparing the performance of Qwen3.8-Flash-Next-NVFP4 and Qwen3.8-27B-FP8 AI models across various tasks, highlighting that Flash-Next is faster with fewer failures but struggles with multi-step symbolic work.

Qwen/Qwen3.8-Flash-Next

Hugging Face Models Trending

Release of Qwen3.8-Flash-Next, an open-weight AI model introducing architectural innovations like Hybrid Attention with QSA and Gated Residual for improved efficiency and scalability, previewing the future Qwen4 architecture.

Is it just me or is Qwen3.8-Flash-Next ... really buggy?

Reddit r/LocalLLaMA

A user questions if others have noticed hallucinations and weird reasoning with the Qwen3.8-Flash-Next model on Mac, reporting issues with AppImage installation and tensor metadata from HuggingFace despite high quantization levels.

Qwen3.8-Flash-Next

Simon Willison's Blog

Qwen has released Qwen3.8-Flash-Next, an open-weights multimodal MoE model with 125B tokens but only 6B active parameters, providing a performance boost and serving as an early preview of the Qwen4 architecture.