First evidence of a pending qwen3.7 open weights release. Qwen3.7-flash is on open router. They referred to Qwen3.6-35b-a3b as Qwen3.6 flash so this is likely a small MoE. The prices are substantially cheaper than 3.6 flash with a native 1M context window.
Summary
Evidence of an upcoming Qwen3.7 open weights release, with a flash variant (likely small MoE) listed on OpenRouter featuring a 1M context window and cheaper pricing than Qwen3.6 flash.
View Cached Full Text
Cached at: 07/28/26, 06:35 AM
Similar Articles
Has anyone tried Qwen3.7 flash on openrouter? How does it compare to our Qwen 3.6 27B?
A user asks about experiences with Qwen3.7 flash on OpenRouter and how it compares to Qwen 3.6 27B, suggesting it might be the next open-weight release from the Qwen team.
Qwen3.8-Flash-Next
Qwen has released Qwen3.8-Flash-Next, an open-weights multimodal MoE model with 125B tokens but only 6B active parameters, providing a performance boost and serving as an early preview of the Qwen4 architecture.
Qwen/Qwen3.6-35B-A3B-FP8
Alibaba releases Qwen3.6-35B-A3B-FP8, an open-weight quantized variant of Qwen3.6 with 35B parameters and 3B activated via MoE, featuring improved agentic coding capabilities and thinking preservation for iterative development.
Qwen3.8-Flash-Next on MLX-serve, 1m context is released!
The article announces the release of Qwen3.8-Flash-Next on MLX-serve, supporting 1 million token context with efficient performance on M5 Max hardware using quantized weights.
Qwen4's architecture is here early, firing 6B parameters out of 125B (3 minute read)
Alibaba's Qwen team released Qwen3.8-Flash-Next, an open-weight preview of the Qwen4 architecture that activates only 6B parameters out of 125B to reduce inference costs and address hardware limitations.