Did Alibaba abandon 35B A3B?
Summary
The article questions whether Alibaba has abandoned its 35B A3B MoE model, noting the absence of new small MoE model announcements alongside the Qwen 3.8 release.
Similar Articles
Any word on Qwen 3.7 9B? (Also looking for 9B-class alternatives to Qwen 3.5)
Discussion about Alibaba's potential open-weights release of Qwen 3.7 9B and seeking alternatives to Qwen 3.5 9B for local deployment.
Are we getting Qwen 3.8 35-A3B?
Speculation about the upcoming Qwen 3.8 release, questioning whether it will be a dense 27B model or a MoE variant like the previous 35B-A3B, with discussion of performance implications for local hardware.
@rohanpaul_ai: Alibaba dropped the weights for Qwen3.8-27B as a 27B open-weight multimodal model built for local deployment. - Apache …
Alibaba released Qwen3.8-27B, a 27B open-weight multimodal model for local deployment, which shows frontier-class performance on coding benchmarks like SWE-bench Pro and OSWorld, outperforming some larger models.
Qwen 3.8
Alibaba launches Qwen3.8, a 2.4 trillion parameter model, with an open-weight release planned soon. A preview version, Qwen3.8-Max-Preview, is now available on Alibaba's Token Plan platform, Qoder, and QoderWork.
@rohanpaul_ai: Alibaba released Qwen3.8-Max, a 2.4 trillion parameter model that activates only about 95 bn parameters per token. A th…
Alibaba released Qwen3.8-Max, a 2.4 trillion-parameter sparse MoE model with 95B active parameters per token, 1M token context, and strong agentic and benchmark results, including autonomously coding for days, circuit design, and outperforming rivals on Terminal Bench and PaperBench.