@Chinazhidx: Ant Group just released Ling-3.0-flash • 124B MoE • 5.1B active params/token • 256K context, expandable to 1M Just 1/8 …
Summary
Ant Group released Ling-3.0-flash, a 124B MoE model with 5.1B active parameters per token and 256K context expandable to 1M, matching or outperforming their 1T flagship model on most benchmarks.
View Cached Full Text
Cached at: 07/24/26, 07:17 PM
🔥Ant Group just released Ling-3.0-flash
• 124B MoE • 5.1B active params/token • 256K context, expandable to 1M
Just 1/8 the total and 1/12 the active parameters of Ant’s 1T flagship model—yet matches or outperforms it on most benchmarks. #AntLing https://t.co/Rycwi6GoVC
Similar Articles
@NousResearch: Ling-3.0-flash, the new MoE model from @AntLingAGI, is now free in Nous Portal for the next week! At 124B parameters an…
AntLingAGI releases Ling-3.0-flash, a 124B-parameter MoE model with 5.1B active parameters, now free for a week on Nous Portal. It matches or beats their 1T flagship on many benchmarks, designed for agent workloads like coding and tool use.
@AntLingAGI: Introducing Ling-2.6-flash, an instruct model with 104B total parameters and 7.4B active parameters. Ling-2.6-flash is …
Ling-2.6-flash is a 104B-total/7.4B-active sparse instruct model optimized for token efficiency, aiming to cut costs and boost throughput on agent tasks.
@0x0SojalSec: Final take : Tencent recently drop a 295B parameter model that only activates 21B params per token. While most labs are…
Tencent released Hy3, a 295B parameter MoE model with 21B active parameters per token, competitive with larger models on agentic coding and tool use tasks, with Apache 2.0 weights.
@thesupermanmx: China just changed the game They just dropped an open-sourced model that burns only 1% of the tokens compared to your f…
China released Ling-2.6-1T, a 1 trillion parameter open-source model that achieves 72.2% on SWE-bench with a 256K context window, claiming high efficiency and compatibility with Claude Code.
Benchmarks: AntLing-3.0-flash a hybrid-reasoning MoE model built for production-scale agents.
AntLing-3.0-flash is a hybrid-reasoning mixture-of-experts model designed for production-scale agent applications, as shown by benchmarks.