@NousResearch: Ling-3.0-flash, the new MoE model from @AntLingAGI, is now free in Nous Portal for the next week! At 124B parameters an…
Summary
AntLingAGI releases Ling-3.0-flash, a 124B-parameter MoE model with 5.1B active parameters, now free for a week on Nous Portal. It matches or beats their 1T flagship on many benchmarks, designed for agent workloads like coding and tool use.
View Cached Full Text
Cached at: 07/24/26, 07:14 PM
Ling-3.0-flash, the new MoE model from @AntLingAGI, is now free in Nous Portal for the next week!
At 124B parameters and 5.1B active, it’s quick to run and built for agent workloads: coding, search, research, and tool use.
Try it free today at https://t.co/aHGXAcLs93 https://t.co/cLcjsnrYjK
Ant Ling (@AntLingAGI): Today, we’re releasing Ling-3.0-flash—a hybrid-reasoning MoE model built for production-scale agents.
124B parameters. Just 5.1B active per token.
With 1/8 of the total and 1/12 of the active parameters, it matches or beats our 1T flagship model on most benchmarks shown.
Similar Articles
@NousResearch: Step 3.7 Flash is now free for 30 days via Nous Portal It is a new MoE vision-language model focused on agent efficienc…
Step 3.7 Flash, a new MoE vision-language model focused on agent efficiency, coding, search, and multimodal workflows, is now free for 30 days via Nous Portal.
@Chinazhidx: Ant Group just released Ling-3.0-flash • 124B MoE • 5.1B active params/token • 256K context, expandable to 1M Just 1/8 …
Ant Group released Ling-3.0-flash, a 124B MoE model with 5.1B active parameters per token and 256K context expandable to 1M, matching or outperforming their 1T flagship model on most benchmarks.
ling 3.0 flash/tiny base models
InclusionAI has open-sourced the Ling-3.0 series, featuring highly efficient language models with sparse MoE architecture and hybrid linear attention, providing checkpoints at various training stages to support research and innovation.
New model release: Ling-3.0-tiny: 7.9B total parameters, with only 1.3B active per token- free for a week
Release of Ling-3.0-tiny, a hybrid reasoning model with 7.9B total parameters and only 1.3B active per token, free for a week.
@AntLingAGI: Introducing Ling-2.6-flash, an instruct model with 104B total parameters and 7.4B active parameters. Ling-2.6-flash is …
Ling-2.6-flash is a 104B-total/7.4B-active sparse instruct model optimized for token efficiency, aiming to cut costs and boost throughput on agent tasks.