AntLing-3.0-flash is now live on OpenRouter, and free to use through August 3, 2026
Summary
AntLing-3.0-flash has been released on OpenRouter and is free to use until August 3, 2026.
Similar Articles
Not local, but a heads-up for the cheap+fast crowd: Ling-3.0-flash is on OpenRouter, free till Aug. 3
Ling-3.0-flash, a cheap and fast AI model, is now available on OpenRouter for free until August 3.
@NousResearch: Ling-3.0-flash, the new MoE model from @AntLingAGI, is now free in Nous Portal for the next week! At 124B parameters an…
AntLingAGI releases Ling-3.0-flash, a 124B-parameter MoE model with 5.1B active parameters, now free for a week on Nous Portal. It matches or beats their 1T flagship on many benchmarks, designed for agent workloads like coding and tool use.
@Chinazhidx: Ant Group just released Ling-3.0-flash • 124B MoE • 5.1B active params/token • 256K context, expandable to 1M Just 1/8 …
Ant Group released Ling-3.0-flash, a 124B MoE model with 5.1B active parameters per token and 256K context expandable to 1M, matching or outperforming their 1T flagship model on most benchmarks.
Ling-3.0-flash weights: SGLang says day-0, vLLM says when they land, llama.cpp closed the 2.6 request as not_planned
The article details the current support status for the Ling-3.0-flash model weights across inference engines: SGLang commits to day-0 integration, vLLM awaits open weights, and llama.ccp lacks conversion for the Bailing MoE variant. It notes that the release pattern involves a free API window followed by open-sourcing, as seen with Ling-2.6-flash.
Benchmarks: AntLing-3.0-flash a hybrid-reasoning MoE model built for production-scale agents.
AntLing-3.0-flash is a hybrid-reasoning mixture-of-experts model designed for production-scale agent applications, as shown by benchmarks.