Ling 3.0 (on Nous)
Summary
This article discusses user-reported issues with the Ling 3.0 AI model from Nous, focusing on problems with text output and context length limitations, potentially linked to quantization or rate limits.
Similar Articles
@NousResearch: Ling-3.0-flash, the new MoE model from @AntLingAGI, is now free in Nous Portal for the next week! At 124B parameters an…
AntLingAGI releases Ling-3.0-flash, a 124B-parameter MoE model with 5.1B active parameters, now free for a week on Nous Portal. It matches or beats their 1T flagship on many benchmarks, designed for agent workloads like coding and tool use.
Ling-3.0-tiny is a very interesting model. Run on NVIDIA Orin Nano Super 8GB at 128K context with IQ4_NL quant.
The article demonstrates running the Ling-3.0-tiny AI model on an NVIDIA Orin Nano Super 8GB device with IQ4_NL quantization, achieving 33 tok/s decode speed and full 128K context, showcasing practical edge AI deployment.
ling 3.0 flash/tiny base models
InclusionAI has open-sourced the Ling-3.0 series, featuring highly efficient language models with sparse MoE architecture and hybrid linear attention, providing checkpoints at various training stages to support research and innovation.
For Ling-2.6-1T, what would make the size feel justified first: quality per token, local serving reality, or long context stability?
The article questions whether the Ling-2.6-1T model's size is justified by quality, local serving feasibility, or long context stability, describing it as an open-source MoE model with 1T total params and up to 1M native context.
Is Ling 3 tiny underrated for its size?
A user discusses the Ling 3 tiny model's benchmarks, comparing it to Qwen3.5 9b and questioning if other open-source models are being overlooked.