The death of SLMs?

Reddit r/LocalLLaMA News

Summary

The author reflects on whether small language models under 27B are being overshadowed by larger models like Qwen 3.5 and Gemma 4, and asks the community for capable SLMs for agentic coding tasks.

I love to see these impressive models coming out that compete with the giants from companies like Z.ai, Moonshot, Alibaba, etc. A win for the open source/weight community is always welcome. While I am grateful, I worry we might be seeing the slow death of models smaller than 27B. The ones released paling in comparison to Qwen 3.5 4B/9B and Gemma 4 12B. Especially for agentic coding and agentic assistance tasks. Is this because we’ve really hit the limit of what we can accomplish with models in the 3B-12B weight class? Or is it because such models aren’t as profitable as their gargantuan counterparts that attempting to improve them to match isn’t viable? Have I been missing these impressive smaller model in lieu of the larger ones taking the headlines? If so, please let me know what models within the SLM weight class you are running for tasks like agentic coding, agentic assistance, or both. I also hear agentic coding is not feasible under 27B. I’m not asking for a model that can one shot an ultra realistic multiplayer call of duty clone in a single html file. Just something the least bit capable in real workflows like the aforementioned.
Original Article

Similar Articles

Has anyone here used SLMs inside agent workflows?

Reddit r/AI_Agents

A user asks the community about using small/local language models within agent workflows for specific tasks like routing, classification, and extraction, and shares thoughts on whether larger models are always necessary.

No more SLM open-source??

Reddit r/LocalLLaMA

The post questions whether small language models (SLMs) will no longer be open-sourced, sparking discussion about the future of open-source AI.

Does size really matter? (LLMs vs. SLMs)

Reddit r/artificial

Discusses the trade-offs between large language models (LLMs) and small language models (SLMs), questioning whether larger models are always necessary for production use cases and exploring the future of AI deployment.

Are super tiny LLMs any good?

Reddit r/singularity

Explores whether very small language models can handle casual conversations adequately, and what training factors differentiate the better ones.