Are super tiny LLMs any good?

Reddit r/singularity Models

Summary

Explores whether very small language models can handle casual conversations adequately, and what training factors differentiate the better ones.

If you’re not coding, not asking complex logical questions, but still want a model that isn’t completely stupid for casual conversations, are there any super tiny models out there that do an ok job? Which ones, and what makes them good, how were they trained and weighted that made them better than other tiny models?
Original Article

Similar Articles

Does size really matter? (LLMs vs. SLMs)

Reddit r/artificial

Discusses the trade-offs between large language models (LLMs) and small language models (SLMs), questioning whether larger models are always necessary for production use cases and exploring the future of AI deployment.

Small LLMs: Pruning vs. Training from Scratch

arXiv cs.LG

This paper empirically compares pruning vs. training small language models from scratch, finding that pruning provides a strong advantage under limited token budgets but that the advantage diminishes as training scales, especially with coarse pruning.

What would optimal use of LLMs even look like?

Reddit r/singularity

Explores the speculative idea of optimizing human interaction with LLMs by conforming to their native communication patterns, such as using neuralese, rather than forcing them to adapt to human language.

The death of SLMs?

Reddit r/LocalLLaMA

The author reflects on whether small language models under 27B are being overshadowed by larger models like Qwen 3.5 and Gemma 4, and asks the community for capable SLMs for agentic coding tasks.