Tag
Explores whether very small language models can handle casual conversations adequately, and what training factors differentiate the better ones.