Tag
Laguna s2.1 launched with impressive benchmarks but suffered from looping and tool usage issues; recent updates may have fixed them, and users are asking if it is now stable.
Modified SGLang to support Qwen and Laguna models on V100 GPUs using custom FlashAttention and Marlin kernels, achieving decent throughput on 4xV100 hardware.
Compares two new AI models for agentic workloads: the compact Nanbeige4.2-3B with a looped transformer architecture and the large Mixture-of-Experts Laguna S2.1, both released on Hugging Face.
Eiso Kant, co-founder of Poolside AI, discusses their 'Model Factory' approach to rapidly training frontier models, the release of Laguna S 2.1, and the economics of AI model development.
This pull request adds support for Laguna XS.2 & M.1 hardware in llama.cpp, expanding compatibility.