Tag
A user with limited VRAM hardware reviews the Laguna XS 2.1 model, finding it performant and better than alternatives like Qwen and Gemma for their use case, though not as polished.
LottoLabs announces LiquidAI's LFM2.5-8B-A1B-GGUF model, an 8B parameter model trained on a massive token count and optimized for fast inference on limited GPU hardware, with support for llama.cpp, Ollama, vLLM, and more.