@liquidai: Introducing LFM2.5-230M: our smallest model yet, built to run fast anywhere (CPUs, NPUs, and GPUs) to enable agentic ta…
Summary
Liquid AI releases LFM2.5-230M, a small 230M parameter model optimized for fast inference on CPUs, NPUs, and GPUs, targeting agentic tasks on devices like phones and robots.
View Cached Full Text
Cached at: 06/25/26, 03:25 PM
Introducing LFM2.5-230M: our smallest model yet, built to run fast anywhere (CPUs, NPUs, and GPUs) to enable agentic tasks on phones, robots, home and network automation devices.
230M parameters, built on the LFM2 architecture Pre-trained on 19T tokens, with a 32K context extension Post-trained with distillation from LFM2.5-350M 213 tok/s decode speed on Galaxy S25 Ultra (CPU) 42 tok/s on a Raspberry Pi 5 (CPU) Competes with and often beats models more than twice its size on instruction following, data extraction, and tool use. use it for large-scale data extraction pipelines or lightweight on-device agentic workloads.
Similar Articles
LiquidAI/LFM2.5-230M
Liquid AI released LFM2.5-230M, a compact 230M-parameter hybrid model optimized for on-device deployment with fast edge inference speeds (213 tok/s on Galaxy S25 Ultra) and built for agentic tasks via reinforcement learning.
LiquidAI/LFM2.5-2.6B
Liquid AI released LFM2.5-2.6B, a 2.6B-parameter hybrid model optimized for on-device deployment with 128K context, agentic post-training, and fast inference (220 tok/s on Apple M5 Max) under 2.5GB memory.
When you don't have a data center GPU
LiquidAI releases LFM2.5-230M, a 230M parameter language model designed to run on limited hardware, with support for transformers, vLLM, and SGLang.
Deploy local agents everywhere with LFM2.5-2.6B
Liquid AI releases LFM2.5-2.6B, a compact agentic model designed for on-device deployment, supporting tool calling and multi-step workflows with efficient inference on CPUs and GPUs.
Liquid AI Releases Liquid Foundation Models 2.5 230M (3 minute read)
Liquid AI releases LFM2.5-230M, a lightweight foundation model that runs on devices from cloud GPUs to CPUs and Raspberry Pi, with strong performance on tool use and data extraction tasks.