@liquidai: Introducing LFM2.5-Embedding-350M and LFM2.5-ColBERT-350M: two multilingual retrieval models built for ultra-fast and a…
Summary
Liquid AI introduces LFM2.5-Embedding-350M and LFM2.5-ColBERT-350M, two multilingual retrieval models optimized for fast and accurate search across 11 languages, with latency as low as 1.5ms.
View Cached Full Text
Cached at: 06/18/26, 04:08 PM
Introducing LFM2.5-Embedding-350M and LFM2.5-ColBERT-350M: two multilingual retrieval models built for ultra-fast and accurate search across 11 languages.
End-to-end retrieval latency as low as 1.5ms with our enterprise stack!
Consistently best-in-class multilingual and cross-lingual performance across Arabic, German, English, Spanish, French, Italian, Japanese, Korean, Norwegian, Portuguese, and Swedish.
Similar Articles
LiquidAI/LFM2.5-ColBERT-350M
LiquidAI releases LFM2.5-ColBERT-350M, a late-interaction multilingual retrieval model, along with a dense bi-encoder variant, both built on LFM2.5-350M-Base, supporting 11 languages and designed as drop-in replacements for RAG pipelines.
LiquidAI/LFM2.5-Embedding-350M
Liquid AI releases LFM2.5-Embedding-350M, a dense bi-encoder for multilingual retrieval supporting 11 languages, as a drop-in replacement for RAG pipelines.
@maximelabonne: We just released two new encoder models (MLM) in 2026 They're super fast, easy to train, and strongly multilingual. Try…
Liquid AI released two new encoder models, LFM2.5-Encoder-230M and LFM2.5-Encoder-350M, that are fast, easy to train, and strongly multilingual, with speed benchmarks showing over 3.7x improvement on CPU compared to ModernBERT-base.
LFM2.5-Encoders for Fast Long-Context Inference on CPU
Liquid AI releases LFM2.5-Encoders (230M and 350M), efficient encoder models optimized for long-context inference on CPU, matching or beating larger encoders on benchmarks with 3.7x speedup over ModernBERT-base.
LiquidAI/LFM2.5-Encoder-350M
Liquid AI releases LFM2.5-Encoder-350M, a multilingual bidirectional encoder built on the LFM2 architecture, offering strong quality for its size, 8k context, and efficient on-device performance across 15 languages.