@AdinaYakup: MiniCPM V4.6 a 1B MLLM that actually runs on your phone, just released by @OpenBMB 1B - Apache2.0 Runs on iOS, Android,…

X AI KOLs Following Models

Summary

OpenBMB has released MiniCPM V4.6, a 1B-parameter multimodal large language model optimized for mobile devices under the Apache 2.0 license. It features mixed visual token compression and claims approximately 1.5x faster throughput than Qwen3.5 0.8B while running natively on iOS, Android, and HarmonyOS.

MiniCPM V4.6 🔥 a 1B MLLM that actually runs on your phone, just released by @OpenBMB ✨ 1B - Apache2.0 ✨ Runs on iOS, Android, HarmonyOS ✨ ~1.5× faster throughput than Qwen3.5 0.8B ✨ Mixed 4x/16x visual token compression https://t.co/KBUal2oUf2
Original Article
View Cached Full Text

Cached at: 05/11/26, 04:42 PM

MiniCPM V4.6 🔥 a 1B MLLM that actually runs on your phone, just released by @OpenBMB

✨ 1B - Apache2.0 ✨ Runs on iOS, Android, HarmonyOS ✨ ~1.5× faster throughput than Qwen3.5 0.8B ✨ Mixed 4x/16x visual token compression https://t.co/KBUal2oUf2

Similar Articles

MiniCPM-V 4.6

Product Hunt

MiniCPM-V 4.6 is an ultra-efficient 1.3B vision-language model optimized for mobile devices.

MiniCPM4: Ultra-Efficient LLMs on End Devices

Papers with Code Trending

MiniCPM4 is a highly efficient large language model designed for end devices, achieving strong performance with 0.5B and 8B parameter versions through innovations in sparse attention, data filtering, training algorithms, and inference systems.

MiniCPM-V 4.5: Cooking Efficient MLLMs via Architecture, Data, and Training Recipe

Papers with Code Trending

MiniCPM-V 4.5 is an 8B multimodal large language model that achieves high efficiency and strong performance through a unified 3D-Resampler architecture, a novel data strategy, and a hybrid reinforcement learning approach. The model reportedly surpasses larger proprietary and open-source benchmarks while significantly reducing GPU memory usage and inference time.