openbmb

Tag

Cards List
#openbmb

openbmb released MiniCPM5-2B, not yet available at huggingface

Reddit r/LocalLLaMA · 2026-07-20

OpenBMB released MiniCPM5-2B, a 2 billion parameter language model, currently not yet available on Hugging Face.

0 favorites 0 likes
#openbmb

@VukRosic99: Long-context Transformers hit two walls: quadratic attention compute and a KV cache that reaches hundreds of GB at 1M t…

X AI KOLs Timeline · 2026-07-11 Cached

MiniCPM-SALA is a 9B-parameter hybrid attention model that interleaves sparse and linear attention to overcome the quadratic compute and large KV cache bottlenecks of long-context Transformers. It achieves 3.5x faster inference than Qwen3-8B at 256K tokens and supports up to 1M tokens on consumer GPUs, with a cost-effective continual training approach that reduces training costs by ~75%.

0 favorites 0 likes
#openbmb

@OpenBMB: Just a quick reminder: Build Small Hackathon sign-up closes on June 3! Total cash prizes: ~$40K $10K @OpenBMB Special A…

X AI KOLs Following · 2026-06-01 Cached

OpenBMB is hosting the Build Small Hackathon with $40k+ in prizes, focusing on building apps using small models (≤32B parameters) with Gradio on Hugging Face Spaces. Registration closes June 3, 2026.

0 favorites 0 likes
#openbmb

MiniCPM5-1B Shows Why the Small-Model Race Isn't Over

Reddit r/ArtificialInteligence · 2026-05-31 Cached

MiniCPM5-1B is a 1B parameter model from OpenBMB that achieves impressive scores on AIME 2025 and τ2-Bench Telecom, outperforming larger models. It features both fast and reasoning modes from a single checkpoint, enabled by a three-stage post-training process including supervised fine-tuning, reinforcement learning, and on-policy distillation.

0 favorites 0 likes
#openbmb

OpenBMB presents the model BitCPM-CANN 1.58 bit

Reddit r/LocalLLaMA · 2026-05-22

OpenBMB introduced BitCPM-CANN, a 1.58-bit model being tested on Huawei Ascend 910B hardware.

0 favorites 0 likes
← Back to home

Submit Feedback