Xiaomi quietly uploaded MiMo-V2.5-DFlash — official DFlash weights are now on Hugging Face
Summary
Xiaomi has quietly released MiMo-V2.5-DFlash, a 300B-parameter model on Hugging Face, with DFlash potentially doubling inference speed.
Similar Articles
XiaomiMiMo/MiMo-V2.5-Pro-FP4-DFlash
XiaomiMiMo releases MiMo-V2.5-Pro-FP4-DFlash, an FP4-quantized MoE model with block-diffusion speculative decoding to reduce memory and bandwidth for trillion-parameter inference.
Xiaomi is now serving MiMo V2.5 at 1000-3000tps using DFlash & Persistent kernel. DFLash model is out, open-source release promised coming soon
Xiaomi has released MiMo V2.5 with DFlash and Persistent kernel, achieving 1000-3000 tps. The DFlash model is now available and open-source release is promised soon.
@zhijianliu_: DFlash for Qwen3.6-35B-A3B just dropped The community was running the day-1 preview before we even finished training. N…
Z-lab releases DFlash for Qwen3.6-35B-A3B, a model fine-tuning/compression technique, with training complete and weights now available on GitHub and HuggingFace.
Xiaomi Mimo-V2.5 Released, looks like today is big day for Open-Weight releases
Xiaomi released Mimo-V2.5, an open-weight AI model, adding to today’s string of open model drops alongside Qwen-27B.
XiaomiMiMo/MiMo-V2.5-Pro
Xiaomi releases MiMo-V2.5-Pro, an open-source MoE language model with 1.02T total parameters and 1M token context, optimized for complex agentic and software engineering tasks.