Tag
oMLX 0.3.9.dev2 is released with improved Gemma 4 support, DFlash engine integration, and ParoQuant capabilities for local LLM inference on Apple Silicon.
Community testers evaluate quantized versions of Qwen3.6, ZAYA1, and other models for SVG chessboard generation accuracy using local inference frameworks like MLX.
OpenBMB has launched the MiniCPM-V 4.6 vision language model, which features immediate day-0 support on the MLX-VLM package for high-speed inference on Apple Silicon Macs.