Tag
Prism-ML announces that their ternary Bonsai models can now be fine-tuned, with example code and a recommendation to use a high learning rate.
Prism-ML published benchmarks for their Bonsai-27B model.
Prism ML is in talks with Apple to deploy model-shrinking technology that would allow powerful AI models to run directly on iPhones, improving privacy and reducing reliance on cloud inference.
Prism ML released Ternary-Bonsai-27B, a ternary-quantized version of Qwen3.6-27B that retains 95% of FP16 intelligence at a ~7.2 GB footprint, enabling full 27B-class reasoning on laptops and single GPUs with speeds up to 26 tok/s on Apple M5 Pro.
Prism ML releases Ternary-Bonsai-27B-mlx-2bit, a ternary-quantized 27B-parameter language model that achieves ~95% of FP16 performance while fitting in ~7.2 GB, enabling full reasoning on laptops.