@0x0SojalSec: You can Run locally Bonsai 2-27B uncensored on MacBook. - tok/s on a MacBook with 24 GB. - but base model is not gd wha…
Summary
A Twitter post discusses running the uncensored Ternary-Bonsai-27B AI model locally on a MacBook with 24 GB RAM, highlighting its performance in coding tasks.
View Cached Full Text
Cached at: 09/19/26, 09:01 AM
You can Run locally Bonsai 2-27B uncensored on MacBook.
- tok/s on a MacBook with 24 GB.
- but base model is not gd what i hope https://t.co/qW2FhwTMHa
Md Ismail Šojal 🕷️ (@0x0SojalSec): You can run locally a new Ternary-Bonsai-27B Uncensored on 7.2GB
- Refusals almost gone parent-level coding,
- abliterated the packed ternary weights
- Low-deg abliteration to max-HC scar models
- coding still 19/20
Research/offline use Still a general model run it under your
Similar Articles
I ran Ternary-Bonsai-27B (2-bit) and Bonsai-27B (1-bit) on Terminal-Bench 2.0, in 8GB VRAM
A user tested quantized 1-bit and 2-bit versions of the 27B-parameter Bonsai model on Terminal-Bench 2.0, achieving results within 8GB VRAM.
Using the Bonsai 27b 1b quant locally - regularly.
A user shares their positive experience using the 1-bit quantized version of Bonsai 27b locally on a 16GB MacBook Air for casual chat, tutoring in Go, and analyzing personal notes, praising its intelligence and small footprint.
@sudoingX: every day someone asks how i'm running bonsai 27b on hardware that shouldn't handle it. so here's the whole thing in on…
A detailed guide on running the 27B Bonsai model on hardware with only 8GB VRAM using a 1-bit quantized version and the PrismML fork of llama.cpp, including exact server commands and configuration.
Bonsai 27B: 1-bit dense LLM running locally in your browser using custom WebGPU kernels
Bonsai 27B is a 1-bit dense large language model that can run locally in a browser using custom WebGPU kernels, enabling efficient on-device inference.
@TheAhmadOsman: HOLYYYY 27B model under 6GBs and 4GBs Local AI will be the default P.S. We are gonna get this optimized in ODS by @Osma…
Ternary Bonsai 27B, a large language model, is demonstrated running locally on an NVIDIA RTX 5090 GPU, requiring under 6GB of memory and enabling end-to-end agentic workflows on consumer hardware.