bonsai

Tag

Cards List
#bonsai

you can now fine tune Prism-ML's ternary Bonsai models

Reddit r/LocalLLaMA · 4d ago

Prism-ML announces that their ternary Bonsai models can now be fine-tuned, with example code and a recommendation to use a high learning rate.

0 favorites 0 likes
#bonsai

@sudoingX: every day someone asks how i'm running bonsai 27b on hardware that shouldn't handle it. so here's the whole thing in on…

X AI KOLs Timeline · 4d ago Cached

A detailed guide on running the 27B Bonsai model on hardware with only 8GB VRAM using a 1-bit quantized version and the PrismML fork of llama.cpp, including exact server commands and configuration.

0 favorites 0 likes
#bonsai

1-Bit LLM in the Browser

Hacker News Top · 2026-07-16 Cached

A 1-bit LLM (Bonsai) is now runnable in the browser via WebGPU, enabling efficient on-device inference.

0 favorites 0 likes
#bonsai

@sudoingX: this lab took qwen 3.6 27b, the model i've been calling king of the 24gb tier all month, and crushed it down to 3.9gb. …

X AI KOLs Timeline · 2026-07-16 Cached

PrismML announces Bonsai 27B, a binary-quantized version of Qwen3.6 27B that runs on a phone using only 1.125 bits per weight, claiming 89.5% intelligence retention. The model is being independently tested by @sudoingX to verify performance.

0 favorites 0 likes
#bonsai

For those with 12GB GPUs, you can now run QWEN 3.6 27B wth little loss via the new Ternary version.

Reddit r/ArtificialInteligence · 2026-07-15

A new ternary quantized version of Qwen3.6 27B, called Bonsai 27B, allows running the model on 12GB GPUs with 10x less memory and 95% of original performance, making it accessible for local deployment.

0 favorites 0 likes
#bonsai

So what's the consensus on 1bit models? Is it still a pipe dream?

Reddit r/LocalLLaMA · 2026-07-14

Discussion on the feasibility of 1-bit models like Bonsai 8b and 27b, which achieve small file sizes while remaining functional, questioning their primary use cases and future.

0 favorites 0 likes
#bonsai

Prism-ML's Bonsai-27B Benchmarks

Reddit r/LocalLLaMA · 2026-07-14

Prism-ML published benchmarks for their Bonsai-27B model.

0 favorites 0 likes
#bonsai

@TheAhmadOsman: HOLYYYY 27B model under 6GBs and 4GBs Local AI will be the default P.S. We are gonna get this optimized in ODS by @Osma…

X AI KOLs Timeline · 2026-07-14 Cached

Ternary Bonsai 27B, a large language model, is demonstrated running locally on an NVIDIA RTX 5090 GPU, requiring under 6GB of memory and enabling end-to-end agentic workflows on consumer hardware.

0 favorites 0 likes
#bonsai

@populartourist: I can run a 27B instead of a 9B on my budget laptop. That's mind blowing.

X AI KOLs Timeline · 2026-07-14 Cached

PrismML announces Bonsai 27B, a multimodal model based on Qwen3.6 27B that can run on a phone, enabling local multi-step reasoning, tool use, and long-context workflows.

0 favorites 0 likes
#bonsai

Bonsai 27B (1-bit LLM): The First 27B-Class Model to Run on a Phone

Hacker News Top · 2026-07-14 Cached

PrismML announces Bonsai 27B, a 1-bit and ternary quantized version of Qwen3.6 27B that runs on phones and laptops, retaining 90-95% of baseline performance with a 3.9GB footprint, enabling agentic and multimodal on-device AI.

0 favorites 0 likes
#bonsai

Prism-ML Bonsai Qwen 3.6 27B

Reddit r/LocalLLaMA · 2026-07-14 Cached

Prism ML released Ternary-Bonsai-27B, a ternary-quantized version of Qwen3.6-27B that retains 95% of FP16 intelligence at a ~7.2 GB footprint, enabling full 27B-class reasoning on laptops and single GPUs with speeds up to 26 tok/s on Apple M5 Pro.

0 favorites 0 likes
#bonsai

prism-ml/Ternary-Bonsai-27B-mlx-2bit

Hugging Face Models Trending · 2026-07-04 Cached

Prism ML releases Ternary-Bonsai-27B-mlx-2bit, a ternary-quantized 27B-parameter language model that achieves ~95% of FP16 performance while fitting in ~7.2 GB, enabling full reasoning on laptops.

0 favorites 0 likes
#bonsai

Strace-ui, Bonsai_term, and the TUI renaissance

Hacker News Top · 2026-06-02 Cached

Jane Street engineers introduce strace-ui, an interactive terminal UI for strace that simplifies syscall debugging with filtering, PID tracking, and man page integration, and highlight the TUI renaissance enabled by their Bonsai framework.

0 favorites 0 likes
#bonsai

PrismML-Eng/Bonsai-demo

GitHub Trending (daily) · 2026-07-16 Cached

PrismML releases Bonsai 27B, a vision-language model with agentic tool calling and long context, along with 1-bit and ternary variants. The demo repository allows running these models locally on various hardware.

0 favorites 0 likes
← Back to home

Submit Feedback