27b-model

Tag

Cards List
#27b-model

@rohanpaul_ai: Beautiful visual of somebody running, qwen 3.8 27B locally on a RTX 5090 32 GB VRAM system with 115 tokens/sec note, Qw…

X AI KOLs Timeline · yesterday Cached

Tweet highlights running the Qwen 3.8 27B model locally on an RTX 5090 system with 32GB VRAM, achieving 115 tokens/sec, and notes the official BF16 checkpoint is 55.6GB.

0 favorites 0 likes
#27b-model

Qwen3.8-27B at 256K on a 24GB RTX PRO 4000 SFF (432 GB/s): 50 tok/s with MTP

Hacker News Top · yesterday Cached

The article details an experiment achieving 50 tokens per second inference with Qwen3.8-27B at 256K context on a 24GB GPU using Multi-Token Prediction and custom optimizations.

0 favorites 0 likes
#27b-model

Try out this "high" reasoning mode for 27B (tested on VLLM)

Reddit r/LocalLLaMA · 3d ago

The author experimented with the 27B model on VLLM and created a 'high' reasoning mode by blending prompts from low and xhigh modes, resulting in more efficient and enjoyable reasoning output.

0 favorites 0 likes
#27b-model

Got a 27B model running locally on a Jetson Orin NX 16GB (1-bit). still kind of amazed it works

Reddit r/LocalLLaMA · 2026-07-25

User reports successfully running a 27B parameter model quantized to 1-bit on a Jetson Orin NX 16GB edge device, expressing amazement at the feasibility.

0 favorites 0 likes
#27b-model

@TheAhmadOsman: HOLYYYY 27B model under 6GBs and 4GBs Local AI will be the default P.S. We are gonna get this optimized in ODS by @Osma…

X AI KOLs Timeline · 2026-07-14 Cached

Ternary Bonsai 27B, a large language model, is demonstrated running locally on an NVIDIA RTX 5090 GPU, requiring under 6GB of memory and enabling end-to-end agentic workflows on consumer hardware.

0 favorites 0 likes
#27b-model

@populartourist: I can run a 27B instead of a 9B on my budget laptop. That's mind blowing.

X AI KOLs Timeline · 2026-07-14 Cached

PrismML announces Bonsai 27B, a multimodal model based on Qwen3.6 27B that can run on a phone, enabling local multi-step reasoning, tool use, and long-context workflows.

0 favorites 0 likes
#27b-model

I built an autonomous dev pipeline and ran the same project head to head: a 27B local on a modded 4090, then again on cheap cloud LLMs

Reddit r/LocalLLaMA · 2026-06-30

The author built an autonomous development pipeline and benchmarked it by running the same project using a local 27B model on a modified RTX 4090 versus cheap cloud LLM APIs.

0 favorites 0 likes
#27b-model

@DJLougen: Proud to introduce a new 27B post-trained model After being impressed by both Fable and Kimi 2.7 Coder, I wanted to see…

X AI KOLs Timeline · 2026-06-20 Cached

Introduces a new 27B post-trained model that distills positives from Fable and Kimi 2.7 Coder, with links to download.

0 favorites 0 likes
← Back to home

Submit Feedback