27b

Tag

Cards List
#27b

Qwen 3.8 27b with tools and directed search on a non-coding professional suite

Reddit r/LocalLLaMA · 2026-08-25

A user's evaluation demonstrates that Qwen 3.8 27b performs excellently on professional certification tests in real estate and finance, achieving scores up to 98.44% with tools and guided search.

0 favorites 0 likes
#27b

Bro wtf, Qwen Lab cooked with Qwen 3.8 27B, it's so fucking good

Reddit r/LocalLLaMA · 2026-08-22

Qwen Lab has released Qwen 3.8 27B, which shows significant improvement over previous versions like Qwen 3.6 27B and other open-source models. The author hopes that Qwen publishes papers to help other labs develop similar high-quality small models.

0 favorites 0 likes
#27b

@no_stp_on_snek: Awe yeah… queuing up tests now

X AI KOLs Following · 2026-08-04 Cached

A developer excitedly queues tests for Qwen's new 27B model, which Qwen promises brings a whole new level of capability.

0 favorites 0 likes
#27b

Daniel Han of Unsloth validates Qwen3.8-27B will run only 17GB VRAM

Reddit r/LocalLLaMA · 2026-08-03

Daniel Han of Unsloth validates that Qwen3.8-27B will run in only 17GB VRAM, making it accessible for local inference.

0 favorites 0 likes
#27b

Can't wait to see Qwen3.8-27B

Reddit r/LocalLLaMA · 2026-08-03

Qwen announced Qwen3.8, including a new 27B model, generating excitement for local deployment.

0 favorites 0 likes
#27b

I built a tool to actually test which weights matter before quantizing, instead of guessing (Qwen3.6-27B, 3 builds: Bedrock/Tightrope/Gambit)

Reddit r/LocalLLaMA · 2026-07-28

A developer built a testing harness that measures KL divergence per weight group during quantization, leading to three custom quantized builds of Qwen3.6-27B (Bedrock, Tightrope, Gambit) with optimized compression. Tool calling is identified as the first capability to degrade under quantization.

0 favorites 0 likes
#27b

DavidAU/Qwen3.6-27B-Fable-Fusion-711-Uncensored-Heretic-NM-DAU-NEO-MAX-MTP-GGUF

Hugging Face Models Trending · 2026-07-17 Cached

DavidAU releases Qwen3.6-27B-Fable-Fusion-711, a multi-stage fine-tune of Qwen 3.6 27B that claims to exceed 700 ARC-C benchmark, surpassing base models and matching closed-source models, available in GGUF format for consumer hardware.

0 favorites 0 likes
#27b

DFlash makes Qwen3.6 27B 2.2x faster with no quality loss

Reddit r/LocalLLaMA · 2026-07-16

DFlash is a method that accelerates Qwen3.6 27B model inference by 2.2x without quality degradation.

0 favorites 0 likes
#27b

Prism-ML's Bonsai-27B Benchmarks

Reddit r/LocalLLaMA · 2026-07-14

Prism-ML published benchmarks for their Bonsai-27B model.

0 favorites 0 likes
#27b

Ternary Qwen3.6 27B Tested on 3090!

Reddit r/LocalLLaMA · 2026-07-14

User tests ternary quantized Qwen3.6 27B on an RTX 3090, achieving 60 tk/s with two slots and 100k KV cache using 21GB VRAM, with good quality and stable tool calls.

0 favorites 0 likes
#27b

Qwen3.6 27B on a 5090, 6.4k sample tok/s distribution after tuning MTP/cache settings

Reddit r/LocalLLaMA · 2026-07-04

Running Qwen3.6 27B on an RTX 5090, achieving 6.4k tokens per second after tuning MTP and cache settings, demonstrating optimization techniques for inference.

0 favorites 0 likes
#27b

prism-ml/Ternary-Bonsai-27B-mlx-2bit

Hugging Face Models Trending · 2026-07-04 Cached

Prism ML releases Ternary-Bonsai-27B-mlx-2bit, a ternary-quantized 27B-parameter language model that achieves ~95% of FP16 performance while fitting in ~7.2 GB, enabling full reasoning on laptops.

0 favorites 0 likes
#27b

Dspark with Qwen 3.6 27b?

Reddit r/LocalLLaMA · 2026-07-03

Mention of Qwen 3.6 27b model in context of Dspark.

0 favorites 0 likes
#27b

@SlimTradeyBaby: Drop your GPU below and I’ll tell you exactly what model and config to run on it. JOKES. No need. Qwen 3.6 27b @Unsloth…

X AI KOLs Timeline · 2026-06-20 Cached

A tweet promoting the Qwen 3.6 27b model and recommending UnslothAI for running it on any GPU.

0 favorites 0 likes
#27b

@WaleedAhmad1a10: Check out the Qwen 3.5 27B MoQ GGUFs :

X AI KOLs Following · 2026-06-16 Cached

A Hugging Face repository (kaitchup/Qwen3.6-27B-GGUF-MoQ) provides GGUF quantized weights for the Qwen3.6-27B MoQ model, enabling local inference with tools like llama.cpp and Ollama.

0 favorites 0 likes
#27b

Jackrong/Qwopus3.6-27B-Coder-MTP-GGUF

Hugging Face Models Trending · 2026-06-11 Cached

A GGUF quantized version of the Qwopus3.6-27B-Coder-MTP model is released on Hugging Face, optimized for local inference and compatible with Transformers, vLLM, SGLang, and Unsloth Studio.

0 favorites 0 likes
#27b

MooreThreads/MusaCoder-27B • Huggingface

Reddit r/LocalLLaMA · 2026-06-10

MooreThreads releases MusaCoder-27B, a 27-billion-parameter code generation model, accompanied by a paper on arXiv.

0 favorites 0 likes
#27b

Okay 27B made me a believer

Reddit r/LocalLLaMA · 2026-05-26

User shares experience with Qwen3.6 27B model, which successfully generated a complete HTML5 breakout game in one shot, showing impressive coherence and attention to detail beyond typical LLM outputs.

0 favorites 0 likes
#27b

@DeepTechTR: Qwen 3.6 27B is incredibly fast with 16 GB VRAM! The impact of Pure Quant The era of the 27B model that runs seamlessly…

X AI KOLs Timeline · 2026-05-24 Cached

Qwen 3.6 27B runs fast on 16 GB VRAM thanks to 'Pure Quant' technology, achieving 40 tokens/s with MTP and supporting 64k contexts, enabling local AI on consumer GPUs like RTX 4060 Ti.

0 favorites 0 likes
#27b

Qwen will release another 27B with high probability

Reddit r/LocalLLaMA · 2026-05-20

Qwen is highly likely to release a 27B parameter model, though the exact roadmap is still pending.

0 favorites 0 likes
Next →
← Back to home

Submit Feedback