Daniel Han of Unsloth validates Qwen3.8-27B will run only 17GB VRAM

Reddit r/LocalLLaMA Models

Summary

Daniel Han of Unsloth validates that Qwen3.8-27B will run in only 17GB VRAM, making it accessible for local inference.

Super excited about this release for the new 27B. Who else is with me. Only 17GB VRAM needed 😍😍
Original Article

Similar Articles

My Qwen 3.8 27B tests on limited VRAM (16-20GB)

Reddit r/LocalLLaMA

This article tests various quantized versions of the Qwen 3.8 27B AI model on limited VRAM setups, comparing their performance on tasks like animation generation, app development, and word generation.

Running Qwen3.6 35b a3b on 8gb vram and 32gb ram ~190k context

Reddit r/LocalLLaMA

The author shares a high-performance local inference configuration for running Qwen3.6 35B A3B on limited hardware (8GB VRAM, 32GB RAM) using a modified llama.cpp with TurboQuant support, achieving ~37-51 tok/sec with ~190k context.