Anyone running qwen 3.8 27b on 5070ti (16GB)?

Reddit r/LocalLLaMA News

Summary

A user asks if running the Qwen 3.8 27b model on a 5070ti GPU with 16GB VRAM is feasible using quantization for agentic coding purposes.

Hi everyone. I recently decided to shell out a few bucks and upgrade my 4070ti (12GB) to a 5070ti (16GB). I'm wondering if there's a reasonable quant that I could run the new qwen 3.8 27b on and get decent tp/s, for agentic coding mainly. I heard that some 4bit quants are decent enough. Or am I still in the no-go territory? Is anyone rocking this card? 5070ti 16GB VRAM 32GB RAM DDR4
Original Article

Similar Articles

Ternary Qwen3.6 27B Tested on 3090!

Reddit r/LocalLLaMA

User tests ternary quantized Qwen3.6 27B on an RTX 3090, achieving 60 tk/s with two slots and 100k KV cache using 21GB VRAM, with good quality and stable tool calls.