@no_stp_on_snek: Very nice. Huge for team 3090. And TurboQuant+ is already implemented in a bunch of inference engines.

X AI KOLs Following Models

Summary

A reply celebrates Unsloth AI's upcoming Qwen3.8-27B model, which will run on 17GB RAM/VRAM setups, and notes TurboQuant+ is already integrated into many inference engines — great news for RTX 3090 users.

Very nice. Huge for team 3090. And TurboQuant+ is already implemented in a bunch of inference engines.
Original Article
View Cached Full Text

Cached at: 08/03/26, 03:46 PM

Very nice. Huge for team 3090. And TurboQuant+ is already implemented in a bunch of inference engines.

Unsloth AI (@UnslothAI): Qwen3.8-27B is coming! 🔥

Will run locally on 17GB RAM/VRAM setups.

Similar Articles

Wow! Qwen 3.6:35b-a3b on a 3090... pretty amazing.

Reddit r/artificial

A user shares impressive results running a quantized Qwen 3.6:35b-a3b model on a used RTX 3090, achieving 160 tokens per second output after fitting the model into VRAM, and demonstrates vision capabilities with a 75-second video processing time.