Devs - you have 64gb of VRAM - which model do you use for coding?
Summary
A developer with 64GB VRAM shares their preference for an unsloth version of Qwen 3.5 122b-a10b for coding and asks the community for their recommendations.
Similar Articles
High VRAM local coding model — still Qwen 3.6 27B?
The user discusses their experience with Qwen 3.6 27B for local coding tasks and asks for recommendations for larger models (100B+) suitable for systems with 224GB of VRAM.
5090 + 96GB RAM, any better choice than Qwen3.8-27B for coding?
A user is asking for a better AI model than Qwen3.8-27B for coding tasks, noting its limitations in higher-level reasoning, system architecture, separation of concerns, and abstractions.
16 GB VRAM purgatory discussion thread
A discussion thread sharing configurations and tips for running AI models like Qwen3.8-27B on 16 GB VRAM Windows systems, focusing on memory optimization techniques.
Best general purpose uncensored or censored coding model with 6GB VRAM and 64GB of RAM?
A user seeks recommendations for the best uncensored or censored AI coding model to run locally on hardware with limited VRAM, aiming for faster response times and integration with development tools like Visual Studio and VS Code.
How many people have 24gb over gpu here?
The author discusses the low adoption of the qwen 3.8 27b model based on download counts and estimates that very few users have the high-VRAM GPUs needed for productive local LLM development.