The dream is to reach 200GB VRAM

Reddit r/LocalLLaMA News

Summary

A user outlines a step-by-step plan to achieve 200GB of VRAM by combining multiple NVIDIA GPUs in a custom PC build, addressing purchase, installation, and power management.

Step 1) Find 16k ASAP before it goes up to 20k after a few months Step 2) Buy RTX PRO 6000 (MAXQ) Step 3) Remove RTX PRO 5000 in pcie_1 slot. Replace w/ RTX PRO 6000 Step 4) Buy a NVME to PCIE converter and HPPLEX 500W then move RTX PRO 5000 there Step 5) Power limit RTX PRO 6000, RTX 5090 and RTX PRO 4000 so it fits 1300W PSU ATX 3.1 4 GPUS RTX PRO 6000 (MAXQ) (96GB) gen5 x8 RTX 5090 (32GB) gen5 x8 RTX PRO 5000 (48GB) gen4 x4 RTX PRO 4000 (24GB) gen4 x4 =200GB VRAM !!! How to finish Step 1??
Original Article

Similar Articles

70-class VRAM stagnation

Reddit r/LocalLLaMA

The author observes that Nvidia's desktop 70-class GPUs have stayed at 12GB VRAM across two generations, and suggests Nvidia may be intentionally limiting memory to preserve demand for higher-margin AI-focused hardware.

RTX 2080 Ti Memory Upgrade to 22 GB

Hacker News Top

GPU Solutions offers a service to upgrade NVIDIA GeForce RTX 2080 Ti VRAM from 11 GB to 22 GB by replacing GDDR6 modules, including BIOS configuration and stability testing.

12GB VRAM gang, what's our plan?

Reddit r/LocalLLaMA

Discussion about running LLMs on 12GB VRAM, noting current focus on dense models like Muse Glimmer 30B and Qwen 3.8 27B, and questioning whether upgrading to 24GB VRAM is needed.

Ultra budget 20GB vram with 448GB/s for $100 bucks.

Reddit r/LocalLLaMA

Demonstrates achieving 20GB VRAM and 448GB/s bandwidth for around $100 using two NVIDIA P102-100 cards, running a llama.cpp server with a Qwen model and supporting 3 concurrent users with large context.