NVIDIA drops DGX Station for Windows (1-Trillion Parameter desktop). Who else is ready to run LLaMA-Behemoth locally?

Reddit r/ArtificialInteligence Products

Summary

NVIDIA announced the DGX Station for Windows, a desktop supercomputer designed to run trillion-parameter AI models locally, with humorous commentary on quantization and community reaction.

Jensen just blessed us, folks. NVIDIA just announced a "desktop" supercomputer for Windows that can natively run a 1-Trillion parameter AI. They say it’s for "enterprise data scientists," but we all know what this is actually for: running uncensored Waifu chatbots at 500 tokens per second. Here is the **TL;DR** of the hardware specs: * **VRAM:** Enough to make a grown man cry (and finally stop daisy-chaining used Tesla P40s with zip-ties). * **Cooling:** Liquid-cooled. Doubles as a space heater. It will completely solve the winter heating bill for your entire neighborhood. * **Power:** Requires a direct line to your local nuclear power plant. * **Price:** Just your soul, your house, and a 50-year enterprise mortgage. # 🦙 The Real Question: Running LLaMA-Behemoth We all know Meta is going to drop **LLaMA-Behemoth-1T-Instruct** any day now. But let's be real about how this sub is actually going to handle it. Even with a multi-hundred-thousand-dollar DGX workstation on our desks, we are **still** going to aggressively quantize it because we refuse to close our 400 Chrome tabs while inferencing. **The** r/LocalLLaMA **Quantization Roadmap for LLaMA-Behemoth-1T:** |**Quantization Level**|**VRAM Needed**|**Intelligence Level**|r/LocalLLaMA **Verdict**| |:-|:-|:-|:-| |**FP16 (Unquantized)**|2000 GB|Absolute AGI. Cures cancer.|*"Waste of VRAM. Can't fit my 8k system prompt."*| |**Q4\_K\_M (GGUF)**|600 GB|Smarter than you.|*"Decent, but I want higher tokens/sec."*| |**IQ2\_XXS**|250 GB|High school dropout.|*"The sweet spot! Highly recommend!"*| |**IQ0\_0.001\_K\_Madness**|8 GB|Hallucinates that it is a toaster. Speaks only in binary.|*"Perfect! Runs flawlessly on my base M1 Mac at 120 t/s!"*| I'm already selling my kidneys to afford the down payment on this DGX Station. Can't wait to run the 1-bit quantization of Behemoth so it can confidently explain to me why 2+2=5 in 40 different languages simultaneously. Who else is pre-ordering?
Original Article

Similar Articles

NVIDIA AI Supercomputer Comes Online at Naval Postgraduate School

NVIDIA Blog

NVIDIA founder Jensen Huang commissions a DGX GB300 AI supercomputer at the Naval Postgraduate School, providing on-premises large-scale AI computing for students and faculty to advance research in weather prediction, cybersecurity, and disaster response.