What is the biggest dense model that would fit into 128 GB RAM (at MXFP4)?

Reddit r/LocalLLaMA News

Summary

Discusses the largest dense model that can be loaded in 128 GB RAM using MXFP4 quantization.

No content available
Original Article

Similar Articles