@aisearchio: Minimax H3 GGUFs are here! Q2 is only 8.49 GB, so you could fit it in lower-end GPUs. https://huggingface.co/realrebela…

X AI KOLs Timeline Models

Summary

Announcement that GGUF quantizations of MiniMax H3 are available, with the Q2 version being only 8.49 GB for lower-end GPUs.

Minimax H3 GGUFs are here! Q2 is only 8.49 GB, so you could fit it in lower-end GPUs. https://huggingface.co/realrebelai/MiniMax-H3_GGUFs/tree/main…
Original Article
View Cached Full Text

Cached at: 08/03/26, 11:55 PM

Minimax H3 GGUFs are here! Q2 is only 8.49 GB, so you could fit it in lower-end GPUs. https://huggingface.co/realrebelai/MiniMax-H3_GGUFs/tree/main…


realrebelai/MiniMax-H3_GGUFs at main

Source: https://huggingface.co/realrebelai/MiniMax-H3_GGUFs/tree/main realrebelai’s picture

realrebelai

Upload MiniMax-H3-FL2VA-Q2_K-(Mixed_Precision).gguf

bc780d9

verified

about 9 hours ago

Similar Articles

realrebelai/MiniMax-H3_GGUFs

Hugging Face Models Trending

Hugging Face repository providing GGUF quantizations of MiniMax-H3 models for use with ComfyUI, including directory structure and links to required VAEs.

Unsloth Minimax M3 GGUF

Reddit r/LocalLLaMA

Unsloth is uploading a GGUF quantized version of the MiniMax M3 model to Hugging Face.

sakamakismile/Qwen3-VL-32B-Heretic-MiniMax-H3-NVFP4

Hugging Face Models Trending

Release of an NVFP4-quantized uncensored MiniMax-H3 text encoder (Qwen3-VL-32B Heretic) that fits on a single 16GB GPU and serves as a drop-in replacement in ComfyUI workflows.