Comfy-Org/MiniMax-H3
Summary
Comfy-Org repackaged MiniMax-H3 model files for ComfyUI, including diffusion models, text encoders, and VAEs, with workflow templates for text-to-video, image-to-video, and reference-to-video generation.
View Cached Full Text
Cached at: 08/03/26, 07:30 AM
Comfy-Org/MiniMax-H3 Β· Hugging Face
Source: https://huggingface.co/Comfy-Org/MiniMax-H3 Repackaged model files for ComfyUI.
Original model repository:https://huggingface.co/MiniMaxAI/MiniMax-H3
Place the files in the following folders:
π ComfyUI/
βββ π models/
β βββ π diffusion_models/
β β βββ minimax_h3_fl2va_bf16.safetensors
β β βββ minimax_h3_fl2va_int8_convrot.safetensors
β β βββ minimax_h3_fl2va_pruned_int8_convrot.safetensors
β β βββ minimax_h3_ref2va_bf16.safetensors
β β βββ minimax_h3_ref2va_int8_convrot.safetensors
β β βββ minimax_h3_ref2va_pruned_int8_convrot.safetensors
β βββ π text_encoders/
β β βββ qwen3vl_32b_minimax_h3_bf16.safetensors
β β βββ qwen3vl_32b_minimax_h3_int8_convrot.safetensors
β β βββ qwen3vl_32b_minimax_h3_nvfp4_awq.safetensors
β βββ π vae/
β β βββ minimax_h3_audio_vae_fp32.safetensors
β β βββ minimax_h3_video_vae_fp16.safetensors
https://huggingface.co/Comfy-Org/MiniMax-H3#workflowsWorkflows
Similar Articles
Kijai/MiniMax-H3_comfy
A Hugging Face repository hosting MiniMax-H3 models converted for ComfyUI usage, along with a Lightx2v distill LoRA for faster inference.
Comfy-Org/Mage-Flow
Repackaged model files for ComfyUI from the Microsoft Mage-Flow model, including diffusion models, text encoder, and VAE.
realrebelai/MiniMax-H3_GGUFs
Hugging Face repository providing GGUF quantizations of MiniMax-H3 models for use with ComfyUI, including directory structure and links to required VAEs.
Comfy-Org/ComfyUI
ComfyUI is an open-source, modular node-based AI engine for generating images, videos, 3D models, and audio, with support for state-of-the-art open-source models and cloud/desktop deployment.
drbaph/MiniMax-H3-Turbo-Lora-ComfyUI
Third-party ComfyUI-compatible LoRA conversions for MiniMax-H3 Turbo 4-step audio-video generation, including further-trained checkpoint-500 variants and an example workflow.