Tag
A Hugging Face repository hosting MiniMax-H3 models converted for ComfyUI usage, along with a Lightx2v distill LoRA for faster inference.
LightX2V releases an open, local LoRA prompt rewriter for MiniMax-H3 text-to-audio-video generation, fine-tuned on Qwen3.6-27B to expand short prompts into structured audio-video descriptions.
An early-preview LoRA for MiniMax-H3 enables 4-step audio-video generation instead of ~20 steps, yielding roughly 5x faster sampling, though quality is still immature.
Third-party ComfyUI-compatible LoRA conversions for MiniMax-H3 Turbo 4-step audio-video generation, including further-trained checkpoint-500 variants and an example workflow.
Sharing a fun website animation trick: use the AI video generation model MiniMax H3 to create a 240-frame sprite sheet, switch background position based on mouse angle, achieve a smooth dog-head-follows-mouse effect, and extend it to creative uses like scroll-controlled playback/rewind.
An early-preview LoRA for MiniMax-H3 that enables joint video and synchronized audio generation in 4 sampling steps instead of ~20, offering roughly 5x faster sampling, with ComfyUI custom nodes and three bf16 checkpoints.
This article introduces a "6+5" system for controlling AI characters' micro-expression performances in videos. It includes 6 action parameters and 5 emotional curves, and provides ready-to-copy prompt templates, case studies, and troubleshooting methods.
Detailed explanation of how to run the MiniMax H3 video generation model locally on an RTX 4080 16GB using ComfyUI's native workflow, including model download, directory configuration, parameter tuning, and common pitfalls.
Release of an NVFP4-quantized uncensored MiniMax-H3 text encoder (Qwen3-VL-32B Heretic) that fits on a single 16GB GPU and serves as a drop-in replacement in ComfyUI workflows.
MiniMax releases H3, a powerful open-weight 33B omni-modal video model supporting text, images, video, and audio, with local deployment on consumer GPUs and significantly lower cost than competitors like Seedance.
MiniMax released the open-weights H3 video model, the first open model to top an AI video ranking, ranking first in video editing and second in text-to-video. The 33B parameter model handles text, images, video, and audio, with some components like 2K resolution and H3-Context-IR remaining closed.
The author built a local video generation studio using MiniMax's H3 model, thanking MiniMax for the state-of-the-art model.
Announcement that GGUF quantizations of MiniMax H3 are available, with the Q2 version being only 8.49 GB for lower-end GPUs.
MiniMax H3, a next-generation open-weights video model capable of generating 2K video with native stereo audio from text, images, video, or audio, launched with day-0 ComfyUI support and optimizations that allow it to run on consumer GPUs.
The author ran same-conditions tests of MiniMax H3 and LTX 2.3 Eros video generation models on DGX Spark. Results show H3 is stronger in narrative and instruction following and includes audio, while LTX is about 1.76x faster.
This Hugging Face repository provides community-compiled quantized and pruned weights for MiniMax H3 (Hailuo 3.0), enabling local text/image/audio-to-video generation on consumer GPUs with 16-24GB VRAM. It includes INT4, INT8, and NVFP4 variants with hardware-specific guides.
This repository provides INT8 ConvRot-quantized ComfyUI safetensors of Qwen3-VL-32B, including a MiniMax-H3 conditioning encoder with layers 0-49 and an optional prompt-enhancement tail for layers 50-63, designed for use in ComfyUI on 32GB GPUs.
Comfy-Org repackaged MiniMax-H3 model files for ComfyUI, including diffusion models, text encoders, and VAEs, with workflow templates for text-to-video, image-to-video, and reference-to-video generation.