China's MiniMax H3 is the first open model to top an AI video ranking (2 minute read)

TLDR AI Models

Summary

MiniMax released the open-weights H3 video model, the first open model to top an AI video ranking, ranking first in video editing and second in text-to-video. The 33B parameter model handles text, images, video, and audio, with some components like 2K resolution and H3-Context-IR remaining closed.

Artificial Analysis ranks MiniMax H3 first in Video Editing, second in Text-to-Video, and third in Image-to-Video.
Original Article
View Cached Full Text

Cached at: 08/04/26, 01:30 PM

# China's MiniMax H3 is the first open model to top an AI video ranking Source: [https://the-decoder.com/chinas-minimax-h3-is-the-first-open-model-to-top-an-ai-video-ranking/](https://the-decoder.com/chinas-minimax-h3-is-the-first-open-model-to-top-an-ai-video-ranking/) **MiniMax releases H3 video model weights, putting an open model at the top of a video ranking for the first time\.**[Artificial Analysis](https://x.com/ArtificialAnlys/status/2083042088338538594)ranks H3 first in Video Editing, second in Text\-to\-Video, and third in Image\-to\-Video\. The 33\-billion\-parameter model processes text, images, video, and audio together, generating four\- to 15\-second clips with stereo sound\.[According to the model card](https://huggingface.co/MiniMaxAI/MiniMax-H3), a single prompt can include up to nine reference images, three video clips, and three audio clips\. *Video by**MiniMax H3* Two pieces remain closed, though\. The 2K resolution module and H3\-Context\-IR, which translates prompts and reference material into a structured intermediate format, aren't included\. Running H3 locally in ComfyUI tops out at 768p, and users will need to handle context prep themselves using MiniMax's published prompting guides\. The open weights do allow fine\-tuning on custom footage, characters, or a specific visual style\. One catch on the license side: commercial use is only permitted for companies making under $20 million in revenue\. ByteDance released its closed[Seedance 2\.5](https://the-decoder.com/bytedances-seedance-2-5-generates-30-second-video-clips-with-built-in-audio/)the same day, which generates 30\-second clips with built\-in audio\. ### AI News Without the Hype – Curated by Humans Subscribe to THE DECODER for ad\-free reading, a weekly AI newsletter, our exclusive "AI Radar" frontier report six times a year, full archive access, and access to our comment section\. [Subscribe now](https://the-decoder.com/subscription/)

Similar Articles

MiniMax H3 (10 minute read)

TLDR AI

MiniMax launches H3, an open multimodal generation model that handles text, images, video, and audio, generating up to 15 seconds of 2K video with native stereo sound, and plans to open-source the weights.

MiniMax M3 (2 minute read)

TLDR AI

MiniMax introduces M3, the first open-weights model to combine coding, agentic, and multimodal capabilities with up to 1M context via sparse attention.