MiniMax H3 becomes first open-weight AI video model to lead a benchmark ranking
AI video model releases
Sign in to create alerts.
MiniMax H3 becomes first open-weight AI video model to lead a benchmark ranking
MiniMax H3, a 33-billion-parameter video model, ranks first in Video Editing, second in Text-to-Video, and third in Image-to-Video on Artificial Analysis, the first open model to top such a ranking.
H3 processes text, images, video, and audio together, generating clips from 4 to 15 seconds with stereo sound, and a single prompt can include up to 9 reference images, 3 video clips, and 3 audio clips.
Two components stay closed: the 2K resolution module and H3-Context-IR, which converts prompts and reference material into a structured intermediate format; running H3 locally in ComfyUI caps out at 768p.
The open weights allow fine-tuning on custom footage, characters, or visual styles, but the license restricts commercial use to companies earning under $20M in revenue.
ByteDance released its closed Seedance 2.5 model the same day, generating 30-second clips with built-in audio.