MiniMax H3 open weights on Hugging Face: what is in the release
MiniMax H3 has a model card on Hugging Face: two checkpoints, a 33B-parameter Transformer, a community license and a 4-GPU serving example. What to check first.

Yes, MiniMax H3 has published weights: the MiniMaxAI/MiniMax-H3 page on Hugging Face describes a 33B-parameter dense, single-stream Transformer released under the MiniMax H3 Community License Agreement, with two task-specific checkpoints. The license has an application form for the USA, EU, UK and South Korea, so read it before you plan production use.
Everything below is from MiniMax's model card, its announcement and the ComfyUI docs, read 2026-09-29. For a hosted call instead of your own GPUs, Sume lists the model as minimax-h3 in the Video generation docs.
What is in the release?
The card describes three modules: H3-Context-IR for input preprocessing, H3-Base for 768p generation and H3-Regenerate-2K for higher resolution. It says the 2K step does not use a conventional super-resolution module; the base model regenerates its own low-resolution result in context, reusing the original multimodal context.
| Checkpoint | Inputs it supports |
|---|---|
| H3-Base-FL2VA | Zero, one or two input images: text-to-video, first-frame or last-frame video, or first-and-last-frame video |
| H3-Base-Ref2VA | Up to 9 images, up to 3 video clips, up to 3 audio clips, at most 12 files across all types |
What does it take to run?
The card's serving examples use SGLang with --num-gpus 4 and state BF16 precision; it gives no VRAM figure. It lists SGLang, vLLM, Diffusers and ComfyUI as frameworks. The ComfyUI docs describe a separate text encoder and video and audio VAEs plus optional turbo LoRAs that use 4 to 8 steps, and a default of 20 steps.
What were the announced dates?
MiniMax's announcement is dated 2026-07-31 and says MiniMax planned to open the weights in the coming days, subject to applicable laws and regulations. The model card carries no announcement date, so this post gives none for the weights.
When is a hosted API the easier route?
When you do not want to hold four GPUs or manage a license application. On Sume, POST /v1/videos with model: "minimax-h3" returns a job you poll, at native 480p or 768p with 2K and 4K as priced upscales. That is a different product from the weights: check the license terms for your region either way.
Sources
Related posts
More in Models
- AI video editing with a text prompt: MiniMax H3 and Sume's edit path
MiniMax H3 is described as editing existing video from instructions. What the vendor says, what Sume's H3 ids accept, and the id with an edit field.
- Nano Banana 2 Lite API: what Google shipped and what Sume lists
Nano Banana 2 Lite is Google's gemini-3.1-flash-lite-image model. What Google says it does, and which Nano Banana models Sume's image API lists today.
- Nano Banana Pro vs Nano Banana 2: which id to send on Sume
Nano Banana Pro and Nano Banana 2 share tiers and 10 references on Sume; they differ in extra aspect ratios and in which tiers change the price.
- Open-source video model vs API: which to use for Wan-class video
Weights you run yourself, or a hosted video API? What the choice changes for cost, setup and limits, with Wan 3.0's per-second API price as the worked example.
Written by Sume