FastH3 Trim: MiniMax H3 on an 8 GB GPU, and who may use it
FastH3 Trim prunes MiniMax H3 to 42 blocks and 8 steps. File sizes, the 8 GB claim, the quality trade, and what the H3 Community License allows.

FastH3 Trim is an experimental, pruned version of MiniMax H3 that keeps 42 of the original 50 transformer blocks and samples in 8 steps, so video with synchronized audio can run on a single consumer NVIDIA GPU; one write-up reports it working on a memory-capped 8 GB RTX 4090. It is released under the MiniMax H3 Community License, which has territory and revenue conditions you should read before you ship clips made with it.
What was cut from MiniMax H3?
According to the model card, the FastVideo team removed eight blocks whose removal changed the video and audio predictions least, replaced each block's timestep conditioning with a shared rank-16 basis, and distilled the result to eight sampling steps. The card calls it 4.2 times smaller than base H3 and faster than FastH3 V2.
The same page labels it an experiment. The ComfyUI Wiki write-up quotes the developers saying that removing blocks makes the model faster and smaller but costs some quality. Treat it as a draft engine, not a replacement for the full model.
Which file fits which GPU?
The ComfyUI repack ships four formats. It needs ComfyUI 0.36.0 or later, the FastVideo FastH3 text-to-video template, a Qwen3VL text encoder and separate audio and video VAE files.
The 8 GB figure is not a file size. The NVFP4 file is 11.9 GB, so the 8 GB case relies on the memory-capped test the ComfyUI Wiki describes, where performance degrades. The same write-up gives roughly 43 to 83 seconds for a 5-second 480p clip on an RTX 4090, depending on the memory cap.
| Format | File size | Target hardware |
|---|---|---|
| NVFP4 | 11.9 GB | NVIDIA Blackwell (RTX 50 series, RTX PRO 6000) |
| FP8 | 19.7 GB | NVIDIA Ada and newer (RTX 40 series) |
| INT8 ConvRot | 18.9 GB | Recent NVIDIA GPUs (template default) |
| BF16 | 37.5 GB | Reference quality and format conversion |
Can you use the clips commercially?
The repack lists the minimax-h3-community-license-agreement. In the license text I read, use is limited to an Applicable Territory that excludes the EU, UK, South Korea and the USA, and it says you may not use the works or their Outputs outside that territory. Companies above $20 million in annual revenue need separate written authorization from MiniMax. Users must prominently display "MiniMax H3" on product interfaces, and Outputs cannot be used to improve another AI model.
That is a summary, not legal advice. If your team or your viewers are in an excluded region, get the license reviewed or contact MiniMax before you build on the weights.
When is a hosted workflow the practical choice?
If the box under your desk is not an NVIDIA card, or you need a clip you can ship without a territory check on your own infrastructure, a hosted route is simpler. Sume's Video Router lists minimax-h3, which the docs say accepts 5 to 15 seconds at native 480p or 768p. Sume bills at the provider list price with a 1.25 house margin, and you read the live capabilities from the catalog rather than assuming them.
That is a different product from FastH3 Trim: it is the hosted H3 row, not these pruned weights, and it has no local setup. Run Trim when you want free local drafts on your own GPU and have cleared the license. Call the hosted row when you want the clip and nothing else to maintain.
Sources
- FastVideo FastH3 Trim ComfyUI repack, Hugging Face (read 2026-10-11)
- FastVideo FastH3 Trim 8-Step model card, Hugging Face (read 2026-10-11)
- ComfyUI Wiki: FastH3 Trim (read 2026-10-11)
- MiniMax-H3 model card, Hugging Face (read 2026-10-11)
- MiniMax H3 Community License (read 2026-10-11)
- Video Router models
Related posts
More in Models
- Image releases of Oct 6-10, 2026: which have a Sume model id
Nano Banana 2.1 is google/nano-banana-2.1 on Sume. Qwen-Image-2.1-Turbo and Kroma have no id; qwen/qwen-image is the nearest Qwen row. Check yours in code.
- Kandinsky 6.0 video with sound: is it on Sume, and what to use
Kandinsky 6.0 makes video and audio together, MIT-licensed. Sume does not list it. Here is what Sume lists for synced sound, and when to self-host Kandinsky.
- Kandinsky 6.0 Video is MIT: can you use the clips commercially?
Kandinsky 6.0 Video Pro (29B) and Lite (3B) are MIT-licensed with joint audio. What MIT covers, what to check, and the hosted alternative on Sume.
- Kimi K3 on Sume: no catalog row, and what to pick for a video agent
Kimi K3 is not in Sume's agent model catalog. Moonshot says it takes text, images and video; here is the verified alternative and how Kimi can still call Sume.
Written by Sume