Open-weights video or hosted API? Four checks, Wan, LTX, Runway
Self-host Wan2.2 or LTX-2.5, or pay Runway, Luma or Sume per clip? Four checks from the vendors' own pages: hardware, license, price per clip and operations.

Choose open weights when you need control over the model, data path or fine-tuning and can run GPUs; choose a hosted API when you want clips on demand without operating infrastructure. Four checks decide it: hardware, license, price per clip, and who runs the service. The vendor pages below give the real numbers for Wan2.2, LTX-2.5, Luma and Runway.
Check one: hardware
The Wan2.2 README (read 2026-10-02) says the TI2V-5B model needs at least 24 GB of VRAM and the 14B models want 80 GB or more on one GPU, or multiple GPUs with FSDP. The LTX-2.5 card (read 2026-10-02) lists a 22B transformer and recommends Python 3.12, CUDA 12.7 and PyTorch about 2.7.
Check two: license
Wan2.2 is listed under Apache 2.0 on its repository. LTX-2.5 is free for commercial use under $10M annual revenue under the LTX-2.x Community License, with a paid agreement above that. Open weights are not all open on the same terms, so read each one.
Check three: price per clip
For self-hosting, price is GPU hours divided by clips per hour. Measure clips per hour on your own prompts before you trust any estimate.
| Vendor | What the page lists |
|---|---|
| Luma API, Ray 3.2 | $0.30 for a 5 s 720p text or image to video clip; $1.20 at 1080p |
| Runway, Standard monthly | $15 for 625 credits per month |
| Runway, Max monthly | $95 for 9,500 credits per month |
| Sume | Per-model prices on the API pricing page |
Check four: who runs it
Hosted APIs bring queues, storage, retries and moderation. Sume jobs are asynchronous: submit to POST /v1/videos, poll the job, then download the result, per the video docs. Self-hosting means building those yourself.
Sume's catalog does not include Wan2.2 or LTX weights; it hosts catalog models such as Wan 3.0 (2 to 30 seconds per clip).
A rule of thumb
Recheck every number here on the vendor page before you decide.
- Prototype on a hosted API; the first clips cost cents, not a GPU contract.
- Move to open weights when volume, privacy or customization justify the operations work.
- Keep your prompts and references portable so you can switch.
Sources
Related posts
More in Comparisons
- OpenAI Batch 200 MB input file vs Sume's 4 MiB body: size your items
OpenAI's Batch API takes a 200 MB JSONL file. A Sume create body is capped at 4 MiB and input at 2 MiB, so media goes by URL and bulk items stay small.
- OpenAI 2,000 batches an hour vs Sume's write budget per minute
OpenAI's Batch API allows 2,000 batch creations per hour. Sume budgets requests per minute by plan: 120 writes on Free to 1,200 on Scale.
- OpenAI custom voices: consent phrase, 30 s sample, 20 voice cap
OpenAI custom voices need a recorded consent phrase, a sample of 30 seconds or less and sales enablement. What that means next to cloning a voice in Sume.
- OpenRouter key credit limit and 402 vs a Sume per-run spend cap
OpenRouter caps a key's credits and reports limit_remaining. A Sume Format run carries its own per-run generation cap. Where each guardrail sits and what fails.
Written by Sume