MiniMax H3
Open-weights video model from MiniMax generating native 2K clips up to 15 seconds with synchronized audio and multi-modal reference conditioning.
Updated 2026-08-01
Yes, within limits. MiniMax H3 runs a free tier with paid plans above it.
Listed pricing: Free tier + token-based API pricing (check site for current rates).
Overview
MiniMax H3 is the Chinese lab's next-generation video model, released on July 31, 2026. It generates clips at native 2K up to roughly 15 seconds, and the headline capability is that it takes text, image, video and audio as inputs to the same prompt rather than making you pick a mode. You can hand it a reference image for a character, a reference clip for motion or framing, and a text instruction for what should happen — the model resolves all of them together. MiniMax calls this omni-reference; in practice it's the feature that separates H3 from the text-to-video-plus-optional-first-frame pattern most competitors still use.
The second thing that matters is audio. H3 produces sound synchronized to the generated footage in the same pass, which puts it in the small group of models — Veo 3 being the obvious comparison — that don't leave you compositing a soundtrack afterwards. Native 2K also means you aren't upscaling 720p output before it's usable anywhere near a real timeline. It's already available through third-party inference hosts including fal.ai, so you can hit it via API without going through MiniMax's own platform.
The audience here is developers and studios building video generation into a product, not casual users looking for a web app to play with — this is a model release first. Open weights are the strategic differentiator: if you can run it yourself, you can fine-tune on your own footage, keep client material off a third-party server, and avoid per-generation API costs at volume. That's a meaningfully different proposition from Veo 3 or Kling, both of which are closed. The tradeoff is that a 2K video model with audio is not something you casually self-host, and MiniMax's tooling and documentation remain thinner than what US labs ship.
Is MiniMax H3 free?
Yes, within limits. MiniMax H3 runs a free tier with paid plans above it.
What the free tier covers: Yes — a limited free tier is available for evaluation. Check minimax.io for current generation limits.
Listed pricing: Free tier + token-based API pricing (check site for current rates).
Pricing on this page has not been re-verified. The entry was last edited on , and no separate pricing check has been run since. Treat the figures as a record of what was published then and confirm on the official site.
What the free tier leaves out
Read straight off the plan list below. Vendors move features between tiers, so check the current split before you pay.
- API (token-based) Usage-based Pay-per-generation pricing through MiniMax's platform and third-party hosts such as fal.ai. Rates vary by resolution and clip length — check the official pricing page for current numbers.
- Self-hosted Your own compute Open weights can be downloaded and run on your own GPUs with no per-generation fee. Hardware requirements for 2K video with audio are substantial.
MiniMax H3 pricing
| Plan | Price | What's included |
|---|---|---|
| Free tier | $0 | Limited trial generations through MiniMax's platform to evaluate output quality before committing to API spend. |
| API (token-based) | Usage-based | Pay-per-generation pricing through MiniMax's platform and third-party hosts such as fal.ai. Rates vary by resolution and clip length — check the official pricing page for current numbers. |
| Self-hosted | Your own compute | Open weights can be downloaded and run on your own GPUs with no per-generation fee. Hardware requirements for 2K video with audio are substantial. |
Limited trial generations through MiniMax's platform to evaluate output quality before committing to API spend.
Pay-per-generation pricing through MiniMax's platform and third-party hosts such as fal.ai. Rates vary by resolution and clip length — check the official pricing page for current numbers.
Open weights can be downloaded and run on your own GPUs with no per-generation fee. Hardware requirements for 2K video with audio are substantial.
2 plans carry no published number: API (token-based), recorded as “Usage-based”; Self-hosted, recorded as “Your own compute”. Nothing is estimated in its place. The vendor's own pricing page is the only source for those figures.
Pricing on this page has not been re-verified. The entry was last edited on , and no separate pricing check has been run since. Treat the figures as a record of what was published then and confirm on the official site.
Is MiniMax H3 worth it?
You can test that on the free tier before paying anything. The recorded trade-offs are listed below, and any one of them can settle the question on its own.
The 8.2/10 AI Score is an editorial read of published capability, price and shipping pace. Nobody here has hands-on hours with MiniMax H3. How we verify.
Worth it if
The strengths recorded against this entry.
- Open weights make self-hosting and fine-tuning possible — rare for a video model at this capability level
- Combines text, image, video and audio references in one prompt instead of forcing a single conditioning mode
- Native 2K output skips the upscaling step most generators still require
- Generates synchronized audio in the same pass, not as a separate job
- Already served by third-party inference hosts like fal.ai, so you can test it without onboarding to MiniMax's platform
Not worth it if
Any one of these blocks your use case.
- The roughly 15-second ceiling means anything longer has to be stitched from multiple generations, with the usual continuity problems
- Self-hosting a 2K video model with audio needs serious GPU capacity — the open weights are practically out of reach for individuals
- Token-based API pricing makes per-project cost hard to forecast compared to a flat monthly plan
- Documentation and developer tooling are thinner than Google's or Runway's, and Chinese-lab origin raises compliance questions for some enterprise buyers
Key features
Omni-reference inputs
Accepts text, image, video and audio references in a single prompt and reconciles them, so you can lock a character's appearance from a still while borrowing camera motion from a clip. Most rival models handle one conditioning signal at a time.
Native 2K generation
Outputs at 2K resolution directly rather than generating at 720p or 1080p and upscaling. Fewer artifacts survive into the final frame, and the footage drops into a real edit without an intermediate enhancement step.
Synchronized audio
Generates a matched audio track alongside the video in the same pass instead of leaving sound design as a separate job. Puts H3 in the same category as Veo 3 and ahead of most open video models, which are silent.
Open weights
The model is released with downloadable weights, so teams can self-host, fine-tune on proprietary footage, and keep sensitive material off external APIs — a route that closed models like Veo 3, Kling and Runway simply don't offer.
How it compares
| Tool | Best for | Pricing | Score |
|---|---|---|---|
| MiniMax H3 | — | Free tier + token-based API pricing (check site for current rates) | 8.2/10 |
| Runway vs Runway → | Filmmakers, advertisers, and content creators who need cinematic AI video with realistic motion and fine-grained creative control. | Free tier + Standard $12/mo + Pro $28/mo + Max $76/mo + Enterprise contact sales | 9.3/10 |
| Veo 3 vs Veo 3 → | Filmmakers, content creators, and marketing teams who need production-quality cinematic video with synced audio, without a production budget. | Free via Gemini + Vertex AI pay-per-use | 9.1/10 |
| Seedance 2.0 vs Seedance 2.0 → | Video creators and developers who need short AI-generated clips with synchronized audio, lip-sync, and cinematic camera control. | Free daily credits (Dreamina) + paid ~$15-$70/mo; API from ~$0.08/s | 9/10 |
Compare head-to-head
Related reading
Mistral €3bn Round Meets OpenAI Wiki Disclosure
Mistral closed a €3bn round above €21bn valuation while OpenAI outlined next steps on agent misalignment reporting from its recent incident.
Mistral Closes €3B Series D at €21B Valuation
Mistral raised €3 billion in Europe's largest tech equity round, led by Samsung at a post-money valuation above €21 billion, to expand sovereign open-weight
Claude 13 Million Line Lean Code Audited
Anthropic September 4 post records Claude writing 13 million lines of Lean code for Fermat’s Last Theorem formalization over 11 days. Audit covers source
Ready to try MiniMax H3?
Head to the official site to start with MiniMax H3 — pricing and plans are listed above.
Visit MiniMax H3

