Evaluate capability and live price
All three workflows share one model ID. Pricing comes from the current HiAPI billing configuration rather than a hard-coded content claim.
View live MiniMax H3 usage pricingMiniMax H3 combines native 2K text-to-video, first and last frame control, and multimodal reference video under one model ID. Compare the workflows, check live pricing, and validate a result in Playground before integrating it.
Provider: MiniMax
Category: video generation
Endpoint: /v1/tasks
Status: Available
Cost: See live page pricing
MiniMax H3 is a native 2K multimodal video model for text-to-video, first/last-frame control, and reference-driven generation with images, video, and audio across 4 to 15 seconds.
Pricing
Standard Usage
Use images to control the first and last frame.
Upload the first-frame image that controls how the video begins
Upload the last-frame image that controls how the video ends
Choose 4-15 seconds in 1-second steps
Upload reference audio files up to the model limit
Upload reference videos up to the model limit
Whether to add an AI-generated watermark
Ready for generation
Configure your parameters and click "Run" to see the output here.
CAPABILITY & ADOPTION
MiniMax H3 combines native 2K text-to-video, first and last frame control, and multimodal reference video under one model ID. Compare the workflows, check live pricing, and validate a result in Playground before integrating it.
All three workflows share one model ID. Pricing comes from the current HiAPI billing configuration rather than a hard-coded content claim.
View live MiniMax H3 usage pricingThe API reference owns parameters, field combinations, media limits, callbacks, and copyable request examples.
Read the MiniMax H3 API parameter and schema referenceThe tutorial focuses on workflow choice, task states, error diagnosis, idempotency, and output storage without duplicating the schema.
Follow the MiniMax H3 integration and debugging tutorialIt is prompt-led and can also use endpoint frames or image, video, and audio references according to the active contract.
4-15 seconds, 2K, and 21:9, 16:9, 4:3, 1:1, 3:4, 9:16, adaptive.
Optional image, video, and audio references; generated audio
Billing follows output duration and the live tier. Resolution, mode, or audio differences are reflected by the page price and task record.
POST an asynchronous task to /v1/tasks with model set to minimax-h3 and place parameters inside input.