minimax/minimax-h3-t2v

Docs
Schema

[Core Function] MiniMax H3 T2V is a text-to-video generation model that creates video from a text prompt only. [Strengths] It supports 4-15 second clips, 768P or 2K output, and concrete aspect ratios from cinematic ultrawide to vertical. [Best For] Highly recommended for: prompt-only storyboards, character-driven shorts, cinematic B-roll from text, and high-resolution drafts without image inputs. [Limitations] Do NOT use this model if you need to condition on images, first or last frames, or reference videos. prompt, duration, resolution, and ratio are required; ratio must be one of the documented aspect ratios. [Routing] Choose MiniMax H3 T2V for text-only MiniMax H3 video. If the user provides a start or end frame, use MiniMax H3 FL2V. If they provide reference images, use MiniMax H3 I2V. If they provide reference videos, use MiniMax H3 V2V.

$0.0920~$0.1495/sec
text-to-video

Input

Video content description.
Required video duration in seconds.
Required video aspect ratio. One of the listed values.
Required output resolution.

Result

No results yet

Run the model to preview the output here.

Next:

README

Attributes

Parameters

Name Description Type Required Enums
prompt Video content description. string Yes -
duration Required video duration in seconds. integer Yes 4, 5, 6, 7, 8, 9, 10, 11, 12, 13, 14, 15
ratio Required video aspect ratio. One of the listed values. string Yes 21:9, 16:9, 4:3, 1:1, 3:4, 9:16
resolution Required output resolution. string Yes 768P, 2K

Pricing

Unit: $/sec

Dimension Pricing
resolution: 768P 0.0920
resolution: 2K 0.1495