系列

Hailuo 2.3 AI 模型家族

3 个模型更新于 Apr 2026
Hailuo 2.3 AI 模型家族

关于 Hailuo 2.3 模型

MiniMax's Hailuo 2.3 series elevates cinematic AI video gen with 4K T2V/I2V, hyper-realistic physics/motion, extended clips, and advanced character consistency.

全部 Hailuo 2.3 模型

minimax/hailuo-2.3-fast-i2v

minimax/hailuo-2.3-fast-i2v

image-to-video

[Core Function] Hailuo 2.3 Fast I2V is a high-speed, cost-effective image-to-video generation model. [Strengths] It excels at generating videos from images much faster and at roughly 50% lower cost than the standard 2.3 model, while still maintaining the 2.3 architecture's strength in human motion. [Best For] Highly recommended for: rapid prototyping, batch social media creation, and cost-sensitive video generation pipelines. [Limitations] Do NOT use this model for text-to-video (it only accepts image inputs). Do NOT use when absolute maximum visual fidelity is the primary requirement. [Routing] Choose this model when the user emphasizes 'fast', 'quick', or 'cost-effective' image-to-video generation. For maximum quality, use the standard Hailuo 2.3 I2V.

minimax/hailuo-2.3-i2v

minimax/hailuo-2.3-i2v

image-to-video

[Core Function] Hailuo 2.3 I2V is a flagship image-to-video generation model optimized for character animation. [Strengths] It excels at animating human characters from a single image, maintaining consistent facial features, producing natural micro-expressions, and handling stylized artwork seamlessly. [Best For] Highly recommended for: animating character concept art, bringing portraits to life, and creating stylized/anime motion sequences. [Limitations] Do NOT use this model for last-frame conditioning (it does not support FL2V). Do NOT use if you need 1080p resolution for 10 seconds (1080p is capped at 6s). [Routing] Use this model by default for high-quality image-to-video tasks involving people or art. For physical realism or 10s at 1080p, route to Hailuo 02 I2V. For cost-effective/faster generation, route to Hailuo 2.3 Fast I2V.

minimax/hailuo-2.3-t2v

minimax/hailuo-2.3-t2v

text-to-video

[Core Function] Hailuo 2.3 T2V is a flagship text-to-video generation model optimized for human performance and stylization. [Strengths] It excels at capturing intricate human motion, nuanced facial micro-expressions, prompt adherence, and applying highly stylized aesthetics (e.g., anime, ink wash, game CG) to video. [Best For] Highly recommended for: character-driven storytelling, close-up emotional shots, stylized artistic videos, and dialogue scenes. [Limitations] Do NOT use this model if you need native 1080p resolution for 10 full seconds (1080p is capped at 6 seconds; generating 10s forces 768p resolution). [Routing] Use this model by default for text-to-video requests involving humans, faces, or specific art styles. If the user requires strict physical realism/world dynamics or native 1080p for 10 seconds, route to Hailuo 02 T2V instead.

没有找到需要的模型? 告诉我们。

探索更多

图生视频
分类56 个模型
文生视频
分类25 个模型
视频生视频
分类27 个模型
Wan
Wan
系列11 个模型
Gemini Omni
Gemini Omni
系列4 个模型
AI Animation Generator
AI Animation Generator
集合16 个模型