Modellix Logo Aurora Nasdaq Full Gray Font SVG
モデル
料金CLI
ドキュメント
ブログ
Modellixについて
日本語
始める
モデル/検索
カテゴリー

すべてのモデルを探す

(34件)
alibaba/qwen-image-3.0-pro
alibaba/qwen-image-3.0-pro
text-to-image

[Core Function] Qwen Image 3.0 Pro is Alibaba's latest text-to-image model with strong prompt following and photorealism. [Strengths] It supports free-form output size (width*height), optional negative prompts, intelligent prompt rewrite, batch generation of 1-6 images, and long structured prompts for complex layouts. [Best For] Highly recommended for: photorealistic stills, marketing posters with readable text, detailed scene compositions, multi-panel layouts, product hero shots, and multi-variant creative exploration (n up to 6). [Limitations] Do NOT use this if the user needs native 4K output, thinking-mode reasoning, or image editing with reference images (use Qwen Image 3.0 Pro Edit for edits). Keep total pixels within 512*512 to 2048*2048. Very long prompts combined with a long negative_prompt may exceed the model input capacity (about 4.5k tokens total). [Routing] Prefer this over Qwen Image 2.0 Pro for new Qwen Image text-to-image work. If the user provides reference image(s) to edit, route to Qwen Image 3.0 Pro Edit instead.

$0.0460/img
bytedance/seedream-5.0-pro
bytedance/seedream-5.0-pro
text-to-image

[Core Function] Seedream 5.0 Pro is ByteDance's flagship professional-grade Text-to-Image (T2I) generation model. [Strengths] It delivers top-tier image quality with enhanced precision control over positions and elements, superior prompt adherence, and improved generation consistency for professional scenarios. [Best For] Highly recommended for: professional design assets, high-fidelity photorealistic imagery, precisely controlled compositions, and brand or commercial visuals where quality matters most. [Limitations] Do NOT use this model for batch image generation or streaming output; it generates exactly one image per request and supports up to 2K resolution (no 3K/4K). [Routing] Choose this Pro model when the user emphasizes ultimate quality or precision. Choose Seedream 5.0 Lite when real-time web knowledge, batch generation, or 3K resolution is needed. To edit an existing image use Seedream 5.0 Pro Edit; to blend multiple reference images use Seedream 5.0 Pro Multi-Reference.

$0.1035/img
google/nano-banana-2-lite
google/nano-banana-2-lite
text-to-image

[Core Function] Nano Banana 2 Lite (Gemini 3.1 Flash Lite Image) is the lightweight, cost-efficient text-to-image model of the Nano Banana 2 family. [Strengths] It generates images even faster and more cheaply than Nano Banana 2, well suited to high-volume creative and stylized output at lower cost. [Best For] Highly recommended for: high-volume batch generation, quick drafts and thumbnails, and cost-sensitive creative iteration. [Limitations] Do NOT use this model when you need maximum detail, high-end photorealism, or the richest quality; use Nano Banana 2, Nano Banana Pro, or the Imagen 4 series instead. [Routing] Choose the Lite variant when cost and throughput matter more than peak quality; step up to Nano Banana 2 for richer results.

$0.0274/img
xai/grok-imagine-image
xai/grok-imagine-image
text-to-image

[Core Function] Grok Imagine Image is xAI's standard text-to-image generation model. [Strengths] It excels at quickly generating solid, visually appealing images from a text prompt across a wide range of aspect ratios. [Best For] Highly recommended for: rapid prototyping, social media content, and general-purpose image generation. [Limitations] Do NOT use this model when maximum detail or fidelity is required; the Quality variant produces richer detail. [Routing] Choose this model for fast, general image generation. When the user demands maximum fidelity, route to Grok Imagine Image (Quality).

$0.0230/img
xai/grok-imagine-image-quality
xai/grok-imagine-image-quality
text-to-image

[Core Function] Grok Imagine Image (Quality) is xAI's high-fidelity text-to-image generation model. [Strengths] It excels at producing richly detailed, high-quality images from a text prompt, with flexible aspect ratios and an optional 2K resolution. [Best For] Highly recommended for: detailed concept art, marketing visuals, and any scenario where image quality is prioritized over generation speed. [Limitations] Do NOT use this model when latency is critical, as generation is slower than the standard model. [Routing] Use this model by default when the user emphasizes quality or detail. For faster, lighter generation use Grok Imagine Image (standard).

$0.0575~$0.0805/img
microsoft/mai-image-2.5-flash
microsoft/mai-image-2.5-flash
text-to-image

[Core Function] MAI Image 2.5 Flash is Microsoft's fast, cost-efficient text-to-image generation model. [Strengths] It excels at quickly generating solid images from a text prompt with the same dimension controls as MAI Image 2.5. [Best For] Highly recommended for: rapid prototyping, batch generation, and cost-sensitive workloads. [Limitations] Do NOT request dimensions below 768 on any side or a width x height product above 1,048,576 pixels; output is always PNG, and maximum fidelity is lower than MAI Image 2.5. [Routing] Choose this model when speed or cost matters more than maximum fidelity. For the highest quality, use MAI Image 2.5.

$0.0388/img
microsoft/mai-image-2.5
microsoft/mai-image-2.5
text-to-image

[Core Function] MAI Image 2.5 is Microsoft's flagship text-to-image generation model. [Strengths] It excels at producing high-quality, detailed images from a text prompt with precise control over output dimensions. [Best For] Highly recommended for: concept art, marketing visuals, and high-fidelity image generation. [Limitations] Do NOT request dimensions below 768 on any side or a width x height product above 1,048,576 pixels (e.g. beyond 1024x1024); output is always PNG. [Routing] Use this model by default for quality-sensitive generation. For faster, cheaper generation, route to MAI Image 2.5 Flash.

$0.0553/img
openai/gpt-image-1.5
openai/gpt-image-1.5
text-to-image

[Core Function] GPT Image 1.5 is a versatile text-to-image generation model. [Strengths] It balances solid visual performance with crucial utility features, notably its native support for generating images with transparent backgrounds. [Best For] Highly recommended for: creating UI icons, standalone logos, game assets, and any graphic design elements that require a transparent background. [Limitations] Do NOT use this model if you need 2K or 4K resolution. Its maximum supported resolution is 1536x1024. [Routing] Choose this model specifically when the user asks for 'transparent background', 'no background', or 'PNG icon'. For standard, high-fidelity, or 4K image generation, use GPT Image 2 instead.

$0.0138~$0.2334/img
openai/gpt-image-2
openai/gpt-image-2
text-to-image

[Core Function] GPT Image 2 is a state-of-the-art text-to-image generation model. [Strengths] It excels at generating highly detailed, photorealistic images from text descriptions, with native support for ultra-high resolutions including 2K and 4K (up to 3840x2160). [Best For] Highly recommended for: cinematic landscapes, detailed character portraits, high-end commercial concept art, and any scenario requiring maximum resolution and visual fidelity. [Limitations] Do NOT use this model if you need a transparent background (e.g., for icons or UI assets), as it does not support the `background: transparent` parameter. [Routing] Use this model by default for all high-quality image generation requests. If the user explicitly asks for an image with a transparent background, route to GPT Image 1.5 instead.

$0.0041~$0.3943/img
必要なモデルが見つかりませんか? ご要望をお聞かせください。
Modellix Logo Full Colored

あらゆるAIメディア生成を、1つの統合APIで。

プロダクト
モデルモデルを探す注目モデル
リソース
ドキュメントブログ料金CLI
会社情報
Modellixについてサービス利用規約プライバシーポリシー
シリーズ
CosyVoiceFunGemini OmniGPT ImageGrok ImagineGrok VoiceHailuo 02Hailuo 2.3HappyHorseImagenKling V3MAI ImageMiniMax Speech
Nano BananaPixVerse C1PixVerse V6Qwen AudioQwen ImageSeedanceSeedreamSkyReelsVeo 3Veo 3.1Vidu Q3Wan 2.6Wan 2.7
プロバイダー
AlibabaBytedanceGoogleKlingMicrosoftMiniMaxOpenAIPixVerseReveSkyworkViduxAI
コレクション
AI Animation GeneratorAI Anime GeneratorAI Art GeneratorAI Avatar GeneratorAI Portrait GeneratorAI Style TransferAI Video GeneratorColorize PhotoDigital HumanFace SwapImage UpscaleLip SyncLogo Generator
One ClickPhoto RestorationVirtual Try-OnVoice Cloning
カテゴリー
テキストから画像画像編集テキストから動画画像から動画動画から動画テキスト読み上げ音声認識音声変換

Copyright © 2026 METAVERSE CLOUD PTE. LTD. All rights reserved.

所在地: 60 Paya Lebar Road #12-03, Paya Lebar Square, Singapore 409051

メール: support@modellix.ai

  • English
  • 简体中文
  • 日本語
最新リリース
最新リリースはまだありません
画像 · 69モデル
GPT Image 2NEWNano Banana 2HOTSeedream 5.0 LiteHOT
画像モデルをすべて表示
動画 · 137モデル
Wan2.7 T2VNEWVeo 3.1 Lite T2VNEWHailuo 2.3 T2VHOTSeedance 2.0 I2VHOTHappyhorse 1.0 T2VHOT
動画モデルをすべて表示
注目
Seedream 5.0 ProGemini Omni Flash T2VSeedance 2.0 T2VHappyhorse 1.1 T2VGPT Image 2
220+モデルをすべて見る
はじめに
AI OnboardingQuick StartModel ProvidersPricing
利用方法
REST APICLIMCPAgent Skill
リソース
Product UpdatesModel UpdatesSupportDiscord Community
ドキュメントをすべて見る
トレンド
GPT Image 2 API Guide: Pricing & VariantsSeedream API Guide: 4K ImagesKling AI API: Pricing & IntegrationSeedance 2.0 API: Global Access
最新記事
HappyHorse-1.0: Open-Source AI Video ModelGoogle AI Suite Lands on ModellixModellix CLI: Generate from TerminalIntroducing Modellix
すべての記事を見る