Modellix Logo Aurora Nasdaq Full Gray Font SVG
料金
始める
モデル/検索
カテゴリー

すべてのモデルを探す

(49件)
microsoft/mai-image-2.6-flash-edit
microsoft/mai-image-2.6-flash-edit
image-to-image

[Core Function] MAI Image 2.6 Flash Edit is the faster, lower-cost variant of MAI Image 2.6 image editing. [Strengths] It applies prompt-guided edits to a single source image with lower latency than MAI Image 2.6 Edit. [Best For] Highly recommended for: high-throughput edit APIs and production retouching pipelines. [Limitations] Do NOT use this if the caller provides more than one source image, a data URI, or a URL that is not publicly reachable over HTTP or HTTPS. It accepts exactly one JPEG or PNG image URL, and output is always PNG. auto_aspect_ratio and web_grounding are optional booleans; omit them unless the caller sets them. [Routing] Use this model when edit speed or cost matters more than maximum 2.6 quality. For the highest fidelity, route to MAI Image 2.6 Edit.

$0.0220/img
microsoft/mai-image-2.6-flash
microsoft/mai-image-2.6-flash
text-to-image

[Core Function] MAI Image 2.6 Flash is the faster, lower-cost variant of MAI Image 2.6 text-to-image generation. [Strengths] It targets similar quality to MAI Image 2.6 with lower latency for high-throughput workloads. [Best For] Highly recommended for: production pipelines, batch generation, and latency-sensitive image APIs. [Limitations] Do NOT use this if the requested width or height is below 768, or if width x height exceeds 1,048,576 pixels. Output is always PNG. auto_aspect_ratio and web_grounding are optional booleans; omit them unless the caller sets them. [Routing] Use this model when speed or cost matters more than maximum 2.6 quality. For the highest fidelity, route to MAI Image 2.6.

$0.0195/img
microsoft/mai-image-2.6-edit
microsoft/mai-image-2.6-edit
image-to-image

[Core Function] MAI Image 2.6 Edit is Microsoft's latest prompt-guided image editing model in the MAI Image family. [Strengths] It applies targeted edits to a single source image with the same quality gains as MAI Image 2.6 generation. [Best For] Highly recommended for: object edits, layout changes, text cleanup, and iterative photorealistic retouching. [Limitations] Do NOT use this if the caller provides more than one source image, a data URI, or a URL that is not publicly reachable over HTTP or HTTPS. It accepts exactly one JPEG or PNG image URL, and output is always PNG. auto_aspect_ratio and web_grounding are optional booleans; omit them unless the caller sets them. [Routing] Use this model when the latest MAI edit quality is required. For faster or lower-cost edits, route to MAI Image 2.6 Flash Edit.

$0.0471/img
microsoft/mai-image-2.6
microsoft/mai-image-2.6
text-to-image

[Core Function] MAI Image 2.6 is Microsoft's latest text-to-image generation model in the MAI Image family. [Strengths] It improves text rendering, portraits, 3D imagery, and commercial photorealistic output compared with MAI Image 2.5. [Best For] Highly recommended for: marketing visuals, product hero shots, portraits, and prompts that need accurate on-image text. [Limitations] Do NOT use this if the requested width or height is below 768, or if width x height exceeds 1,048,576 pixels. Output is always PNG. auto_aspect_ratio and web_grounding are optional booleans; omit them unless the caller sets them. [Routing] Use this model when the latest MAI quality is required. For lower latency or cost, route to MAI Image 2.6 Flash.

$0.0389/img
xai/grok-imagine-image-2.0-edit
xai/grok-imagine-image-2.0-edit
image-to-image

[Core Function] Grok Imagine Image 2.0 Edit edits one to three source images from a text prompt. [Strengths] It supports quality (low or medium), 1k or 2k resolution, and the same aspect ratios as Grok Imagine Image 2.0. [Best For] Highly recommended for: restyling, combining up to 3 references, and iterative refinement. [Limitations] Requires at least one source image; at most 3 images per request. This model does not generate from text alone. [Routing] For text-to-image generation, use Grok Imagine Image 2.0.

$0.0840~$0.1320/img
xai/grok-imagine-image-2.0
xai/grok-imagine-image-2.0
text-to-image

[Core Function] Grok Imagine Image 2.0 generates images from a text prompt. [Strengths] It supports quality (low or medium), 1k or 2k resolution, a wide range of aspect ratios including 21:9 and 5:2, and up to 10 images per request. [Best For] Highly recommended for: concept art, marketing visuals, cinematic banners, and batch generation. [Limitations] This model does not edit existing images. [Routing] For image editing, use Grok Imagine Image 2.0 Edit.

$0.0480~$0.0960/img
microsoft/mai-image-2.5-pro-edit
microsoft/mai-image-2.5-pro-edit
image-to-image

[Core Function] MAI Image 2.5 Pro Edit is Microsoft's highest-fidelity image editing model in the MAI 2.5 family. [Strengths] It excels at applying premium, prompt-guided edits and transformations to a single source image. [Best For] Highly recommended for: hero asset retouching, high-fidelity restyling, and detailed prompt-driven edits. [Limitations] Do NOT provide more than one source image or non-JPEG/PNG inputs; it accepts exactly one JPEG or PNG image, and output is always PNG. [Routing] Use this model when maximum edit quality is required. For faster or lower-cost edits, route to MAI Image 2.5 Edit or MAI Image 2.5 Flash Edit.

$0.1167/img
microsoft/mai-image-2.5-pro
microsoft/mai-image-2.5-pro
text-to-image

[Core Function] MAI Image 2.5 Pro is Microsoft's highest-fidelity text-to-image generation model in the MAI 2.5 family. [Strengths] It excels at producing exceptionally detailed, high-quality images from a text prompt with the same dimension controls as MAI Image 2.5. [Best For] Highly recommended for: premium marketing visuals, hero assets, and maximum-fidelity generation. [Limitations] Do NOT request dimensions below 768 on any side or a width x height product above 1,048,576 pixels; output is always PNG. [Routing] Use this model when maximum quality is required. For faster or lower-cost generation, route to MAI Image 2.5 or MAI Image 2.5 Flash.

$0.1085/img
alibaba/qwen-image-3.0-edit
alibaba/qwen-image-3.0-edit
image-to-image

[Core Function] Qwen Image 3.0 Edit is the standard image-to-image editing model for instruction-based edits and multi-image fusion. [Strengths] It accepts 1-3 reference images plus an edit instruction, optional negative prompts, free-form output size (width*height), intelligent prompt rewrite (direct mode), and 1-6 outputs while preserving subject identity. [Best For] Background replacement, outfit or style changes, multi-image fusion, and iterative retouching when Pro-tier quality is not required. [Limitations] Do NOT use this for pure text-to-image with no reference images (use Qwen Image 3.0 instead). Keep output total pixels within 512*512 to 2048*2048. prompt_extend_mode only supports direct (agent is T2I-only). [Routing] Route here when the user provides reference image(s) and wants balanced Qwen 3.0 edit quality. Prefer Qwen Image 3.0 Pro Edit for higher quality edits.

$0.0297/img
必要なモデルが見つかりませんか? ご要望をお聞かせください。
Modellix Logo Full Colored

あらゆるAIメディア生成を、1つの統合APIで。

モデル
すべて探す注目LLM
リソース
料金CLIドキュメントブログGitHub
会社情報
営業窓口Modellixについてサービス利用規約プライバシーポリシー
シリーズ
CosyVoiceFunGemini OmniGPT ImageGrok ImagineGrok VoiceHailuo 02Hailuo 2.3HappyHorseImagenKling V3MAI ImageMiniMax Speech
Nano BananaPixVerse C1PixVerse V6Qwen AudioQwen ImageSeedanceSeedreamSkyReelsVeo 3Veo 3.1Vidu Q3Wan
プロバイダー
AlibabaBytedanceGoogleKlingMicrosoftMiniMaxOpenAIPixVerseReveSkyworkViduxAI
コレクション
AI Animation GeneratorAI Anime GeneratorAI Art GeneratorAI Avatar GeneratorAI Portrait GeneratorAI Style TransferAI Video GeneratorColorize PhotoDigital HumanFace SwapImage UpscaleLip SyncLogo Generator
One ClickPhoto RestorationVirtual Try-OnVoice Cloning
カテゴリー
テキストから画像画像編集テキストから動画画像から動画動画から動画テキスト読み上げ音声認識音声変換

Copyright © 2026 METAVERSE CLOUD PTE. LTD. All rights reserved.

所在地: 60 Paya Lebar Road #12-03, Paya Lebar Square, Singapore 409051

メール: support@modellix.ai

  • English
  • 简体中文
  • 日本語
最新リリース
最新リリースはまだありません
モデル
すべて探す注目LLM
リソース
CLIドキュメントブログGitHub
会社情報
営業窓口Modellixについて