Modellix Logo Aurora Nasdaq Full Gray Font SVG
模型
价格CLI
文档
博客
关于我们
简体中文
立即开始
模型/搜索
分类

探索所有模型

(结果:68)
alibaba/qwen-image-3.0-edit
alibaba/qwen-image-3.0-edit
image-to-image

[Core Function] Qwen Image 3.0 Edit is the standard image-to-image editing model for instruction-based edits and multi-image fusion. [Strengths] It accepts 1-3 reference images plus an edit instruction, optional negative prompts, free-form output size (width*height), intelligent prompt rewrite (direct mode), and 1-6 outputs while preserving subject identity. [Best For] Background replacement, outfit or style changes, multi-image fusion, and iterative retouching when Pro-tier quality is not required. [Limitations] Do NOT use this for pure text-to-image with no reference images (use Qwen Image 3.0 instead). Keep output total pixels within 512*512 to 2048*2048. prompt_extend_mode only supports direct (agent is T2I-only). [Routing] Route here when the user provides reference image(s) and wants balanced Qwen 3.0 edit quality. Prefer Qwen Image 3.0 Pro Edit for higher quality edits.

$0.0040/img
alibaba/qwen-image-3.0
alibaba/qwen-image-3.0
text-to-image

[Core Function] Qwen Image 3.0 is Alibaba's standard text-to-image model balancing quality and speed. [Strengths] It supports free-form output size (width*height), optional negative prompts, intelligent prompt rewrite (direct/agent modes), batch generation of 1-6 images, and long structured prompts. [Best For] General creative stills, posters with readable text, product shots, and multi-variant exploration (n up to 6) when Pro-tier photorealism is not required. [Limitations] Do NOT use this if the user needs image editing with reference images (use Qwen Image 3.0 Edit). Keep total pixels within 512*512 to 2048*2048. Very long prompts combined with a long negative_prompt may exceed the model input capacity (about 4.5k tokens total). [Routing] Prefer Qwen Image 3.0 Pro for higher photorealism; use this for balanced quality/speed. If the user provides reference image(s) to edit, route to Qwen Image 3.0 Edit instead.

$0.0020/img
alibaba/qwen-image-3.0-pro-edit
alibaba/qwen-image-3.0-pro-edit
image-to-image

[Core Function] Qwen Image 3.0 Pro Edit is an image-to-image editing model for instruction-based edits and multi-image fusion. [Strengths] It accepts 1-3 reference images plus an edit instruction, optional negative prompts, free-form output size (width*height), intelligent prompt rewrite, 1-6 outputs while preserving subject identity, and long structured edit instructions. [Best For] Highly recommended for: background replacement, outfit or style changes, multi-image fusion, identity-preserving portrait edits, and iterative creative retouching. [Limitations] Do NOT use this for pure text-to-image with no reference images (use Qwen Image 3.0 Pro instead), native 4K output, or thinking-mode reasoning. Keep output total pixels within 512*512 to 2048*2048; input images should follow supported formats and size guidance. Do NOT combine very long prompts with multiple reference images and a long negative_prompt if the request may exceed the model input capacity (about 4.5k tokens total across text and images). [Routing] Route here when the user provides reference image(s) and wants Qwen 3.0 edit quality. For text-only generation without images, use Qwen Image 3.0 Pro.

$0.0460/img
alibaba/qwen-image-3.0-pro
alibaba/qwen-image-3.0-pro
text-to-image

[Core Function] Qwen Image 3.0 Pro is Alibaba's latest text-to-image model with strong prompt following and photorealism. [Strengths] It supports free-form output size (width*height), optional negative prompts, intelligent prompt rewrite, batch generation of 1-6 images, and long structured prompts for complex layouts. [Best For] Highly recommended for: photorealistic stills, marketing posters with readable text, detailed scene compositions, multi-panel layouts, product hero shots, and multi-variant creative exploration (n up to 6). [Limitations] Do NOT use this if the user needs native 4K output, thinking-mode reasoning, or image editing with reference images (use Qwen Image 3.0 Pro Edit for edits). Keep total pixels within 512*512 to 2048*2048. Very long prompts combined with a long negative_prompt may exceed the model input capacity (about 4.5k tokens total). [Routing] Prefer this over Qwen Image 2.0 Pro for new Qwen Image text-to-image work. If the user provides reference image(s) to edit, route to Qwen Image 3.0 Pro Edit instead.

$0.0460/img
bytedance/seedream-5.0-pro-multi-reference
bytedance/seedream-5.0-pro-multi-reference
image-to-image

[Core Function] Seedream 5.0 Pro Multi-Reference is a professional-grade multi-reference image generation (I2I) model that creates a single image from 2-10 reference images plus a text prompt. [Strengths] It excels at reference consistency, preserving characters, styles, and objects across multiple input images while following complex blending instructions with professional-grade quality. [Best For] Highly recommended for: keeping character or style consistency across references, combining subjects from different images into one scene, placing products into reference scenes, and IP-consistent content creation. [Limitations] Do NOT use this model with fewer than 2 or more than 10 reference images, and do NOT use it for batch generation or streaming; it outputs exactly one image per request. [Routing] For single-image editing use Seedream 5.0 Pro Edit; for text-only generation use Seedream 5.0 Pro; choose Seedream 5.0 Lite Edit when up to 14 reference images or batch outputs are needed.

$0.1725/img
bytedance/seedream-5.0-pro-edit
bytedance/seedream-5.0-pro-edit
image-to-image

[Core Function] Seedream 5.0 Pro Edit is a professional-grade single-image editing (I2I) model. [Strengths] It supports interactive precise editing: edit locations can be specified via coordinates, selection boxes, or arrows described in the prompt, with strong element-level control and subject consistency. [Best For] Highly recommended for: precise local retouching, adding/removing/replacing objects at exact positions, style transfer of a single photo, and professional post-editing workflows. [Limitations] Do NOT use this model for text-to-image generation (an input image is required) or for blending multiple reference images; it accepts exactly one input image and outputs exactly one image (no batch or streaming). [Routing] For 2-10 reference images use Seedream 5.0 Pro Multi-Reference; for pure text-to-image use Seedream 5.0 Pro; choose Seedream 5.0 Lite Edit when batch outputs or more than 10 input images are required.

$0.1380/img
bytedance/seedream-5.0-pro
bytedance/seedream-5.0-pro
text-to-image

[Core Function] Seedream 5.0 Pro is ByteDance's flagship professional-grade Text-to-Image (T2I) generation model. [Strengths] It delivers top-tier image quality with enhanced precision control over positions and elements, superior prompt adherence, and improved generation consistency for professional scenarios. [Best For] Highly recommended for: professional design assets, high-fidelity photorealistic imagery, precisely controlled compositions, and brand or commercial visuals where quality matters most. [Limitations] Do NOT use this model for batch image generation or streaming output; it generates exactly one image per request and supports up to 2K resolution (no 3K/4K). [Routing] Choose this Pro model when the user emphasizes ultimate quality or precision. Choose Seedream 5.0 Lite when real-time web knowledge, batch generation, or 3K resolution is needed. To edit an existing image use Seedream 5.0 Pro Edit; to blend multiple reference images use Seedream 5.0 Pro Multi-Reference.

$0.1035/img
google/nano-banana-2-lite-edit
google/nano-banana-2-lite-edit
image-to-image

[Core Function] Nano Banana 2 Lite Edit (Gemini 3.1 Flash Lite Image) is the lightweight, cost-efficient image editing model of the Nano Banana 2 family; it transforms one or more input images per a text instruction. [Strengths] It performs fast, low-cost instruction-based editing across up to 14 input images. [Best For] Highly recommended for: high-volume edits, quick style transforms, and batch background or attribute changes where cost and throughput matter. [Limitations] Do NOT use this model when you need the highest edit fidelity or richest detail; use Nano Banana 2 Edit or Nano Banana Pro Edit instead. [Routing] Choose the Lite variant for cost- and throughput-sensitive edits; step up to Nano Banana 2 Edit for higher quality.

$0.0276/img
google/nano-banana-2-lite
google/nano-banana-2-lite
text-to-image

[Core Function] Nano Banana 2 Lite (Gemini 3.1 Flash Lite Image) is the lightweight, cost-efficient text-to-image model of the Nano Banana 2 family. [Strengths] It generates images even faster and more cheaply than Nano Banana 2, well suited to high-volume creative and stylized output at lower cost. [Best For] Highly recommended for: high-volume batch generation, quick drafts and thumbnails, and cost-sensitive creative iteration. [Limitations] Do NOT use this model when you need maximum detail, high-end photorealism, or the richest quality; use Nano Banana 2 or Nano Banana Pro instead. [Routing] Choose the Lite variant when cost and throughput matter more than peak quality; step up to Nano Banana 2 for richer results.

$0.0274/img
...
没有找到需要的模型? 告诉我们。
Modellix Logo Full Colored

一站式聚合全球领先 AI 媒体模型

产品
模型探索模型精选模型
资源
文档博客价格CLI
关于
关于我们服务协议隐私声明
系列
CosyVoiceFunGemini OmniGPT ImageGrok ImagineGrok VoiceHailuo 02Hailuo 2.3HappyHorseImagenKling V3MAI ImageMiniMax Speech
Nano BananaPixVerse C1PixVerse V6Qwen AudioQwen ImageSeedanceSeedreamSkyReelsVeo 3Veo 3.1Vidu Q3Wan 2.6Wan 2.7
供应商
AlibabaBytedanceGoogleKlingMicrosoftMiniMaxOpenAIPixVerseReveSkyworkViduxAI
集合
AI Animation GeneratorAI Anime GeneratorAI Art GeneratorAI Avatar GeneratorAI Portrait GeneratorAI Style TransferAI Video GeneratorColorize PhotoDigital HumanFace SwapImage UpscaleLip SyncLogo Generator
One ClickPhoto RestorationVirtual Try-OnVoice Cloning
分类
文生图图像编辑文生视频图生视频视频生视频文生语音语音转文本语音转语音

Copyright © 2026 METAVERSE CLOUD PTE. LTD. 保留所有权利。

地址: 60 Paya Lebar Road #12-03, Paya Lebar Square, Singapore 409051

邮箱: support@modellix.ai

  • English
  • 简体中文
  • 日本語
最新发布
暂无最新发布
图像 · 68 个模型
GPT Image 2NEWNano Banana 2HOTSeedream 5.0 LiteHOT
查看全部图像模型
视频 · 138 个模型
Wan2.7 T2VNEWVeo 3.1 Lite T2VNEWHailuo 2.3 T2VHOTSeedance 2.0 I2VHOTHappyhorse 1.0 T2VHOT
查看全部视频模型
精选
Seedance 2.5 T2VSeedream 5.0 ProGemini Omni Flash T2VV6 T2VHappyhorse 1.1 T2V
浏览全部 220+ 模型
快速开始
AI OnboardingQuick StartModel ProvidersPricing
使用方式
REST APICLIMCPAgent Skill
资源
Product UpdatesModel UpdatesSupportDiscord Community
查看完整文档
热门
GPT Image 2 API Guide: Pricing & VariantsSeedream API Guide: 4K ImagesKling AI API: Pricing & IntegrationSeedance 2.0 API: Global Access
最新
HappyHorse-1.0: Open-Source AI Video ModelGoogle AI Suite Lands on ModellixModellix CLI: Generate from TerminalIntroducing Modellix
浏览全部文章