alibaba/wan2.1-vace-plus

ドキュメント
スキーマ

[Core Function] Wan 2.1 VACE Plus - Unified Video Editing Model is an older generation image-to-image editing model. [Strengths] Delivered strong Chinese-language prompt understanding and regional aesthetic preferences. Maintained for backward compatibility. [Best For] Existing legacy integrations and workflows that strictly depend on this specific model version's quirks. [Limitations] Do NOT use this for new creations. It is a legacy model maintained for backward compatibility. [Routing] Only use if explicitly requested; otherwise use Wan 2.7 I2I.

$0.0591/sec
video-to-video

入力

Content description in Chinese or English (1-800 characters). Describes the desired video content or editing effect. Required for all functions
Video editing function to use. Determines which parameters are required and applicable: (1) image_reference: Generate video from reference images; (2) video_repainting: Apply style transfer to existing video using control conditions; (3) video_edit: Edit specific regions of video using mask; (4) video_extension: Extend video at start/end using frame/clip references; (5) video_outpainting: Expand video canvas boundaries
Bottom boundary expansion scale for video_outpainting function (1.0-2.0). 1.0 = no expansion (original height), 2.0 = double the bottom area. Default is 1.0. Use values > 1.0 to expand the canvas downward. For example, 1.5 means add 50% more canvas space below the original video. The model will generate content to fill the expanded area naturally
Control condition type for structure preservation. Usage by function: (1) video_repainting: required, determines how video structure is preserved during repainting; (2) video_extension: required when video_url is present, maintains consistency with reference video; (3) video_edit: optional, provides structural guidance for editing; (4) other functions: not used. Options: 'posebodyface' (pose+body+face detection), 'posebody' (pose+body only), 'depth' (depth map), 'scribble' (edge detection), '' (empty string for no extraction)
Video duration in seconds. Fixed at 5 seconds and cannot be modified. The model always generates 5-second videos regardless of this parameter value
Mask expansion mode for video_edit function. Determines how the mask is expanded when expand_ratio > 0. Options: (1) 'hull': convex hull expansion (smooth, rounded boundaries); (2) 'bbox': bounding box expansion (rectangular boundaries); (3) 'original': keep original mask shape while expanding. Default is 'hull'. Use 'hull' for natural objects, 'bbox' for rectangular regions
Mask expansion ratio for video_edit function (0.0-1.0). Expands the mask boundary to include surrounding areas. 0.0 = no expansion (use exact mask), 1.0 = maximum expansion. Default is 0.05 (5% expansion). Useful for ensuring complete coverage of editing region and avoiding edge artifacts
First video clip URL for video_extension function. Provides a video segment to use as the starting portion. At least one of first_frame_url/last_frame_url/first_clip_url/last_clip_url must be provided. Use this when you want more control than a single frame can provide. The model will extend naturally from the end of this clip
ヒント:ファイルをドラッグ&ドロップするか、クリップボード(Ctrl/Cmd+V)またはURLから追加できます。
First frame image URL for video_extension function. Specifies the starting frame when extending video forward. At least one of first_frame_url/last_frame_url/first_clip_url/last_clip_url must be provided. Use this to define the exact starting point of the extended video. The model will generate smooth transition from this frame
ヒント:ファイルをドラッグ&ドロップするか、クリップボード(Ctrl/Cmd+V)またはURLから追加できます。
Last video clip URL for video_extension function. Provides a video segment to use as the ending portion. At least one of first_frame_url/last_frame_url/first_clip_url/last_clip_url must be provided. Use this when you want more control than a single frame can provide. The model will extend naturally to the beginning of this clip
ヒント:ファイルをドラッグ&ドロップするか、クリップボード(Ctrl/Cmd+V)またはURLから追加できます。
Last frame image URL for video_extension function. Specifies the ending frame when extending video backward. At least one of first_frame_url/last_frame_url/first_clip_url/last_clip_url must be provided. Use this to define the exact ending point of the extended video. The model will generate smooth transition to this frame
ヒント:ファイルをドラッグ&ドロップするか、クリップボード(Ctrl/Cmd+V)またはURLから追加できます。
Left boundary expansion scale for video_outpainting function (1.0-2.0). 1.0 = no expansion (original width), 2.0 = double the left area. Default is 1.0. Use values > 1.0 to expand the canvas leftward. For example, 1.5 means add 50% more canvas space to the left of the original video. The model will generate content to fill the expanded area naturally
Frame ID (1-based index) indicating which frame the mask_image_url corresponds to. Only applicable when using mask_image_url in video_edit function. The mask will be propagated from this frame to others based on mask_type. For example, mask_frame_id=1 means the mask corresponds to the first frame of the video. Default is 1 (first frame)
Mask image URL for video_edit function. Defines the region to edit (white=edit, black=keep). Must provide either mask_image_url OR mask_video_url, not both. When using mask_image_url, also specify mask_frame_id to indicate which frame the mask corresponds to. The mask will be propagated to other frames based on mask_type setting
ヒント:ファイルをドラッグ&ドロップするか、クリップボード(Ctrl/Cmd+V)またはURLから追加できます。
Mask propagation type for video_edit function. Options: (1) 'tracking': mask follows object movement across frames (recommended for moving objects); (2) 'fixed': mask stays in same position across all frames (recommended for static scenes or background edits). Default is 'tracking'
Mask video URL for video_edit function. Provides frame-by-frame mask for precise control (white=edit, black=keep). Must provide either mask_image_url OR mask_video_url, not both. Mask video must have the same frame count as the input video. Use this for complex editing that requires different masks for different frames
ヒント:ファイルをドラッグ&ドロップするか、クリップボード(Ctrl/Cmd+V)またはURLから追加できます。
Classification for each reference image: 'obj' (object/subject) or 'bg' (background). Only for image_reference function. Array length must match ref_images_url length. Maximum 1 'bg' element allowed. Helps model distinguish between subject references and background style references
obj_or_bg
Enable intelligent prompt rewriting and enhancement. When true (default), the model will automatically optimize and expand your prompt for better results. When false, uses your prompt exactly as provided. Recommended to keep true unless you need precise control over the exact wording. Applicable to all functions
Reference image URLs array. Usage by function: (1) image_reference: 1-3 images required, used as visual references for video generation; (2) video_repainting: max 1 image optional, provides style reference; (3) video_edit: max 1 image optional, provides style reference; (4) other functions: not used. Images must be publicly accessible HTTP/HTTPS URLs
ヒント:ファイルをドラッグ&ドロップするか、クリップボード(Ctrl/Cmd+V)またはURLから追加できます。
Right boundary expansion scale for video_outpainting function (1.0-2.0). 1.0 = no expansion (original width), 2.0 = double the right area. Default is 1.0. Use values > 1.0 to expand the canvas rightward. For example, 1.5 means add 50% more canvas space to the right of the original video. The model will generate content to fill the expanded area naturally
Random seed for reproducible results (0-2147483647). Using the same seed with identical parameters will produce similar (though not pixel-perfect identical) results. Useful for A/B testing different prompts while keeping other randomness constant. If not specified, a random seed is used each time. Applicable to all functions
Output video resolution in width*height format. Available options: '1280*720' (16:9 landscape, HD), '720*1280' (9:16 portrait, mobile-friendly), '960*960' (1:1 square, social media), '1088*832' (4:3 landscape), '832*1088' (3:4 portrait). Default is '1280*720'. Choose based on your target platform and use case. Applicable to all functions
Top boundary expansion scale for video_outpainting function (1.0-2.0). 1.0 = no expansion (original height), 2.0 = double the top area. Default is 1.0. Use values > 1.0 to expand the canvas upward. For example, 1.5 means add 50% more canvas space above the original video. The model will generate content to fill the expanded area naturally
Input video URL. Usage by function: (1) video_repainting: required, the video to be repainted; (2) video_edit: required, the video to be edited; (3) video_outpainting: required, the video to expand; (4) video_extension: optional, provides reference style when extending; (5) image_reference: not used. Must be publicly accessible HTTP/HTTPS/OSS URL
ヒント:ファイルをドラッグ&ドロップするか、クリップボード(Ctrl/Cmd+V)またはURLから追加できます。

結果

結果はまだありません

モデルを実行すると、ここで出力をプレビューできます。

次へ:

README

属性

パラメーター

名前 説明 必須 列挙値
prompt Content description in Chinese or English (1-800 characters). Describes the desired video content or editing effect. Required for all functions string はい -
function Video editing function to use. Determines which parameters are required and applicable: (1) image_reference: Generate video from reference images; (2) video_repainting: Apply style transfer to existing video using control conditions; (3) video_edit: Edit specific regions of video using mask; (4) video_extension: Extend video at start/end using frame/clip references; (5) video_outpainting: Expand video canvas boundaries string はい image_reference, video_repainting, video_edit, video_extension, video_outpainting
bottom_scale Bottom boundary expansion scale for video_outpainting function (1.0-2.0). 1.0 = no expansion (original height), 2.0 = double the bottom area. Default is 1.0. Use values > 1.0 to expand the canvas downward. For example, 1.5 means add 50% more canvas space below the original video. The model will generate content to fill the expanded area naturally number いいえ -
control_condition Control condition type for structure preservation. Usage by function: (1) video_repainting: required, determines how video structure is preserved during repainting; (2) video_extension: required when video_url is present, maintains consistency with reference video; (3) video_edit: optional, provides structural guidance for editing; (4) other functions: not used. Options: ‘posebodyface’ (pose+body+face detection), ‘posebody’ (pose+body only), ‘depth’ (depth map), ‘scribble’ (edge detection), ‘’ (empty string for no extraction) string いいえ posebodyface, posebody, depth, scribble, ``
duration Video duration in seconds. Fixed at 5 seconds and cannot be modified. The model always generates 5-second videos regardless of this parameter value integer いいえ 5
expand_mode Mask expansion mode for video_edit function. Determines how the mask is expanded when expand_ratio > 0. Options: (1) ‘hull’: convex hull expansion (smooth, rounded boundaries); (2) ‘bbox’: bounding box expansion (rectangular boundaries); (3) ‘original’: keep original mask shape while expanding. Default is ‘hull’. Use ‘hull’ for natural objects, ‘bbox’ for rectangular regions string いいえ hull, bbox, original
expand_ratio Mask expansion ratio for video_edit function (0.0-1.0). Expands the mask boundary to include surrounding areas. 0.0 = no expansion (use exact mask), 1.0 = maximum expansion. Default is 0.05 (5% expansion). Useful for ensuring complete coverage of editing region and avoiding edge artifacts number いいえ -
first_clip_url First video clip URL for video_extension function. Provides a video segment to use as the starting portion. At least one of first_frame_url/last_frame_url/first_clip_url/last_clip_url must be provided. Use this when you want more control than a single frame can provide. The model will extend naturally from the end of this clip string いいえ -
first_frame_url First frame image URL for video_extension function. Specifies the starting frame when extending video forward. At least one of first_frame_url/last_frame_url/first_clip_url/last_clip_url must be provided. Use this to define the exact starting point of the extended video. The model will generate smooth transition from this frame string いいえ -
last_clip_url Last video clip URL for video_extension function. Provides a video segment to use as the ending portion. At least one of first_frame_url/last_frame_url/first_clip_url/last_clip_url must be provided. Use this when you want more control than a single frame can provide. The model will extend naturally to the beginning of this clip string いいえ -
last_frame_url Last frame image URL for video_extension function. Specifies the ending frame when extending video backward. At least one of first_frame_url/last_frame_url/first_clip_url/last_clip_url must be provided. Use this to define the exact ending point of the extended video. The model will generate smooth transition to this frame string いいえ -
left_scale Left boundary expansion scale for video_outpainting function (1.0-2.0). 1.0 = no expansion (original width), 2.0 = double the left area. Default is 1.0. Use values > 1.0 to expand the canvas leftward. For example, 1.5 means add 50% more canvas space to the left of the original video. The model will generate content to fill the expanded area naturally number いいえ -
mask_frame_id Frame ID (1-based index) indicating which frame the mask_image_url corresponds to. Only applicable when using mask_image_url in video_edit function. The mask will be propagated from this frame to others based on mask_type. For example, mask_frame_id=1 means the mask corresponds to the first frame of the video. Default is 1 (first frame) integer いいえ -
mask_image_url Mask image URL for video_edit function. Defines the region to edit (white=edit, black=keep). Must provide either mask_image_url OR mask_video_url, not both. When using mask_image_url, also specify mask_frame_id to indicate which frame the mask corresponds to. The mask will be propagated to other frames based on mask_type setting string いいえ -
mask_type Mask propagation type for video_edit function. Options: (1) ‘tracking’: mask follows object movement across frames (recommended for moving objects); (2) ‘fixed’: mask stays in same position across all frames (recommended for static scenes or background edits). Default is ‘tracking’ string いいえ tracking, fixed
mask_video_url Mask video URL for video_edit function. Provides frame-by-frame mask for precise control (white=edit, black=keep). Must provide either mask_image_url OR mask_video_url, not both. Mask video must have the same frame count as the input video. Use this for complex editing that requires different masks for different frames string いいえ -
obj_or_bg Classification for each reference image: ‘obj’ (object/subject) or ‘bg’ (background). Only for image_reference function. Array length must match ref_images_url length. Maximum 1 ‘bg’ element allowed. Helps model distinguish between subject references and background style references string[] いいえ obj, bg
prompt_extend Enable intelligent prompt rewriting and enhancement. When true (default), the model will automatically optimize and expand your prompt for better results. When false, uses your prompt exactly as provided. Recommended to keep true unless you need precise control over the exact wording. Applicable to all functions boolean いいえ true, false
ref_images_url Reference image URLs array. Usage by function: (1) image_reference: 1-3 images required, used as visual references for video generation; (2) video_repainting: max 1 image optional, provides style reference; (3) video_edit: max 1 image optional, provides style reference; (4) other functions: not used. Images must be publicly accessible HTTP/HTTPS URLs string[] いいえ -
right_scale Right boundary expansion scale for video_outpainting function (1.0-2.0). 1.0 = no expansion (original width), 2.0 = double the right area. Default is 1.0. Use values > 1.0 to expand the canvas rightward. For example, 1.5 means add 50% more canvas space to the right of the original video. The model will generate content to fill the expanded area naturally number いいえ -
seed Random seed for reproducible results (0-2147483647). Using the same seed with identical parameters will produce similar (though not pixel-perfect identical) results. Useful for A/B testing different prompts while keeping other randomness constant. If not specified, a random seed is used each time. Applicable to all functions integer いいえ -
size Output video resolution in widthheight format. Available options: '1280720’ (16:9 landscape, HD), ‘7201280’ (9:16 portrait, mobile-friendly), '960960’ (1:1 square, social media), ‘1088832’ (4:3 landscape), '8321088’ (3:4 portrait). Default is ‘1280*720’. Choose based on your target platform and use case. Applicable to all functions string いいえ 1280*720, 720*1280, 960*960, 832*1088, 1088*832
top_scale Top boundary expansion scale for video_outpainting function (1.0-2.0). 1.0 = no expansion (original height), 2.0 = double the top area. Default is 1.0. Use values > 1.0 to expand the canvas upward. For example, 1.5 means add 50% more canvas space above the original video. The model will generate content to fill the expanded area naturally number いいえ -
video_url Input video URL. Usage by function: (1) video_repainting: required, the video to be repainted; (2) video_edit: required, the video to be edited; (3) video_outpainting: required, the video to expand; (4) video_extension: optional, provides reference style when extending; (5) image_reference: not used. Must be publicly accessible HTTP/HTTPS/OSS URL string いいえ -

料金

単位: $/sec

料金
$0.0591/sec