Nano Banana 2.1 現已上線 — Google 最新模型 | 立即體驗 →
Wan 3.0 Models

Wan 3.0 Models

Wan 3.0 delivers cinematic AI video generation for text-to-video, image-to-video, and reference-to-video workflows

Wan 3.0 delivers cinematic AI video generation for text-to-video, image-to-video, and reference-to-video workflows

所有模型

10 個模型
alibaba/wan-3.0-prime/image-to-video5% OFF
image-to-video$0.7500$0.7125

alibaba/wan-3.0-prime/image-to-video

Wan 3.0 Prime Image to Video is an accelerated variant that animates a first-frame image into a cinematic video, with optional last-frame guidance, flexible 2-30 second duration, aspect ratio control, optional audio, and deep-thinking controls for high-quality motion and scene continuity. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

alibaba/wan-3.0-prime/text-to-video5% OFF
text-to-video$0.7500$0.7125

alibaba/wan-3.0-prime/text-to-video

Wan 3.0 Prime Text to Video is an accelerated variant that generates cinematic videos from text prompts, with flexible 2-30 second duration, aspect ratio control, optional audio, and deep-thinking controls for high-quality video generation. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

alibaba/wan-3.0-prime/reference-to-video5% OFF
reference-to-video$0.7500$0.7125

alibaba/wan-3.0-prime/reference-to-video

Wan 3.0 Prime Reference to Video is an accelerated variant that creates coherent videos from prompts and multimodal references, including images, videos, and audio, with flexible 2-30 second duration and aspect ratio control for subject consistency, motion guidance, timing control, and scene continuity. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

alibaba/wan-3.0-prime/video-edit
video-to-video$0.7500

alibaba/wan-3.0-prime/video-edit

Wan 3.0 Prime Video Edit edits existing videos with text prompts and optional reference images or audio, supporting prompt-guided scene changes, visual refinements, and multimodal video editing. Inputs longer than 15 seconds are trimmed to the first 15 seconds, with output aspect ratio automatically matched to the input video. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

alibaba/wan-3.0-prime/video-extend
video-extend$0.7500

alibaba/wan-3.0-prime/video-extend

Wan 3.0 Prime Video Extend continues existing videos by generating a new 2-30 second segment from the final frame and appending it to the retained source video. It preserves existing source audio, supports optional audio generation for the new segment, retains up to the last 120 seconds of input, and outputs at 480P / 720P / 1080P. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

alibaba/wan-3.0/text-to-video5% OFF
text-to-video$0.5000$0.4750

alibaba/wan-3.0/text-to-video

Wan 3.0 Text to Video generates cinematic videos from text prompts, with flexible 2-30 second duration, aspect ratio control, optional audio, and deep-thinking controls for high-quality video generation. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

alibaba/wan-3.0/image-to-video5% OFF
image-to-video$0.5000$0.4750

alibaba/wan-3.0/image-to-video

Wan 3.0 Image to Video animates a first-frame image into a cinematic video, with optional last-frame guidance, flexible 2-30 second duration, aspect ratio control, optional audio, and deep-thinking controls for high-quality motion and scene continuity. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

alibaba/wan-3.0/reference-to-video5% OFF
reference-to-video$0.5000$0.4750

alibaba/wan-3.0/reference-to-video

Wan 3.0 Reference to Video creates coherent videos from prompts and multimodal references, including images, videos, and audio, with flexible 2-30 second duration and aspect ratio control for subject consistency, motion guidance, timing control, and scene continuity. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

alibaba/wan-3.0/video-edit
video-to-video$0.5000

alibaba/wan-3.0/video-edit

Wan 3.0 Video Edit edits existing videos with text prompts and optional reference images or audio, supporting prompt-guided scene changes, visual refinements, and multimodal video editing. Inputs longer than 15 seconds are trimmed to the first 15 seconds, with output aspect ratio automatically matched to the input video. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

alibaba/wan-3.0/video-extend
video-extend$0.5000

alibaba/wan-3.0/video-extend

Wan 3.0 Video Extend continues an existing video with a newly generated segment, preserving the source video and its original audio while supporting optional audio generation for the extension. It supports 480P / 720P / 1080P output, optional last-frame guidance, and prompt control for the next action and camera movement. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

Wan 3.0 Models

Alibaba Wan 3.0 provides an advanced AI video generation model suite for text-to-video, image-to-video, and reference-to-video workflows. The collection is designed for creators, developers, marketers, studios, and AI video applications that need cinematic motion, stable subject rendering, strong prompt understanding, and scalable video generation APIs.

Built on the Wan video model family, Wan 3.0 helps users generate videos from text prompts, animate still images into dynamic video clips, and create reference-guided videos that preserve subject identity, visual style, and scene consistency. It is suitable for social media videos, ad creatives, product showcases, character scenes, storytelling, concept visualization, and commercial video production workflows.

Core Model Capabilities

Text-to-Video Generation:

Create cinematic videos directly from natural-language prompts with scene understanding, subject motion, camera movement, lighting direction, and visual style control.

Image-to-Video Generation:

Animate still images into dynamic video clips while preserving the original subject, composition, identity, and visual style.

Reference-to-Video Generation:

Generate videos from reference images or visual inputs while maintaining character identity, object appearance, visual style, and scene continuity.

Video Editing:

Edit and transform existing videos with natural-language instructions while preserving motion continuity, subject structure, scene coherence, and overall visual style. Wan 3.0 Video Edit is suitable for scene refinement, subject modification, visual restyling, object changes, and production-ready creative video adaptation.

Video Extension:

Extend existing video clips into longer continuous videos while preserving motion continuity, character identity, scene coherence, and overall visual style. Wan 3.0 Video Extend is suitable for story continuation, cinematic sequence expansion, short-form content extension, and production-ready creative video workflows.

Cinematic Motion Quality:

Produce videos with smooth motion, coherent scene structure, stable subjects, and natural visual transitions for creative and commercial use cases.

Reference-Based Consistency:

Use reference materials to guide character appearance, object details, style direction, and story continuity across generated clips.

Creative Video Production:

Support short-form videos, product ads, brand visuals, cinematic concepts, storytelling scenes, character-driven clips, and AI-powered video production pipelines.

Developer-Friendly Video API:

Access Wan 3.0 models through scalable APIs for automated video generation, fast creative iteration, and production-ready video workflows.

Wan 3.0 Prime Workflows:

Use Wan 3.0 Prime models for higher-fidelity text-to-video, image-to-video, and reference-to-video generation. Prime workflows provide stronger detail rendering, improved motion quality, more stable subject identity, and more production-ready cinematic results for demanding creative and commercial video projects.

Alibaba Wan 3.0 Models on WaveSpeedAI give creators and developers fast access to text-to-video, image-to-video, reference-to-video, video editing, video extension, and Prime video generation workflows with flexible pricing, scalable API access, and production-ready video quality.

Wan 3.0 Models API — 價格與效能

透過單一 REST API 執行 Wan 3.0 Models 系列中的任何模型。按生成計費 — 無訂閱、無最低消費 — 在可用率 99.9% 的基礎架構上提供業界領先的延遲。

為什麼在 WaveSpeedAI 上執行 Wan 3.0 Models

透明的價格

每個 Wan 3.0 Models 模型都採按呼叫計費。價格列在每個模型的頁面上 — 不會額外加收平台費。

為低延遲最佳化

大多數 Wan 3.0 Models 影像模型在 2 秒內完成。影片與 3D 模型比自架方案快數倍。

99.9% 可用率

多區域故障轉移與自動重試可在供應商故障期間 — 仍將您的生產流量保持線上。

常見問題

Wan 3.0 Models API 多少錢?+

每個模型在其模型頁面上都列有自己的按呼叫價格。我們按每次成功生成計費,沒有訂閱費或最低消費。

Wan 3.0 Models 模型在 WaveSpeedAI 上有多快?+

本系列中的影像模型通常在 2 秒內完成。影片與 3D 模型取決於長度與解析度,但通常比自架執行快數倍。

不用信用卡可以試用 API 嗎?+

符合資格的新帳戶可能獲得 $1 的推廣額度,用於無需信用卡試用 Wan 3.0 Models 模型。並非每次註冊都會獲得試用額度,請在生成前查看帳戶餘額。

有速率限制嗎?+

標準帳戶具有充足的並行任務限制。Enterprise 方案提供自訂 RPM、更高並行性和專屬容量 — 詳情請聯繫業務。