Nano Banana 2.1 现已上线 — Google 最新模型 | 立即体验 →
Wan 3.0 Models

Wan 3.0 Models

Wan 3.0 delivers cinematic AI video generation for text-to-video, image-to-video, and reference-to-video workflows

Wan 3.0 delivers cinematic AI video generation for text-to-video, image-to-video, and reference-to-video workflows

所有模型

10 个模型
alibaba/wan-3.0-prime/image-to-video5% OFF
image-to-video$0.7500$0.7125

alibaba/wan-3.0-prime/image-to-video

Wan 3.0 Prime Image to Video is an accelerated variant that animates a first-frame image into a cinematic video, with optional last-frame guidance, flexible 2-30 second duration, aspect ratio control, optional audio, and deep-thinking controls for high-quality motion and scene continuity. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

alibaba/wan-3.0-prime/text-to-video5% OFF
text-to-video$0.7500$0.7125

alibaba/wan-3.0-prime/text-to-video

Wan 3.0 Prime Text to Video is an accelerated variant that generates cinematic videos from text prompts, with flexible 2-30 second duration, aspect ratio control, optional audio, and deep-thinking controls for high-quality video generation. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

alibaba/wan-3.0-prime/reference-to-video5% OFF
reference-to-video$0.7500$0.7125

alibaba/wan-3.0-prime/reference-to-video

Wan 3.0 Prime Reference to Video is an accelerated variant that creates coherent videos from prompts and multimodal references, including images, videos, and audio, with flexible 2-30 second duration and aspect ratio control for subject consistency, motion guidance, timing control, and scene continuity. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

alibaba/wan-3.0-prime/video-edit
video-to-video$0.7500

alibaba/wan-3.0-prime/video-edit

Wan 3.0 Prime Video Edit edits existing videos with text prompts and optional reference images or audio, supporting prompt-guided scene changes, visual refinements, and multimodal video editing. Inputs longer than 15 seconds are trimmed to the first 15 seconds, with output aspect ratio automatically matched to the input video. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

alibaba/wan-3.0-prime/video-extend
video-extend$0.7500

alibaba/wan-3.0-prime/video-extend

Wan 3.0 Prime Video Extend continues existing videos by generating a new 2-30 second segment from the final frame and appending it to the retained source video. It preserves existing source audio, supports optional audio generation for the new segment, retains up to the last 120 seconds of input, and outputs at 480P / 720P / 1080P. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

alibaba/wan-3.0/text-to-video5% OFF
text-to-video$0.5000$0.4750

alibaba/wan-3.0/text-to-video

Wan 3.0 Text to Video generates cinematic videos from text prompts, with flexible 2-30 second duration, aspect ratio control, optional audio, and deep-thinking controls for high-quality video generation. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

alibaba/wan-3.0/image-to-video5% OFF
image-to-video$0.5000$0.4750

alibaba/wan-3.0/image-to-video

Wan 3.0 Image to Video animates a first-frame image into a cinematic video, with optional last-frame guidance, flexible 2-30 second duration, aspect ratio control, optional audio, and deep-thinking controls for high-quality motion and scene continuity. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

alibaba/wan-3.0/reference-to-video5% OFF
reference-to-video$0.5000$0.4750

alibaba/wan-3.0/reference-to-video

Wan 3.0 Reference to Video creates coherent videos from prompts and multimodal references, including images, videos, and audio, with flexible 2-30 second duration and aspect ratio control for subject consistency, motion guidance, timing control, and scene continuity. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

alibaba/wan-3.0/video-edit
video-to-video$0.5000

alibaba/wan-3.0/video-edit

Wan 3.0 Video Edit edits existing videos with text prompts and optional reference images or audio, supporting prompt-guided scene changes, visual refinements, and multimodal video editing. Inputs longer than 15 seconds are trimmed to the first 15 seconds, with output aspect ratio automatically matched to the input video. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

alibaba/wan-3.0/video-extend
video-extend$0.5000

alibaba/wan-3.0/video-extend

Wan 3.0 Video Extend continues an existing video with a newly generated segment, preserving the source video and its original audio while supporting optional audio generation for the extension. It supports 480P / 720P / 1080P output, optional last-frame guidance, and prompt control for the next action and camera movement. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

Wan 3.0 Models

Alibaba Wan 3.0 provides an advanced AI video generation model suite for text-to-video, image-to-video, and reference-to-video workflows. The collection is designed for creators, developers, marketers, studios, and AI video applications that need cinematic motion, stable subject rendering, strong prompt understanding, and scalable video generation APIs.

Built on the Wan video model family, Wan 3.0 helps users generate videos from text prompts, animate still images into dynamic video clips, and create reference-guided videos that preserve subject identity, visual style, and scene consistency. It is suitable for social media videos, ad creatives, product showcases, character scenes, storytelling, concept visualization, and commercial video production workflows.

Core Model Capabilities

Text-to-Video Generation:

Create cinematic videos directly from natural-language prompts with scene understanding, subject motion, camera movement, lighting direction, and visual style control.

Image-to-Video Generation:

Animate still images into dynamic video clips while preserving the original subject, composition, identity, and visual style.

Reference-to-Video Generation:

Generate videos from reference images or visual inputs while maintaining character identity, object appearance, visual style, and scene continuity.

Video Editing:

Edit and transform existing videos with natural-language instructions while preserving motion continuity, subject structure, scene coherence, and overall visual style. Wan 3.0 Video Edit is suitable for scene refinement, subject modification, visual restyling, object changes, and production-ready creative video adaptation.

Video Extension:

Extend existing video clips into longer continuous videos while preserving motion continuity, character identity, scene coherence, and overall visual style. Wan 3.0 Video Extend is suitable for story continuation, cinematic sequence expansion, short-form content extension, and production-ready creative video workflows.

Cinematic Motion Quality:

Produce videos with smooth motion, coherent scene structure, stable subjects, and natural visual transitions for creative and commercial use cases.

Reference-Based Consistency:

Use reference materials to guide character appearance, object details, style direction, and story continuity across generated clips.

Creative Video Production:

Support short-form videos, product ads, brand visuals, cinematic concepts, storytelling scenes, character-driven clips, and AI-powered video production pipelines.

Developer-Friendly Video API:

Access Wan 3.0 models through scalable APIs for automated video generation, fast creative iteration, and production-ready video workflows.

Wan 3.0 Prime Workflows:

Use Wan 3.0 Prime models for higher-fidelity text-to-video, image-to-video, and reference-to-video generation. Prime workflows provide stronger detail rendering, improved motion quality, more stable subject identity, and more production-ready cinematic results for demanding creative and commercial video projects.

Alibaba Wan 3.0 Models on WaveSpeedAI give creators and developers fast access to text-to-video, image-to-video, reference-to-video, video editing, video extension, and Prime video generation workflows with flexible pricing, scalable API access, and production-ready video quality.

Wan 3.0 Models API — 价格与性能

通过单一 REST API 运行 Wan 3.0 Models 系列中的任意模型。按生成计费 — 无订阅、无最低消费 — 在 99.9% 可用性的基础设施上提供行业领先的延迟。

为什么在 WaveSpeedAI 上运行 Wan 3.0 Models

透明定价

每个 Wan 3.0 Models 模型都有按调用计价。价格在每个模型的页面上列出 — 不收取额外的平台费。

为低延迟优化

大多数 Wan 3.0 Models 图像模型在 2 秒内完成。视频和 3D 模型比自托管方案快数倍。

99.9% 可用性

多区域故障转移和自动重试可确保您的生产流量保持在线 — 即使在供应商故障期间。

常见问题

Wan 3.0 Models API 多少钱?+

每个模型在其模型页面上都列有自己的按调用价格。我们按每次成功生成计费,没有订阅费或最低消费。

Wan 3.0 Models 模型在 WaveSpeedAI 上有多快?+

本系列中的图像模型通常在 2 秒内完成。视频和 3D 模型取决于时长和分辨率,但通常比自托管运行快数倍。

不用信用卡可以试用 API 吗?+

符合条件的新账户可能获得 $1 的推广额度,用于无需信用卡试用 Wan 3.0 Models 模型。并非每次注册都会获得试用额度,请在生成前查看账户余额。

有速率限制吗?+

标准账户有充足的并发任务限制。企业版计划提供自定义 RPM、更高并发和专用容量 — 详情请联系销售。