Seedance 2.5 Now Live | Try in Video Generator →
Wan 2.5 Models

Wan 2.5 Models

Wan 2.5 enables synchronized audio-video generation in a single step, delivering richer detail, smoother motion, and more immersive storytelling.

Wan 2.5 enables synchronized audio-video generation in a single step, delivering richer detail, smoother motion, and more immersive storytelling.

All models

8 models
alibaba/wan-2.5/image-to-video
image-to-video

alibaba/wan-2.5/image-to-video

WAN 2.5 converts text or images into videos (480p/720p/1080p) with synced audio, faster and more affordable than Google Veo3. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

alibaba/wan-2.5/image-to-video-fast
image-to-video

alibaba/wan-2.5/image-to-video-fast

WAN 2.5 Fast converts text or images into synchronized-audio videos in 480p, 720p, or 1080p, offering faster, more affordable generation compared to Google Veo3. REST API, no coldstarts, affordable pricing.

alibaba/wan-2.5/text-to-video-fast
text-to-video

alibaba/wan-2.5/text-to-video-fast

WAN 2.5 Fast creates synchronized-audio videos from text or images in 720p, faster and more affordable than Google Veo3. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

alibaba/wan-2.5/text-to-video
text-to-video

alibaba/wan-2.5/text-to-video

WAN 2.5 makes 480p-1080p text/image-to-video with synced audio and is faster, more affordable than Google Veo3. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

alibaba/wan-2.5/video-extend
video-extend

alibaba/wan-2.5/video-extend

WAN 2.5 Video-Extend turns short clips into longer videos with preserved or generated synchronized audio for continuity. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

alibaba/wan-2.5/video-extend-fast
video-extend

alibaba/wan-2.5/video-extend-fast

WAN 2.5 Fast is an AI-powered video extender that turns short clips into longer videos while preserving audio tracks for seamless extensions. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

alibaba/wan-2.5/text-to-image
text-to-image

alibaba/wan-2.5/text-to-image

WAN 2.5 Text-to-Image turns text prompts into AI-generated images with the WAN 2.5 model for on-demand image creation. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

alibaba/wan-2.5/image-edit
image-to-image

alibaba/wan-2.5/image-edit

Refine existing visuals with WAN 2.5 image-edit using prompt-driven adjustments and stylistic upgrades for photos and graphics. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

Wan 2.5 Models

WAN 2.5 on DashScope: convert text or images into lip-synced HD videos (480p/720p/1080p) in one step — faster and more budget-friendly than Veo 3, perfect for quick, audio-embedded content. Video generation is available for durations between 3s and 10s, with flexible options for each selection.

Model Lineup

  1. wan-2.5/text-to-video
  2. wan-2.5/image-to-video
  3. wan-2.5/text-to-image
  4. wan-2.5/image-edit
  5. wan-2.5/video-extend

Why Wan 2.5?

  1. More affordable — Lower overall cost than Veo 3; efficient for batch production.
  2. One-pass A/V sync — Generate video with voiceover + lip-sync in a single run—no separate VO or manual alignment.
  3. Multilingual that works — Reliable A/V sync for Chinese and minor languages (Veo 3 often shows “unknown language”).
  4. Longer more flexible — Up to 10 seconds (vs. ~8 seconds on Veo 3) and three aspect ratios for different platforms.
  5. Audio-driven control — Use voice/SFX/BGM as references to guide generation (Veo 3 doesn’t support audio references).

See WAN 2.5 vs. Veo 3

Veo3 VS Wan 2.5 effect compare

Great for

  1. Shorts — 3–10s hooks for TikTok/Reels. e.g., “Dynamic city night shot, upbeat VO summarizing three tips.”
  2. Ads & E-commerce — Product hero shots + CTA. e.g., “Rotate sneaker, macro textures, VO: ‘Lightweight, all-day comfort.’”
  3. Explainers/Tutorials — Step-by-step with on-beat VO. e.g., “3-step setup, captions auto-timed to narration.”

Wan 2.5 Models API — pricing & performance

Run any model in the Wan 2.5 Models collection through a single REST API. Pay per generation — no subscriptions, no minimums — with industry-leading latency on a 99.9% uptime infrastructure.

Why run Wan 2.5 Models on WaveSpeedAI

Transparent pricing

Per-call pricing for every Wan 2.5 Models model. The price is listed on each model page — no platform fees on top.

Optimized for low latency

Most Wan 2.5 Models image models complete in under 2 seconds. Video and 3D models run several times faster than self-hosted alternatives.

99.9% uptime

Multi-region failover and automatic retries keep your production traffic online — even during provider outages.

Frequently asked questions

How much does the Wan 2.5 Models API cost?+

Each model has its own per-call price listed on the model page. We bill per successful generation, with no subscription fees or minimums.

How fast are Wan 2.5 Models models on WaveSpeedAI?+

Image models in this collection typically complete in under 2 seconds. Video and 3D models depend on duration and resolution but are usually several times faster than self-hosted runs.

Can I try the API without a credit card?+

Yes — every account gets $1 in free credits on signup, enough to try most Wan 2.5 Models models without a credit card.

Are there rate limits?+

Standard accounts have generous concurrent-job limits. Enterprise plans offer custom RPM, higher concurrency, and dedicated capacity — contact sales for details.