Wan 2.2 API
Alibaba의 Wan 2.2 — WaveSpeedAI에 배포된 35개 이상의 자체 변형을 갖춘 오픈 웨이트 비디오 툴킷입니다. Animate(120초 캐릭터 애니메이션), Video Edit, Speech-to-Video(10분 오디오 기반), Fun-Control(Apache 2.0 라이선스)과 여러 모델 크기(5B, A14B) 및 해상도(480p / 720p)의 이미지-투-비디오와 텍스트-투-비디오를 제공합니다.
WaveSpeedAI가 호스팅하는 변형만 해당합니다. Animate는 최대 120초의 720p 클립을, Speech-to-Video는 최대 10분의 480p 클립을 만들며, Fun-Control은 상업적 사용이 가능한 Apache 2.0 하에서 프리셋 Control Codes를 사용합니다. LoRA 학습 엔드포인트는 몇 분 만에 파인튜닝합니다.
개요
Wan 2.2 API 소개
Wan 2.2의 기능, Alibaba 모델 라인업에서의 위치, 그리고 팀들이 이 모델을 선택하는 이유.
Alibaba의 비디오 생성 모델 Wan 2.2. WaveSpeedAI REST API로 바로 사용할 수 있습니다. Alibaba의 Wan 2.2 — WaveSpeedAI에 배포된 35개 이상의 자체 변형을 갖춘 오픈 웨이트 비디오 툴킷입니다. Animate(120초 캐릭터 애니메이션), Video Edit, Speech-to-Video(10분 오디오 기반), Fun-Control(Apache 2.0 라이선스)과 여러 모델 크기(5B, A14B) 및 해상도(480p / 720p)의 이미지-투-비디오와 텍스트-투-비디오를 제공합니다.
WaveSpeedAI가 호스팅하는 변형만 해당합니다. Animate는 최대 120초의 720p 클립을, Speech-to-Video는 최대 10분의 480p 클립을 만들며, Fun-Control은 상업적 사용이 가능한 Apache 2.0 하에서 프리셋 Control Codes를 사용합니다. LoRA 학습 엔드포인트는 몇 분 만에 파인튜닝합니다.
WaveSpeedAI의 Wan 2.2 제품군은 Image-To-Video, Motion-Control, Video-To-Video, Image-To-Image, Digital-Human, Text-To-Image, Training, Text-To-Video 워크플로를 아우르는 REST 엔드포인트 32개를 제공합니다. 각 변형은 고유한 가격, 파라미터, 예시 출력을 갖고 있으니 입력 방식과 운영 조건에 맞는 것을 고르거나, 같은 API 키로 여러 개를 호출해 다단계 파이프라인을 구성하세요.
Wan 2.2 실행에도 WaveSpeedAI의 다른 1,000개 이상의 AI 모델과 같은 API 키, 결제 계정, 요청 한도 체계를 그대로 사용합니다. 별도의 벤더 설정도, 제공사별 SDK도, 벤더별 요청 한도 체계도 필요 없습니다. 하나의 연동으로 텍스트 → 이미지, 텍스트 → 비디오부터 오디오 합성, 3D 생성, 업스케일링, 편집까지 모두 지원합니다.
엔드포인트
Wan 2.2 API 전체 엔드포인트
Wan 2.2 엔드포인트 32개를 지금 WaveSpeedAI에서 사용할 수 있습니다. 워크플로에 맞는 변형을 고르세요.
/filters:quality(82)/media/images/20260408111216_2fizlz7x.webp)
Wan 2.2 Image To Video Lora
Wan-2.2/image-to-video-lora enables unlimited image-to-video generation from a single image, producing smooth, cinematic motion with clean detail. Supports custom LoRAs for style and character consistency. Ready-to-use REST inference API, best performance, no cold starts, affordable pricing.
/filters:quality(82)/media/images/20260408111152_zehbgma1.webp)
Wan 2.2 Image To Video
Wan 2.2 Image-to-Video turns a single image into smooth, cinematic motion with clean detail—ideal for storyboards, mood shots, and product demos. Ready-to-use REST inference API, best performance, no cold starts, affordable pricing.
/filters:quality(82)/media/images/20260408111211_p8f6ffxe.webp)
Wan 2.2 Animate
Wan2.2-Animate unified character animation & replacement model replicating movement and expression; generates 720p videos up to 120s. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/20260408111241_3bzr5rd1.webp)
Wan 2.2 Video Edit
Wan 2.2 Video Edit lets you modify videos via text prompts (e.g., change clothing or characters). Powered by Wan 2.2, it supports 480p ($0.20/5s) and 720p ($0.40/5s), up to 120s. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/20260408111257_4rmbjsgj.webp)
Wan 2.2 Image To Image
WAN 2.2 (14B) is an image-to-image model for high-resolution photorealistic image editing with exceptional precision and fidelity. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/20260408111154_gvj4f5sj.webp)
Wan 2.2 Speech To Video
Wan-2.2-S2V turns images and speech into high-fidelity videos with realistic face and body motion; supports up to 10-minute clips in 480p, from $0.15/5s. Ready-to-use REST API, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/20260408111206_p4ked9pu.webp)
Wan 2.2 Text To Image Lora
WAN 2.2 generates super-detailed images from text prompts and supports custom LoRAs for fine-grained style and subject control. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/20260408111250_80jkk4zs.webp)
Wan 2.2 Fun Control
Wan2.2-Fun-Control uses Control Codes and multi-modal inputs to generate preset-controlled videos up to 120s at 720p; released under Apache 2.0 for commercial use. Ready-to-use REST API, no coldstarts, affordable.
/filters:quality(82)/media/images/20260408112009_zkynaf2n.webp)
Wan 2.2 Image Lora Trainer
Train custom Wan 2.2 character/style LoRA models 10x faster. Style training, character training, object training. From concept to model in minutes, not hours. Upload a ZIP file containing images to start!
/filters:quality(82)/media/images/20260408111954_4jny6q8q.webp)
Wan 2.2 I2v Lora Trainer
Train custom Wan 2.2 I2V LoRA models 10x faster. Action training, motion training, video efect training. From concept to model in minutes, not hours. Upload a ZIP file containing videos to start!
/filters:quality(82)/media/images/20260408111226_b3370qq2.webp)
Wan 2.2 Text To Image Realism
WAN 2.2 delivers ultra-realistic text-to-image generation, converting prompts into photoreal images with high fidelity and detail. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/20260408111246_kbtz9m3g.webp)
Wan 2.2 I2v 5b 720p Lora
Wan 2.2 i2v-5B-720p is a 5B image-to-video model producing 720p videos with LoRA support for style customization. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/20260408111210_znuyn6r8.webp)
Wan 2.2 T2v 5b 720p Lora
Wan 2.2 T2V 5B is a 5B text-to-video model with LoRA support that generates 720p videos from text prompts for easy personalization. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/20260408111236_0was1j4o.webp)
Wan 2.2 I2v 480p Lora Ultra Fast
Wan 2.2 i2v delivers ultra-fast Image-to-Video at 480p with support for custom LoRAs for tailored styles. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/20260408111252_sejz6k48.webp)
Wan 2.2 I2v 480p Ultra Fast
Wan 2.2 A14B Image-to-Video (i2v-480p) produces ultra-fast 480p videos from single images, enabling unlimited AI video generation with high throughput. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/20260408111149_z1lz7r9s.webp)
Wan 2.2 I2v 720p Ultra Fast
Generate unlimited ultra-fast 720p AI videos from images with Wan 2.2 A14B image-to-video model. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/20260408111259_8bca6477.webp)
Wan 2.2 T2v 480p Lora Ultra Fast
Ultra-fast Wan 2.2 text-to-video model producing 480p videos with custom LoRA support—generate unlimited AI videos with personalized styles. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/20260408111238_h7qxvir0.webp)
Wan 2.2 I2v 720p Lora Ultra Fast
Wan 2.2 i2v 720P is an ultra-fast Image-to-Video model that generates unlimited AI videos and supports custom LoRAs for personalized outputs. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/20260408111248_iyye47z8.webp)
Wan 2.2 T2v 480p Ultra Fast
Wan 2.2 t2v 480p Ultra-Fast generates unlimited AI videos from text prompts at 480p with ultra-fast inference. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/20260408111201_lv1td92d.webp)
Wan 2.2 T2v 5b 720p
Wan 2.2 T2V 5B is a 720P text-to-video model that generates unlimited AI videos from simple text prompts, producing consistent high-quality 720p outputs. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/20260408111221_xo8hpiqi.webp)
Wan 2.2 I2v 5b 720p
Wan 2.2 I2V 5B converts images into high-quality 720P videos using a 5B image-to-video model for AI video generation. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/20260408111214_rct2hj0c.webp)
Wan 2.2 I2v 480p
Wan 2.2 A14B converts images into 480p videos, enabling unlimited AI video generation from single images. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/20260408111234_cplw518o.webp)
Wan 2.2 I2v 480p Lora
WAN 2.2 A14B Image-to-Video model generates unlimited 480p videos from images and supports custom LoRAs for personalized styles and fine-tuning. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/20260408111232_8tgp9v7m.webp)
Wan 2.2 I2v 720p Lora
WAN 2.2 Image-to-Video (i2v) 720p converts images into 720p videos and supports custom LoRAs for style personalization. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/20260408111239_qlur3dn5.webp)
Wan 2.2 I2v 720p
WAN 2.2 A14B i2v-720p converts images into smooth 720p videos, enabling unlimited AI video generation with the Wan 2.2 image-to-video model. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/20260408111156_b3olejjy.webp)
Wan 2.2 T2v 480p
Wan 2.2 t2v-480p generates unlimited AI videos from text prompts at 480p resolution, ideal for rapid prototyping and content creation. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/20260408111208_sw0j3xam.webp)
Wan 2.2 T2v 480p Lora
WAN 2.2 T2V 480p with LoRA generates text-to-video at 480p and supports custom LoRAs for personalized styles. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/20260408111218_00vzkvbl.webp)
Wan 2.2 T2v 720p
Wan 2.2 t2v-720p converts text prompts into native 720P videos, producing high-quality 720P clips from simple prompts. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/20260408111230_mtxgxh1x.webp)
Wan 2.2 T2v 720p Lora
Wan 2.2 T2V 720p with custom LoRA support turns text prompts into 720p AI videos and enables unlimited video generation. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/20260408111158_csf9651o.webp)
Wan 2.2 T2v 720p Lora Ultra Fast
Ultra-fast Wan 2.2 Text-to-Video generates unlimited 720p AI videos with custom LoRAs for personalized styles. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/20260408111244_aqp0x6b0.webp)
Wan 2.2 T2v 720p Ultra Fast
WAN 2.2 T2V 720p Ultra-Fast generates high-quality 720p videos from text prompts with unlimited output and ultra-fast throughput. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/1786269272686834557_4oOX6gqz.webp)
Wan 2.2 Animate 2
Wan 2.2 Animate 2 is the next-generation Wan character animation model: an end-to-end DiT that makes the character in a reference image perform the motion of a driving video, with no pose extraction, prompt-controlled background, and strong identity preservation; generates 480p/720p videos up to 120s. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
예시
Wan 2.2 활용 예시
Wan 2.2 API로 생성한 실제 출력물입니다. 비디오 위에 마우스를 올리면 미리 보고, 클릭하면 전체 크기 뷰어가 열립니다.
사용 방법
Wan 2.2 API 사용 방법
가입부터 결과물 생성까지 네 단계입니다. Python, Node.js, cURL 전체 예시는 아래 API 섹션에 있습니다.
- 01
API 키 받기
WaveSpeedAI 계정을 만들고 대시보드에서 API 키를 복사하세요. 신규 계정에는 무료 스타터 크레딧이 제공되어, 요금이 청구되기 전에 Playground를 수십 번 실행해 볼 수 있습니다.
- 02
Prediction 제출
입력을 JSON으로 https://api.wavespeed.ai/api/v3/wavespeed-ai/wan-2.2/image-to-video에 POST하세요. 엔드포인트가 prediction id를 즉시 반환합니다. 생성은 비동기로 처리되므로 추론 중에 연결을 계속 열어 둘 필요가 없습니다.
- 03
완료될 때까지 폴링
https://api.wavespeed.ai/api/v3/predictions/{request_id}/result로 GET 요청을 보내세요. completed이면 outputs를 반환하고, failed, cancelled, timeout, deleted이면 오류로 종료하며, 그 밖의 모든 status에서는 폴링을 계속합니다.
- 04
출력 URL 읽기
status가 "completed"가 되면 data.outputs[0]에서 URL을 읽으세요. 이 URL은 WaveSpeedAI CDN에 있는 생성된 미디어를 가리키며, 호출한 Wan 2.2 변형에 따라 이미지, 비디오, 오디오 또는 3D 파일입니다.
활용 사례
Wan 2.2 활용 사례
개발자와 크리에이터가 Wan 2.2 API를 사용하는 대표적인 워크플로입니다.
Wan 2.2 Animate — 최대 120초 캐릭터 애니메이션
wavespeed-ai/wan-2.2/animate는 "움직임과 표정을 재현하는 통합 캐릭터 애니메이션 및 교체 모델로, 최대 120초의 720p 비디오를 생성"합니다. 카탈로그 설명입니다. 대부분의 포즈 기반 애니메이션 도구보다 훨씬 깁니다.
최대 10분의 Speech-to-Video
wavespeed-ai/wan-2.2/speech-to-video는 "이미지와 음성을 사실적인 얼굴 및 신체 모션을 갖춘 고충실도 비디오로 변환하며, 480p로 최대 10분 클립을 지원"합니다. 롱폼 토킹 콘텐츠에 유용합니다.
프롬프트 기반 변경을 위한 Video Edit
wavespeed-ai/wan-2.2/video-edit로 텍스트 프롬프트를 통해 비디오를 수정할 수 있습니다(카탈로그 예시: 의상이나 캐릭터 변경). 480p와 720p, 최대 120초를 지원합니다.
Apache 2.0 라이선스의 Fun-Control
wavespeed-ai/wan-2.2/fun-control은 "Control Codes와 멀티모달 입력으로 720p에서 최대 120초의 프리셋 제어 비디오를 생성하며, 상업적 사용이 가능한 Apache 2.0으로 공개"되었습니다. Apache 2.0 라이선스는 상업용 파이프라인에서 실질적인 차별점입니다.
LoRA 학습(10배 빠름)
이미지 LoRA에는 wavespeed-ai/wan-2.2-image-lora-trainer, I2V LoRA에는 wavespeed-ai/wan-2.2-i2v-lora-trainer를 사용합니다. 카탈로그 설명: "10배 빠른 학습". 스타일, 캐릭터, 오브젝트, 모션, 액션, 비디오 효과 학습을 지원합니다. ZIP 파일을 업로드해 시작하세요.
여러 크기(5B / A14B)의 이미지-투-비디오
모델 크기를 고르세요. 속도와 비용에는 5B(더 작음), 최대 품질에는 A14B입니다. 480p용 표준 i2v, LoRA를 지원하는 변형, 초고속 변형이 있습니다.
팁
Wan 2.2 프롬프트 작성 팁
Wan 2.2에서 더 나은 결과를 얻기 위한 실용적인 조언입니다. 운영 파이프라인에서 비디오 모델 전반에 통하는 패턴을 바탕으로 했습니다.
- 01
작업에 맞는 변형을 고르세요
Wan 2.2는 하나의 범용 모델이 아니라 특화된 엔드포인트를 제공합니다. 포즈 기반 모션에는 Animate, 부분 수정에는 video-edit, 토킹 콘텐츠에는 speech-to-video, 일반 생성에는 image-to-video를 사용하세요. 엔드포인트를 작업에 맞추면 범용 모델에 모든 것을 시킬 때보다 출력이 훨씬 좋아집니다.
- 02
프로덕션 규모의 일관성을 위해 LoRA를 학습하세요
Wan 2.2의 LoRA 트레이너 엔드포인트는 부가 도구가 아니라 핵심 API 기능입니다. 수백 번의 생성에 걸쳐 반복되는 정체성(브랜드 마스코트, 고정 캐릭터, 시그니처 스타일)이 필요한 프로덕션에서는 LoRA를 한 번 학습하고 생성마다 LoRA 추론 엔드포인트를 호출하세요.
- 03
오픈 웨이트는 파인튜닝을 가능하게 합니다
클로즈드 웨이트 경쟁 모델은 벤더의 모델 동작에 묶이게 합니다. Wan 2.2의 기반은 오픈 웨이트이므로, 기본 모델이 해결하지 못하는 한계에 부딪히면 파인튜닝이라는 선택지가 있습니다. 대부분의 다른 상용 비디오 API는 이런 경로를 제공하지 않습니다.
- 04
Animate를 Kling Motion Control과 비교해 고르세요
둘 다 서로 다른 절충을 가진 포즈 기반 애니메이션을 제공합니다. Wan 2.2 Animate는 LoRA 지원과 오픈 웨이트를 갖추고 있고, Kling Motion Control은 정체성 유지가 더 강합니다. 파인튜닝이 더 중요한지, 정체성 품질이 더 중요한지에 따라 선택하세요.
가격
Wan 2.2 API 가격
가격은 출력당 책정됩니다. 최종 요금은 각 변형의 Playground에서 설정한 파라미터(해상도, 길이, 출력 수, 레퍼런스)에 따라 달라집니다.
API
Wan 2.2 API 호출하기
wavespeed.ai/accesskey에서 API 키를 발급받은 뒤 REST로 prediction을 제출하세요. Playground에서 입력 조합별로 바로 붙여넣을 수 있는 샘플이 생성됩니다.
POSThttps://api.wavespeed.ai/api/v3/wavespeed-ai/wan-2.2/image-to-video
# 1. Submit the prediction.
SUBMIT_RESPONSE=$(curl --silent --show-error --fail-with-body \
-X POST "https://api.wavespeed.ai/api/v3/wavespeed-ai/wan-2.2/image-to-video" \
-H "Content-Type: application/json" \
-H "Authorization: Bearer $WAVESPEED_API_KEY" \
-d '{
"prompt": "A cinematic shot of a city at sunset, soft golden light",
"image": "https://interactive-examples.mdn.mozilla.net/media/cc0-images/painted-hand-298-332.jpg",
"resolution": "480p",
"duration": 5
}')
TASK=$(printf '%s' "$SUBMIT_RESPONSE" | jq 'if has("data") then .data else . end')
PREDICTION_ID=$(printf '%s' "$TASK" | jq -r '.id')
if [ -z "$PREDICTION_ID" ] || [ "$PREDICTION_ID" = "null" ]; then
printf 'Submission response did not contain a prediction id
' >&2
exit 1
fi
RESULT_URL="https://api.wavespeed.ai/api/v3/predictions/$PREDICTION_ID/result"
# 2. Poll until the prediction finishes.
while true; do
RESPONSE=$(curl --silent --show-error --fail-with-body "$RESULT_URL" \
-H "Authorization: Bearer $WAVESPEED_API_KEY")
RESULT=$(printf '%s' "$RESPONSE" | jq 'if has("data") then .data else . end')
STATUS=$(printf '%s' "$RESULT" | jq -r '.status')
case "$STATUS" in
completed) printf '%s\n' "$RESULT" | jq '.outputs'; break ;;
failed|cancelled|timeout|deleted) printf '%s\n' "$RESULT" | jq . >&2; exit 1 ;;
*) sleep 2 ;;
esac
doneconst submitUrl = "https://api.wavespeed.ai/api/v3/wavespeed-ai/wan-2.2/image-to-video";
const apiKey = process.env.WAVESPEED_API_KEY;
if (!apiKey) throw new Error('Set WAVESPEED_API_KEY');
async function requestJson(url, options = {}) {
const response = await fetch(url, options);
if (!response.ok) throw new Error(await response.text());
return response.json();
}
// 1. Submit the prediction.
const body = await requestJson(submitUrl, {
method: "POST",
headers: {
"Authorization": `Bearer ${apiKey}`,
"Content-Type": "application/json",
},
body: JSON.stringify({
"prompt": "A cinematic shot of a city at sunset, soft golden light",
"image": "https://interactive-examples.mdn.mozilla.net/media/cc0-images/painted-hand-298-332.jpg",
"resolution": "480p",
"duration": 5
}),
});
const task = body.data ?? body;
const resultUrl = `https://api.wavespeed.ai/api/v3/predictions/${task.id}/result`;
// 2. Poll until the prediction finishes.
while (true) {
const resultBody = await requestJson(resultUrl, {
headers: { "Authorization": `Bearer ${apiKey}` },
});
const result = resultBody.data ?? resultBody;
if (result.status === "completed") {
console.log(result.outputs);
break;
}
if (["failed", "cancelled", "timeout", "deleted"].includes(result.status)) throw new Error(JSON.stringify(result));
await new Promise(resolve => setTimeout(resolve, 2000));
}import json
import os
import time
from urllib.request import Request, urlopen
api_key = os.environ["WAVESPEED_API_KEY"]
headers = {"Authorization": f"Bearer {api_key}", "Content-Type": "application/json"}
payload = {
"prompt": "A cinematic shot of a city at sunset, soft golden light",
"image": "https://interactive-examples.mdn.mozilla.net/media/cc0-images/painted-hand-298-332.jpg",
"resolution": "480p",
"duration": 5
}
def request_json(url, data=None):
request = Request(url, data=data, headers=headers, method="POST" if data else "GET")
with urlopen(request) as response:
return json.load(response)
# 1. Submit the prediction.
body = request_json("https://api.wavespeed.ai/api/v3/wavespeed-ai/wan-2.2/image-to-video", json.dumps(payload).encode())
task = body.get("data", body)
result_url = f"https://api.wavespeed.ai/api/v3/predictions/{task['id']}/result"
# 2. Poll until the prediction finishes.
while True:
result_body = request_json(result_url)
result = result_body.get("data", result_body)
status = result.get("status")
if status == "completed":
print(result.get("outputs", []))
break
if status in {"failed", "cancelled", "timeout", "deleted"}:
raise RuntimeError(result)
time.sleep(2)비교
Wan 2.2 vs 대안
비슷한 WaveSpeedAI 모델과 비교한 Wan 2.2의 선택 기준.
Wan 2.2 vs Wan 2.7
Wan 2.7(alibaba/wan-2.7/*)은 Alibaba의 더 새로운 아키텍처로, 한 패밀리에 레퍼런스-투-비디오, video-edit, image-edit, 텍스트-투-이미지가 있어 교차 모달 툴킷이 더 넓습니다. Wan 2.2(WaveSpeedAI 변형)는 2.7에는 없는 Animate(120초), Speech-to-Video(10분), Fun-Control(Apache 2.0), LoRA 트레이너 같은 특화 엔드포인트를 제공합니다.
Wan 2.2 vs Seedance 2.0
Seedance 2.0은 모든 변형에서 네이티브 오디오를 갖춘 할리우드급 출력을 제공합니다. Wan 2.2는 변형 구성(35개 이상의 엔드포인트)과 LoRA 학습에서 앞서며, 범용 비디오 모델이 아니라 특정 특화 기능(Animate, Speech-to-Video, Fun-Control)이 필요할 때 알맞은 선택입니다.
Wan 2.2 vs Kling 3.0 Motion Control
둘 다 포즈 기반 캐릭터 애니메이션을 제공합니다. Wan 2.2 Animate는 120초의 720p 클립을 만들고 LoRA 파인튜닝을 제공합니다. Kling Motion Control은 레퍼런스 비디오 길이(3~30초)로 제한되지만, 더 강한 모션 사전 지식을 위해 Kuaishou의 비디오 코퍼스로 학습되었습니다.
FAQ
Wan 2.2 API — 자주 묻는 질문
가격, 라이선스, 연동 — WaveSpeedAI에서 Wan 2.2 실행에 관한 자주 묻는 질문.
Wan 2.2 API란 무엇인가요?
Wan 2.2 — WaveSpeedAI에서 REST API로 제공되는 Alibaba의 비디오 생성 모델입니다. Alibaba의 Wan 2.2 — WaveSpeedAI에 배포된 35개 이상의 자체 변형을 갖춘 오픈 웨이트 비디오 툴킷입니다. Animate(120초 캐릭터 애니메이션), Video Edit, Speech-to-Video(10분 오디오 기반), Fun-Control(Apache 2.0 라이선스)과 여러 모델 크기(5B, A14B) 및 해상도(480p / 720p)의 이미지-투-비디오와 텍스트-투-비디오를 제공합니다. 프로그래밍 방식으로 호출하거나 위에 링크된 Playground에서 체험해 볼 수 있습니다.
Wan 2.2 API는 어떻게 호출하나요?
WaveSpeedAI 계정을 만들고 /accesskey에서 API 키를 복사한 뒤, 입력을 JSON으로 https://api.wavespeed.ai/api/v3/wavespeed-ai/wan-2.2/image-to-video에 POST하세요. 엔드포인트가 prediction id를 반환합니다. 결과 엔드포인트 폴링은 약 2초마다 시작하고, 오래 걸리는 작업은 간격을 늘리며, 종료 status가 나오면 중단하세요. 운영 환경용 Python / Node.js / cURL 예시는 위에 있습니다.
Wan 2.2 API 요금은 얼마인가요?
Wan 2.2의 가격은 회당 $0.02부터입니다. 정확한 비용은 설정한 파라미터(해상도, 길이, 출력 수, 레퍼런스)에 따라 달라집니다. Playground의 생성 버튼 옆 실시간 비용 미리보기에서 현재 입력 기준의 정확한 가격을 확인할 수 있습니다.
Wan 2.2에는 어떤 변형이 있나요?
WaveSpeedAI에는 라이브 Wan 2.2 엔드포인트가 32개 있습니다: wavespeed-ai/wan-2.2/image-to-video-lora, wavespeed-ai/wan-2.2/image-to-video, wavespeed-ai/wan-2.2/animate, wavespeed-ai/wan-2.2/video-edit, wavespeed-ai/wan-2.2/image-to-image, wavespeed-ai/wan-2.2/speech-to-video, wavespeed-ai/wan-2.2/text-to-image-lora, wavespeed-ai/wan-2.2/fun-control 외 다수. 각 변형에는 고유한 Playground 페이지와 가격이 있습니다.
Wan 2.2 출력물을 상업적으로 사용할 수 있나요?
상업적 사용 권한은 Alibaba 모델 라이선스를 따릅니다. 대부분의 Alibaba 모델은 출력물의 상업적 사용을 허용합니다. 구체적인 라이선스 요약은 각 모델의 Playground 페이지를, 플랫폼 수준의 조건은 WaveSpeedAI 이용약관을 참고하세요.
직접 연동하지 않고 WaveSpeedAI에서 Wan 2.2 모델을 사용하는 이유는 무엇인가요?
API 키 하나와 결제 계정 하나로 Wan 2.2 모델과 다른 제공사의 1,000개 이상의 AI 모델을 모두 사용합니다. 벤더별 SDK 설정도, 별도의 요청 한도 체계도, 벤더마다 연동 코드를 다시 쓸 일도 없습니다. 가격은 대체로 Alibaba의 직접 API와 같거나 더 저렴합니다.
제공사
Alibaba 소개
WaveSpeedAI의 Wan 2.2 및 Alibaba 전체 모델 라인업을 만든 팀.
Alibaba의 Tongyi Lab은 Wan 비디오 모델 패밀리와 Qwen LLM 패밀리를 만듭니다. Wan은 오픈 웨이트로 공개되었고, 폭넓은 변형(텍스트-투-비디오, 이미지-투-비디오, 레퍼런스-투-비디오, video-edit, video-extend, image-edit, 텍스트-투-이미지)을 갖추었으며, 다국어 프롬프트 전반에서 모션 안정성과 프롬프트 반영이 꾸준히 강한 점이 특징입니다.
WaveSpeedAI에서 Wan 2.2 시작하기
가입 시 무료 스타터 크레딧 제공. Alibaba 및 다른 모든 제공사의 1,000개 이상 AI 모델을 API 키 하나로.