Wan 2.2 API
Alibaba の Wan 2.2 は、WaveSpeedAI にデプロイされた、35以上の自社バリアントを持つオープンウェイトの動画ツールキットです。Animate(120秒のキャラクターアニメーション)、Video Edit、Speech-to-Video(10分のオーディオ駆動)、Fun-Control(Apache 2.0ライセンス)に加え、複数のモデルサイズ(5B、A14B)と解像度(480p / 720p)の画像から動画とテキストから動画があります。
WaveSpeedAI でホストされるバリアントのみです。Animate は最長120秒の720pクリップを、Speech-to-Video は最長10分の480pクリップを生成し、Fun-Control は商用利用できる Apache 2.0 のもと、プリセットの Control Codes を使います。LoRA学習のエンドポイントは、数分でファインチューニングできます。
概要
Wan 2.2 APIについて
Wan 2.2でできること、Alibabaのモデルラインナップにおける位置づけ、そして多くのチームに選ばれる理由。
Wan 2.2はAlibabaの動画生成モデルで、WaveSpeedAIのREST APIから利用できます。Alibaba の Wan 2.2 は、WaveSpeedAI にデプロイされた、35以上の自社バリアントを持つオープンウェイトの動画ツールキットです。Animate(120秒のキャラクターアニメーション)、Video Edit、Speech-to-Video(10分のオーディオ駆動)、Fun-Control(Apache 2.0ライセンス)に加え、複数のモデルサイズ(5B、A14B)と解像度(480p / 720p)の画像から動画とテキストから動画があります。
WaveSpeedAI でホストされるバリアントのみです。Animate は最長120秒の720pクリップを、Speech-to-Video は最長10分の480pクリップを生成し、Fun-Control は商用利用できる Apache 2.0 のもと、プリセットの Control Codes を使います。LoRA学習のエンドポイントは、数分でファインチューニングできます。
WaveSpeedAIのWan 2.2ファミリーには、Image-To-Video, Motion-Control, Video-To-Video, Image-To-Image, Digital-Human, Text-To-Image, Training, Text-To-Videoのワークフローをカバーする32個のRESTエンドポイントがあります。各バリアントには、それぞれ独自の料金、パラメーター、作例があります。入力の種類と本番環境の制約に合うものを選ぶか、同じAPIキーで複数を呼び出して、多段階のパイプラインを組み立ててください。
WaveSpeedAIの他の1,000以上のAIモデルと同じAPIキー、請求アカウント、レート制限の枠組みでWan 2.2を実行できます。ベンダーごとの設定も、プロバイダーごとのSDKも、ベンダーごとのレート制限も不要です。1つの連携で、テキストから画像、テキストから動画から、音声合成、3D生成、高画質化、編集まで、すべてカバーできます。
エンドポイント
Wan 2.2のAPIエンドポイント一覧
WaveSpeedAIで現在利用可能なWan 2.2のエンドポイントは32個です。ワークフローに合ったバリアントを選んでください。
/filters:quality(82)/media/images/20260408111216_2fizlz7x.webp)
Wan 2.2 Image To Video Lora
Wan-2.2/image-to-video-lora enables unlimited image-to-video generation from a single image, producing smooth, cinematic motion with clean detail. Supports custom LoRAs for style and character consistency. Ready-to-use REST inference API, best performance, no cold starts, affordable pricing.
/filters:quality(82)/media/images/20260408111152_zehbgma1.webp)
Wan 2.2 Image To Video
Wan 2.2 Image-to-Video turns a single image into smooth, cinematic motion with clean detail—ideal for storyboards, mood shots, and product demos. Ready-to-use REST inference API, best performance, no cold starts, affordable pricing.
/filters:quality(82)/media/images/20260408111211_p8f6ffxe.webp)
Wan 2.2 Animate
Wan2.2-Animate unified character animation & replacement model replicating movement and expression; generates 720p videos up to 120s. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/20260408111241_3bzr5rd1.webp)
Wan 2.2 Video Edit
Wan 2.2 Video Edit lets you modify videos via text prompts (e.g., change clothing or characters). Powered by Wan 2.2, it supports 480p ($0.20/5s) and 720p ($0.40/5s), up to 120s. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/20260408111257_4rmbjsgj.webp)
Wan 2.2 Image To Image
WAN 2.2 (14B) is an image-to-image model for high-resolution photorealistic image editing with exceptional precision and fidelity. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/20260408111154_gvj4f5sj.webp)
Wan 2.2 Speech To Video
Wan-2.2-S2V turns images and speech into high-fidelity videos with realistic face and body motion; supports up to 10-minute clips in 480p, from $0.15/5s. Ready-to-use REST API, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/20260408111206_p4ked9pu.webp)
Wan 2.2 Text To Image Lora
WAN 2.2 generates super-detailed images from text prompts and supports custom LoRAs for fine-grained style and subject control. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/20260408111250_80jkk4zs.webp)
Wan 2.2 Fun Control
Wan2.2-Fun-Control uses Control Codes and multi-modal inputs to generate preset-controlled videos up to 120s at 720p; released under Apache 2.0 for commercial use. Ready-to-use REST API, no coldstarts, affordable.
/filters:quality(82)/media/images/20260408112009_zkynaf2n.webp)
Wan 2.2 Image Lora Trainer
Train custom Wan 2.2 character/style LoRA models 10x faster. Style training, character training, object training. From concept to model in minutes, not hours. Upload a ZIP file containing images to start!
/filters:quality(82)/media/images/20260408111954_4jny6q8q.webp)
Wan 2.2 I2v Lora Trainer
Train custom Wan 2.2 I2V LoRA models 10x faster. Action training, motion training, video efect training. From concept to model in minutes, not hours. Upload a ZIP file containing videos to start!
/filters:quality(82)/media/images/20260408111226_b3370qq2.webp)
Wan 2.2 Text To Image Realism
WAN 2.2 delivers ultra-realistic text-to-image generation, converting prompts into photoreal images with high fidelity and detail. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/20260408111246_kbtz9m3g.webp)
Wan 2.2 I2v 5b 720p Lora
Wan 2.2 i2v-5B-720p is a 5B image-to-video model producing 720p videos with LoRA support for style customization. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/20260408111210_znuyn6r8.webp)
Wan 2.2 T2v 5b 720p Lora
Wan 2.2 T2V 5B is a 5B text-to-video model with LoRA support that generates 720p videos from text prompts for easy personalization. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/20260408111236_0was1j4o.webp)
Wan 2.2 I2v 480p Lora Ultra Fast
Wan 2.2 i2v delivers ultra-fast Image-to-Video at 480p with support for custom LoRAs for tailored styles. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/20260408111252_sejz6k48.webp)
Wan 2.2 I2v 480p Ultra Fast
Wan 2.2 A14B Image-to-Video (i2v-480p) produces ultra-fast 480p videos from single images, enabling unlimited AI video generation with high throughput. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/20260408111149_z1lz7r9s.webp)
Wan 2.2 I2v 720p Ultra Fast
Generate unlimited ultra-fast 720p AI videos from images with Wan 2.2 A14B image-to-video model. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/20260408111259_8bca6477.webp)
Wan 2.2 T2v 480p Lora Ultra Fast
Ultra-fast Wan 2.2 text-to-video model producing 480p videos with custom LoRA support—generate unlimited AI videos with personalized styles. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/20260408111238_h7qxvir0.webp)
Wan 2.2 I2v 720p Lora Ultra Fast
Wan 2.2 i2v 720P is an ultra-fast Image-to-Video model that generates unlimited AI videos and supports custom LoRAs for personalized outputs. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/20260408111248_iyye47z8.webp)
Wan 2.2 T2v 480p Ultra Fast
Wan 2.2 t2v 480p Ultra-Fast generates unlimited AI videos from text prompts at 480p with ultra-fast inference. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/20260408111201_lv1td92d.webp)
Wan 2.2 T2v 5b 720p
Wan 2.2 T2V 5B is a 720P text-to-video model that generates unlimited AI videos from simple text prompts, producing consistent high-quality 720p outputs. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/20260408111221_xo8hpiqi.webp)
Wan 2.2 I2v 5b 720p
Wan 2.2 I2V 5B converts images into high-quality 720P videos using a 5B image-to-video model for AI video generation. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/20260408111214_rct2hj0c.webp)
Wan 2.2 I2v 480p
Wan 2.2 A14B converts images into 480p videos, enabling unlimited AI video generation from single images. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/20260408111234_cplw518o.webp)
Wan 2.2 I2v 480p Lora
WAN 2.2 A14B Image-to-Video model generates unlimited 480p videos from images and supports custom LoRAs for personalized styles and fine-tuning. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/20260408111232_8tgp9v7m.webp)
Wan 2.2 I2v 720p Lora
WAN 2.2 Image-to-Video (i2v) 720p converts images into 720p videos and supports custom LoRAs for style personalization. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/20260408111239_qlur3dn5.webp)
Wan 2.2 I2v 720p
WAN 2.2 A14B i2v-720p converts images into smooth 720p videos, enabling unlimited AI video generation with the Wan 2.2 image-to-video model. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/20260408111156_b3olejjy.webp)
Wan 2.2 T2v 480p
Wan 2.2 t2v-480p generates unlimited AI videos from text prompts at 480p resolution, ideal for rapid prototyping and content creation. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/20260408111208_sw0j3xam.webp)
Wan 2.2 T2v 480p Lora
WAN 2.2 T2V 480p with LoRA generates text-to-video at 480p and supports custom LoRAs for personalized styles. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/20260408111218_00vzkvbl.webp)
Wan 2.2 T2v 720p
Wan 2.2 t2v-720p converts text prompts into native 720P videos, producing high-quality 720P clips from simple prompts. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/20260408111230_mtxgxh1x.webp)
Wan 2.2 T2v 720p Lora
Wan 2.2 T2V 720p with custom LoRA support turns text prompts into 720p AI videos and enables unlimited video generation. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/20260408111158_csf9651o.webp)
Wan 2.2 T2v 720p Lora Ultra Fast
Ultra-fast Wan 2.2 Text-to-Video generates unlimited 720p AI videos with custom LoRAs for personalized styles. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/20260408111244_aqp0x6b0.webp)
Wan 2.2 T2v 720p Ultra Fast
WAN 2.2 T2V 720p Ultra-Fast generates high-quality 720p videos from text prompts with unlimited output and ultra-fast throughput. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/1786269272686834557_4oOX6gqz.webp)
Wan 2.2 Animate 2
Wan 2.2 Animate 2 is the next-generation Wan character animation model: an end-to-end DiT that makes the character in a reference image perform the motion of a driving video, with no pose extraction, prompt-controlled background, and strong identity preservation; generates 480p/720p videos up to 120s. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
作例
Wan 2.2の実力を見る
Wan 2.2 APIで実際に生成された出力です。動画にカーソルを合わせるとプレビュー、クリックすると原寸のビューアで開きます。
使い方
Wan 2.2 APIの使い方
登録から生成完了まで4ステップ。Python、Node.js、cURLの完全なサンプルは、下のAPIセクションにあります。
- 01
APIキーを取得
WaveSpeedAIのアカウントに登録し、ダッシュボードからAPIキーをコピーします。新規アカウントには無料のスタータークレジットが付くので、課金が始まる前にプレイグラウンドを数十回実行できます。
- 02
予測を送信
入力をJSONとしてhttps://api.wavespeed.ai/api/v3/wavespeed-ai/wan-2.2/image-to-videoにPOSTします。エンドポイントはすぐに予測IDを返します。生成は非同期なので、推論中に接続を開いたままにする必要はありません。
- 03
完了までポーリング
https://api.wavespeed.ai/api/v3/predictions/{request_id}/resultにGETします。completedなら出力を返し、failed、cancelled、timeout、deletedならエラーで停止し、それ以外のステータスの間はポーリングを続けます。
- 04
出力URLを読み取る
ステータスが"completed"になったら、data.outputs[0]からURLを読み取ります。URLは、WaveSpeedAIのCDN上にある生成されたメディアを指します。呼び出したWan 2.2のバリアントに応じて、画像、動画、音声、3Dファイルのいずれかです。
活用例
Wan 2.2で作れるもの
開発者やクリエイターがWan 2.2 APIでよく使うワークフロー。
Wan 2.2 Animate — 最長120秒のキャラクターアニメーション
wavespeed-ai/wan-2.2/animate は、「動きと表情を再現する、統合型のキャラクターアニメーション・置き換えモデルで、最長120秒の720p動画を生成」します。これはカタログ上の特長です。ポーズ駆動のほとんどのアニメーションツールよりはるかに長くなっています。
最長10分の Speech-to-Video
wavespeed-ai/wan-2.2/speech-to-video は、「画像と音声を、リアルな顔と体の動きを備えた高精細な動画に変換し、480pで最長10分のクリップに対応」します。長尺のトーク系コンテンツに便利です。
プロンプト駆動の変更ができる Video Edit
wavespeed-ai/wan-2.2/video-edit は、テキストプロンプトで動画を変更できます(カタログの例では、服装やキャラクターの変更)。480pと720p、最長120秒に対応しています。
Apache 2.0 ライセンスの Fun-Control
wavespeed-ai/wan-2.2/fun-control は、「Control Codes とマルチモーダル入力で、720pで最長120秒のプリセット制御の動画を生成し、商用利用できる Apache 2.0 で公開」されています。Apache 2.0 というライセンスは、商用パイプラインにとって大きな差別化ポイントです。
LoRA学習(10倍高速)
画像LoRAには wavespeed-ai/wan-2.2-image-lora-trainer、I2V LoRAには wavespeed-ai/wan-2.2-i2v-lora-trainer を使います。カタログ上の特長は「10倍高速な学習」です。スタイル、キャラクター、オブジェクト、モーション、アクション、動画エフェクトの学習に対応しています。ZIPファイルをアップロードして始めます。
複数サイズの画像から動画(5B / A14B)
モデルサイズを選べます。速度とコストなら5B(より小さい)、フル品質ならA14Bです。480p向けの標準的な i2v、LoRA対応のバリアント、超高速バリアントがあります。
コツ
Wan 2.2のプロンプトのコツ
Wan 2.2からより良い出力を得るための実践的なアドバイス。本番のパイプラインで動画モデル全般に通用するパターンに基づいています。
- 01
タスクに合ったバリアントを選ぶ
Wan 2.2 は、1つの汎用モデルではなく、専用のエンドポイントを提供します。ポーズ駆動のモーションには Animate、狙った変更には video-edit、トーク系のコンテンツには speech-to-video、汎用の生成には image-to-video です。タスクに合ったエンドポイントを選んでください。汎用モデルに何でもさせるより、はるかに良い出力が得られます。
- 02
制作規模での一貫性のために LoRA を学習する
Wan 2.2 の LoRA トレーナーのエンドポイントは、補助的なツールではなく、APIの主要な機能です。数百回の生成にわたって同じアイデンティティが必要な制作(ブランドのマスコット、繰り返し登場するキャラクター、独自のスタイル)では、LoRA を一度学習し、生成ごとに LoRA 推論のエンドポイントを呼び出します。
- 03
オープンウェイトならファインチューニングが可能
クローズドウェイトの競合は、ベンダーのモデルの挙動に縛られます。Wan 2.2 の根幹はオープンウェイトなので、ベースモデルでは解決できない制約に突き当たっても、ファインチューニングという選択肢があります。他のほとんどの商用の動画APIにはない道です。
- 04
Animate と Kling Motion Control を組み合わせる
どちらもポーズ駆動のアニメーションを提供しますが、トレードオフは異なります。Wan 2.2 Animate は LoRA に対応し、オープンウェイトです。Kling Motion Control は、アイデンティティの保持が優れています。ファインチューニングが重要か、アイデンティティの品質が重要かで選んでください。
料金
Wan 2.2 APIの料金
料金は出力ごとです。最終的な請求額は、各バリアントのプレイグラウンドで設定するパラメーター(解像度、長さ、出力数、参照入力)に応じて変わります。
API
Wan 2.2 APIを呼び出す
wavespeed.ai/accesskeyでAPIキーに登録し、RESTで予測を送信します。プレイグラウンドでは、入力の組み合わせに応じて、そのまま貼り付けられるサンプルが生成されます。
POSThttps://api.wavespeed.ai/api/v3/wavespeed-ai/wan-2.2/image-to-video
# 1. Submit the prediction.
SUBMIT_RESPONSE=$(curl --silent --show-error --fail-with-body \
-X POST "https://api.wavespeed.ai/api/v3/wavespeed-ai/wan-2.2/image-to-video" \
-H "Content-Type: application/json" \
-H "Authorization: Bearer $WAVESPEED_API_KEY" \
-d '{
"prompt": "A cinematic shot of a city at sunset, soft golden light",
"image": "https://interactive-examples.mdn.mozilla.net/media/cc0-images/painted-hand-298-332.jpg",
"resolution": "480p",
"duration": 5
}')
TASK=$(printf '%s' "$SUBMIT_RESPONSE" | jq 'if has("data") then .data else . end')
PREDICTION_ID=$(printf '%s' "$TASK" | jq -r '.id')
if [ -z "$PREDICTION_ID" ] || [ "$PREDICTION_ID" = "null" ]; then
printf 'Submission response did not contain a prediction id
' >&2
exit 1
fi
RESULT_URL="https://api.wavespeed.ai/api/v3/predictions/$PREDICTION_ID/result"
# 2. Poll until the prediction finishes.
while true; do
RESPONSE=$(curl --silent --show-error --fail-with-body "$RESULT_URL" \
-H "Authorization: Bearer $WAVESPEED_API_KEY")
RESULT=$(printf '%s' "$RESPONSE" | jq 'if has("data") then .data else . end')
STATUS=$(printf '%s' "$RESULT" | jq -r '.status')
case "$STATUS" in
completed) printf '%s\n' "$RESULT" | jq '.outputs'; break ;;
failed|cancelled|timeout|deleted) printf '%s\n' "$RESULT" | jq . >&2; exit 1 ;;
*) sleep 2 ;;
esac
doneconst submitUrl = "https://api.wavespeed.ai/api/v3/wavespeed-ai/wan-2.2/image-to-video";
const apiKey = process.env.WAVESPEED_API_KEY;
if (!apiKey) throw new Error('Set WAVESPEED_API_KEY');
async function requestJson(url, options = {}) {
const response = await fetch(url, options);
if (!response.ok) throw new Error(await response.text());
return response.json();
}
// 1. Submit the prediction.
const body = await requestJson(submitUrl, {
method: "POST",
headers: {
"Authorization": `Bearer ${apiKey}`,
"Content-Type": "application/json",
},
body: JSON.stringify({
"prompt": "A cinematic shot of a city at sunset, soft golden light",
"image": "https://interactive-examples.mdn.mozilla.net/media/cc0-images/painted-hand-298-332.jpg",
"resolution": "480p",
"duration": 5
}),
});
const task = body.data ?? body;
const resultUrl = `https://api.wavespeed.ai/api/v3/predictions/${task.id}/result`;
// 2. Poll until the prediction finishes.
while (true) {
const resultBody = await requestJson(resultUrl, {
headers: { "Authorization": `Bearer ${apiKey}` },
});
const result = resultBody.data ?? resultBody;
if (result.status === "completed") {
console.log(result.outputs);
break;
}
if (["failed", "cancelled", "timeout", "deleted"].includes(result.status)) throw new Error(JSON.stringify(result));
await new Promise(resolve => setTimeout(resolve, 2000));
}import json
import os
import time
from urllib.request import Request, urlopen
api_key = os.environ["WAVESPEED_API_KEY"]
headers = {"Authorization": f"Bearer {api_key}", "Content-Type": "application/json"}
payload = {
"prompt": "A cinematic shot of a city at sunset, soft golden light",
"image": "https://interactive-examples.mdn.mozilla.net/media/cc0-images/painted-hand-298-332.jpg",
"resolution": "480p",
"duration": 5
}
def request_json(url, data=None):
request = Request(url, data=data, headers=headers, method="POST" if data else "GET")
with urlopen(request) as response:
return json.load(response)
# 1. Submit the prediction.
body = request_json("https://api.wavespeed.ai/api/v3/wavespeed-ai/wan-2.2/image-to-video", json.dumps(payload).encode())
task = body.get("data", body)
result_url = f"https://api.wavespeed.ai/api/v3/predictions/{task['id']}/result"
# 2. Poll until the prediction finishes.
while True:
result_body = request_json(result_url)
result = result_body.get("data", result_body)
status = result.get("status")
if status == "completed":
print(result.get("outputs", []))
break
if status in {"failed", "cancelled", "timeout", "deleted"}:
raise RuntimeError(result)
time.sleep(2)比較
Wan 2.2と他モデルの比較
WaveSpeedAI上の類似モデルではなくWan 2.2を選ぶべきケース。
Wan 2.2 vs Wan 2.7
Wan 2.7(alibaba/wan-2.7/*)は、reference-to-video、video-edit、image-edit、text-to-image を1つのファミリーにまとめた、Alibaba の新しいアーキテクチャで、モダリティをまたぐツールが幅広くなっています。Wan 2.2(WaveSpeedAI のバリアント)には、2.7 にはない専用エンドポイント(Animate(120秒)、Speech-to-Video(10分)、Fun-Control(Apache 2.0)、LoRAトレーナー)があります。
Wan 2.2 vs Seedance 2.0
Seedance 2.0 は、全バリアントでネイティブオーディオを備えた、ハリウッド品質の出力です。Wan 2.2 は、バリアントの幅(35以上のエンドポイント)と LoRA 学習で優位に立ちます。汎用の動画モデルではなく、特定の専用機能(Animate、Speech-to-Video、Fun-Control)が必要なときに適した選択です。
Wan 2.2 vs Kling 3.0 Motion Control
どちらもポーズ駆動のキャラクターアニメーションを提供します。Wan 2.2 Animate は120秒の720pクリップを生成でき、LoRAのファインチューニングも備えています。Kling Motion Control は参照動画の長さ(3〜30秒)に制約されますが、Kuaishou の動画コーパスで学習されており、モーションの事前知識が強力です。
FAQ
Wan 2.2 API — よくある質問
料金、ライセンス、連携など、WaveSpeedAIでWan 2.2を実行する際のよくある質問。
Wan 2.2 APIとは何ですか?
Wan 2.2は、Alibabaの動画生成モデルで、WaveSpeedAI上でREST APIとして提供されています。Alibaba の Wan 2.2 は、WaveSpeedAI にデプロイされた、35以上の自社バリアントを持つオープンウェイトの動画ツールキットです。Animate(120秒のキャラクターアニメーション)、Video Edit、Speech-to-Video(10分のオーディオ駆動)、Fun-Control(Apache 2.0ライセンス)に加え、複数のモデルサイズ(5B、A14B)と解像度(480p / 720p)の画像から動画とテキストから動画があります。プログラムから呼び出すことも、上にリンクされたプレイグラウンドで試すこともできます。
Wan 2.2 APIはどう呼び出しますか?
WaveSpeedAIのアカウントに登録し、/accesskeyからAPIキーをコピーして、入力をJSONとしてhttps://api.wavespeed.ai/api/v3/wavespeed-ai/wan-2.2/image-to-videoにPOSTします。エンドポイントは予測IDを返します。結果のエンドポイントを約2秒ごとにポーリングし、長時間かかるタスクでは間隔を広げ、終端ステータスになったら停止してください。本番運用向けのPython / Node.js / cURLのサンプルは上にあります。
Wan 2.2 APIの料金はいくらですか?
Wan 2.2は1回あたり$0.02からです。実際の料金は、設定するパラメーター(解像度、長さ、出力数、参照入力)に応じて変わります。プレイグラウンドの「生成」ボタンの横に表示されるリアルタイムの料金プレビューで、現在の入力での正確な料金を確認できます。
Wan 2.2にはどのバリアントがありますか?
WaveSpeedAIでは、32個のWan 2.2エンドポイントが利用可能です:wavespeed-ai/wan-2.2/image-to-video-lora, wavespeed-ai/wan-2.2/image-to-video, wavespeed-ai/wan-2.2/animate, wavespeed-ai/wan-2.2/video-edit, wavespeed-ai/wan-2.2/image-to-image, wavespeed-ai/wan-2.2/speech-to-video, wavespeed-ai/wan-2.2/text-to-image-lora, wavespeed-ai/wan-2.2/fun-controlほか。各バリアントには、専用のプレイグラウンドページと料金があります。
Wan 2.2の出力を商用利用できますか?
商用利用の権利は、Alibabaのモデルライセンスに従います。Alibabaのほとんどのモデルは出力の商用利用を認めています。具体的なライセンスの概要は各モデルのプレイグラウンドページで、プラットフォーム全体の条件はWaveSpeedAIの利用規約でご確認ください。
なぜ直接ではなくWaveSpeedAIでWan 2.2を使うのですか?
Wan 2.2と、他のプロバイダーの1,000以上のAIモデルを、1つのAPIキー、1つの請求アカウントで利用できます。ベンダーごとのSDK設定も、個別のレート制限も、ベンダーごとの連携コードの書き直しも不要です。料金は、通常Alibabaの直接のAPIと同等かそれ以下です。
提供元
Alibabaについて
WaveSpeedAIにおけるWan 2.2と、Alibabaのモデルラインナップ全体を手がけるチーム。
Alibaba の Tongyi Lab は、動画モデルの Wan ファミリーと、LLM の Qwen ファミリーを開発しています。Wan は、オープンウェイトで公開されていること、幅広いバリアント(テキストから動画、画像から動画、リファレンスから動画、video-edit、video-extend、image-edit、text-to-image)を網羅していること、多言語のプロンプトにわたってモーションの安定性とプロンプトへの追従性に一貫して強みがあることが特長です。
WaveSpeedAIでWan 2.2を使った開発を始めよう
登録時に無料のスタータークレジットを進呈。Alibabaをはじめ、あらゆるプロバイダーの1,000以上のAIモデルを、1つのAPIキーで。