Vidu Q3 API
生數科技 Vidu Q3 提供文字轉影片、圖片轉影片、參考轉影片(1 至 4 張參考圖片,維持多主體一致性)與 start-end-to-video。分為 Standard、Pro、Turbo 三個等級,部分版本最長可輸出 16 秒。
提供 Standard、Pro(1 至 16 秒輸出)與 Turbo(更快)等級。參考轉影片可接受 1 至 4 張參考圖片,生成多主體一致的影片,解析度 360p 至 1080p、最長 16 秒。Start-end-to-video 可銜接兩張關鍵幀(Pro 為 1 至 16 秒)。image-to-video-pro 版本支援 720p/1080p/2K/4K。
概覽
關於 Vidu Q3 API
Vidu Q3 的功能、它在 Shengshu 模型陣容中的定位,以及團隊選用它的原因。
Vidu Q3 是來自 Shengshu 的影片生成模型,可透過 WaveSpeedAI REST API 使用。生數科技 Vidu Q3 提供文字轉影片、圖片轉影片、參考轉影片(1 至 4 張參考圖片,維持多主體一致性)與 start-end-to-video。分為 Standard、Pro、Turbo 三個等級,部分版本最長可輸出 16 秒。
提供 Standard、Pro(1 至 16 秒輸出)與 Turbo(更快)等級。參考轉影片可接受 1 至 4 張參考圖片,生成多主體一致的影片,解析度 360p 至 1080p、最長 16 秒。Start-end-to-video 可銜接兩張關鍵幀(Pro 為 1 至 16 秒)。image-to-video-pro 版本支援 720p/1080p/2K/4K。
WaveSpeedAI 上的 Vidu Q3 系列提供 13 個 REST 端點,涵蓋 Text-To-Video, Image-To-Video, Reference-To-Video 等工作流程。每個變體都有各自的定價、參數選項與範例輸出 — 請挑選符合您輸入模態與生產限制的版本,或使用同一組 API 金鑰呼叫多個變體,組合成多步驟的處理流程。
使用您呼叫其他 1,000+ 個 WaveSpeedAI AI 模型時相同的 API 金鑰、帳單帳戶與速率限制額度來執行 Vidu Q3。無需另外設定供應商、無需各家 SDK、無需處理各家不同的速率限制 — 一次整合,涵蓋從文字轉圖片、文字轉影片,到音訊合成、3D 生成、放大與編輯的所有功能。
端點
所有 Vidu Q3 API 端點
WaveSpeedAI 目前提供 13 個 Vidu Q3 端點 — 請選擇符合您工作流程的變體。
/filters:quality(82)/media/images/1778682596980020194_WE2bktCL.webp)
Q3 Pro Text To Video
Vidu Q3 Pro Text to Video is a fast AI video generation model that creates high-quality, audio-capable videos from text prompts with support for 1–16 second outputs. Ready-to-use REST inference API for cinematic clips, advertising creatives, social media videos, product visuals, storytelling, and professional text-to-video workflows with simple integration, no coldstarts, and affordable pricing.
/filters:quality(82)/media/images/1778642810217109913_lPenxFPY.webp)
Q3 Pro Start End To Video
Vidu Q3 Pro Start-End-to-Video creates smooth transitions between two keyframes with viduq3-pro (1–16s). Billing follows Vidu's published Q3-pro per-second rates by resolution. Ready-to-use REST inference API on WaveSpeed.
/filters:quality(82)/media/images/20260408105234_f2z15q17.webp)
Q3 Turbo Start End To Video
Vidu Q3 Turbo Start-End-to-Video creates smooth transitions between two images with faster processing. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/20260408105138_ppp4gly8.webp)
Q3 Start End To Video
Vidu Q3 Start End Image-to-Video turns text prompts into high-quality videos with exceptional visual fidelity and diverse motion. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/1783683828943385893_6nuh5TGv.webp)
Q3 Reference To Video
Vidu Q3 Reference-to-Video Mix generates multi-entity consistent videos from 1-4 reference images with text prompt guidance. Supports 360p to 1080p resolutions, up to 16 seconds duration, multiple aspect ratios, and optional audio generation. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/20260408105132_z9o7hzdv.webp)
Q3 Image To Video Pro
Vidu Q3 Image-to-Video Pro generates high-resolution videos (720p/1080p/2K/4K) from images with exceptional visual fidelity and diverse motion. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/1778642590678181581_rTjsBLV5.webp)
Q3 Pro Image To Video
Vidu Q3 Pro Image-to-Video animates still images with high-quality motion via viduq3-pro (1–16s). Billing follows Vidu's published Q3-pro per-second rates by resolution. Ready-to-use REST inference API on WaveSpeed.
/filters:quality(82)/media/images/20260408105229_yjachyfs.webp)
Q3 Turbo Image To Video
Vidu Q3 Turbo Image-to-Video animates static images with high-quality motion and faster processing. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/20260408105141_ud2162ut.webp)
Q3 Text To Video
Vidu Q3 Text-to-Video turns text prompts into high-quality videos with exceptional visual fidelity and diverse motion. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/1782498276203721773_heCfpzIR.webp)
Q3 Drama
Vidu Q3 Drama generates complete script-driven drama videos from scripts and structured assets, including characters, scenes, tools, and references. It plans the narrative structure, scene pacing, and transitions to create a story-driven drama in one request, supporting up to 180 seconds. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/20260408105128_fw6o586f.webp)
Q3 Image To Video
Vidu Q3 Image-to-Video turns text prompts into high-quality videos with exceptional visual fidelity and diverse motion. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/1782880699830272243_SRdcBUgG.webp)
Q3 Ad
Vidu Q3 Ad Video generates commercial ad videos from 1 to 7 reference images with prompt guidance, supporting 720P / 1080P output and synchronized audio for product ads, brand campaigns, marketing creatives, and promotional videos. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/1782715804387581475_DcHQ19js.webp)
Q3 Drama Clip
Vidu Q3 Drama Clip generates 8-12 second script-driven drama videos from structured assets, including characters, scenes, and tools. It is ideal for compact story scenes, storyboard shots, and focused narrative moments. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
範例
看看 Vidu Q3 的實際效果
由 Vidu Q3 API 生成的真實輸出。將滑鼠移到任一影片上即可預覽,點擊可開啟完整尺寸檢視器。
使用方式
如何使用 Vidu Q3 API
從註冊到完成生成,只要四個步驟。完整的 Python、Node.js 與 cURL 範例請見下方的 API 區段。
- 01
取得 API 金鑰
註冊 WaveSpeedAI 帳號,並從控制台複製您的 API 金鑰。新帳號會獲贈免費入門點數 — 足以在開始計費前執行數十次 Playground。
- 02
提交 prediction
將您的輸入以 JSON 格式 POST 到 https://api.wavespeed.ai/api/v3/vidu/q3/text-to-video。端點會立即回傳 prediction id — 生成是非同步進行的,因此推論期間您不需要保持連線。
- 03
輪詢完成狀態
GET https://api.wavespeed.ai/api/v3/predictions/{request_id}/result。狀態為 completed 時回傳輸出;狀態為 failed、cancelled、timeout 或 deleted 時停止並回報錯誤;其他狀態則持續輪詢。
- 04
讀取輸出網址
當狀態為 "completed" 時,從 data.outputs[0] 讀取網址。該網址指向 WaveSpeedAI CDN 上您生成的媒體 — 依您呼叫的 Vidu Q3 變體,可能是圖片、影片、音訊或 3D 檔案。
應用情境
您可以用 Vidu Q3 打造什麼
開發者與創作者使用 Vidu Q3 API 的常見工作流程。
支援 1 至 4 張參考的參考轉影片
vidu/q3/reference-to-video 可依文字提示詞引導,從 1 至 4 張參考圖片生成多主體一致的影片,支援 360p 至 1080p、最長 16 秒與多種長寬比。適合需要讓多個參考主體保持連貫的群像場景。
首尾關鍵幀插值
vidu/q3/start-end-to-video 在兩張關鍵幀之間插值生成影片。Pro 與 Turbo 等級同樣支援 1 至 16 秒的首尾幀生成。適合已有起始與結束靜態圖的動態分鏡腳本。
最長 16 秒的輸出
目錄說明,Pro 與參考轉影片版本支援 1 至 16 秒的時長,比許多競品影片模型 5 至 8 秒的範圍更長,適合在單次生成中完成完整的敘事段落。
高解析度圖片轉影片
vidu/q3/image-to-video-pro 支援由圖片生成 720p / 1080p / 2K / 4K 的影片。從靜態圖出發時,無須放大處理即可獲得交付等級的輸出。
Pro 等級(最便宜)
vidu/q3-pro/* 是 Vidu Q3 中最便宜的等級,適合大量作業。涵蓋圖片轉影片、文字轉影片與 start-end-to-video,輸出 1 至 16 秒。
技巧
Vidu Q3 提示詞技巧
讓 Vidu Q3 產出更好結果的實用建議 — 整理自在實際生產流程中,影片模型通用的有效做法。
- 01
以 1 至 4 張參考圖片支援多主體場景
vidu/q3/reference-to-video 接受 1 至 4 張參考圖片,並依文字提示詞引導,生成多主體一致的影片。適合群像演員場景、團體產品展示與多主體分鏡。
- 02
為關鍵幀工作流程使用首尾幀插值
vidu/q3/start-end-to-video 以生成的動態銜接兩張靜態圖。特別適合動態分鏡、以關鍵姿勢驅動的分鏡腳本,以及將概念圖串成動態,而無須為每個片段重新撰寫提示詞。
- 03
有意識地選擇等級
Standard 是預設的交付等級;Pro 等級(1 至 16 秒輸出)定位為大量作業;Turbo 等級則以速度優先。請查看本頁的最新價格表了解各等級的目前成本,Vidu Q3 的 Pro 等級價格特別具競爭力。
- 04
部分版本最長可輸出 16 秒
目錄說明,Pro 與參考轉影片版本支援 1 至 16 秒的時長。比許多競品影片模型 5 至 8 秒的範圍更長,適合在單次生成中完成完整的敘事段落。
- 05
image-to-video-pro 用於高解析度輸出
vidu/q3/image-to-video-pro 支援由圖片生成 720p / 1080p / 2K / 4K 解析度的影片。從靜態圖出發時,無須放大處理即可獲得交付等級的輸出。
定價
Vidu Q3 API 定價
依輸出計費。最終費用會依您在各變體 Playground 中設定的參數(解析度、時長、輸出數量、參考素材)而調整。
| 端點 | 類型 | 起始價格 |
|---|---|---|
| vidu/ | text-to-video | $0.25 |
| vidu/ | image-to-video | $0.25 |
| vidu/ | image-to-video | $0.30 |
| vidu/ | image-to-video | $0.35 |
| vidu/ | reference-to-video | $0.35 |
| vidu/ | image-to-video | $0.45 |
| vidu/ | image-to-video | $0.25 |
| vidu/ | image-to-video | $0.30 |
| vidu/ | text-to-video | $0.35 |
| vidu/ | image-to-video | $1.12 |
| vidu/ | image-to-video | $0.35 |
| vidu/ | image-to-video | $0.15 |
| vidu/ | image-to-video | $1.12 |
API
呼叫 Vidu Q3 API
請到 wavespeed.ai/accesskey 註冊取得 API 金鑰,然後透過 REST 提交 prediction。Playground 會為任何輸入組合產生可直接貼上的範例。
POSThttps://api.wavespeed.ai/api/v3/vidu/q3/text-to-video
# 1. Submit the prediction.
SUBMIT_RESPONSE=$(curl --silent --show-error --fail-with-body \
-X POST "https://api.wavespeed.ai/api/v3/vidu/q3/text-to-video" \
-H "Content-Type: application/json" \
-H "Authorization: Bearer $WAVESPEED_API_KEY" \
-d '{
"prompt": "A cinematic shot of a city at sunset, soft golden light",
"style": "general",
"resolution": "720p",
"duration": 5,
"aspect_ratio": "4:3",
"movement_amplitude": "auto",
"generate_audio": true,
"bgm": true
}')
TASK=$(printf '%s' "$SUBMIT_RESPONSE" | jq 'if has("data") then .data else . end')
PREDICTION_ID=$(printf '%s' "$TASK" | jq -r '.id')
if [ -z "$PREDICTION_ID" ] || [ "$PREDICTION_ID" = "null" ]; then
printf 'Submission response did not contain a prediction id
' >&2
exit 1
fi
RESULT_URL="https://api.wavespeed.ai/api/v3/predictions/$PREDICTION_ID/result"
# 2. Poll until the prediction finishes.
while true; do
RESPONSE=$(curl --silent --show-error --fail-with-body "$RESULT_URL" \
-H "Authorization: Bearer $WAVESPEED_API_KEY")
RESULT=$(printf '%s' "$RESPONSE" | jq 'if has("data") then .data else . end')
STATUS=$(printf '%s' "$RESULT" | jq -r '.status')
case "$STATUS" in
completed) printf '%s\n' "$RESULT" | jq '.outputs'; break ;;
failed|cancelled|timeout|deleted) printf '%s\n' "$RESULT" | jq . >&2; exit 1 ;;
*) sleep 2 ;;
esac
doneconst submitUrl = "https://api.wavespeed.ai/api/v3/vidu/q3/text-to-video";
const apiKey = process.env.WAVESPEED_API_KEY;
if (!apiKey) throw new Error('Set WAVESPEED_API_KEY');
async function requestJson(url, options = {}) {
const response = await fetch(url, options);
if (!response.ok) throw new Error(await response.text());
return response.json();
}
// 1. Submit the prediction.
const body = await requestJson(submitUrl, {
method: "POST",
headers: {
"Authorization": `Bearer ${apiKey}`,
"Content-Type": "application/json",
},
body: JSON.stringify({
"prompt": "A cinematic shot of a city at sunset, soft golden light",
"style": "general",
"resolution": "720p",
"duration": 5,
"aspect_ratio": "4:3",
"movement_amplitude": "auto",
"generate_audio": true,
"bgm": true
}),
});
const task = body.data ?? body;
const resultUrl = `https://api.wavespeed.ai/api/v3/predictions/${task.id}/result`;
// 2. Poll until the prediction finishes.
while (true) {
const resultBody = await requestJson(resultUrl, {
headers: { "Authorization": `Bearer ${apiKey}` },
});
const result = resultBody.data ?? resultBody;
if (result.status === "completed") {
console.log(result.outputs);
break;
}
if (["failed", "cancelled", "timeout", "deleted"].includes(result.status)) throw new Error(JSON.stringify(result));
await new Promise(resolve => setTimeout(resolve, 2000));
}import json
import os
import time
from urllib.request import Request, urlopen
api_key = os.environ["WAVESPEED_API_KEY"]
headers = {"Authorization": f"Bearer {api_key}", "Content-Type": "application/json"}
payload = {
"prompt": "A cinematic shot of a city at sunset, soft golden light",
"style": "general",
"resolution": "720p",
"duration": 5,
"aspect_ratio": "4:3",
"movement_amplitude": "auto",
"generate_audio": True,
"bgm": True
}
def request_json(url, data=None):
request = Request(url, data=data, headers=headers, method="POST" if data else "GET")
with urlopen(request) as response:
return json.load(response)
# 1. Submit the prediction.
body = request_json("https://api.wavespeed.ai/api/v3/vidu/q3/text-to-video", json.dumps(payload).encode())
task = body.get("data", body)
result_url = f"https://api.wavespeed.ai/api/v3/predictions/{task['id']}/result"
# 2. Poll until the prediction finishes.
while True:
result_body = request_json(result_url)
result = result_body.get("data", result_body)
status = result.get("status")
if status == "completed":
print(result.get("outputs", []))
break
if status in {"failed", "cancelled", "timeout", "deleted"}:
raise RuntimeError(result)
time.sleep(2)比較
Vidu Q3 與其他選擇比較
在 WaveSpeedAI 上,何時該選擇 Vidu Q3 而非類似的模型。
Vidu Q3 vs Seedance 2.0
Seedance 2.0 在所有版本皆具原生音訊合成,並有 Turbo 等級。Vidu Q3 價格明顯更低,並將首尾幀插值列為 Seedance 所沒有的核心端點。
Vidu Q3 vs Kling 3.0
Kling 3.0 涵蓋 Standard、Pro 與 4K,並以動作控制作為子端點。Vidu Q3 在多數等級較便宜,並將 start-end-to-video 與參考轉影片(1 至 4 張參考)列為核心版本。
Vidu Q3 vs Wan 2.7
Wan 2.7 在同一系列提供參考轉影片、video-edit、video-extend、image-edit 與文字轉圖片。Vidu Q3 則專注於影片生成,等級更便宜,並提供首尾幀插值工作流程。
常見問題
Vidu Q3 API — 常見問題
定價、授權、整合 — 關於在 WaveSpeedAI 上執行 Vidu Q3 的常見問題。
什麼是 Vidu Q3 API?
Vidu Q3 是 Shengshu 的影片生成模型,在 WaveSpeedAI 上以 REST API 提供。生數科技 Vidu Q3 提供文字轉影片、圖片轉影片、參考轉影片(1 至 4 張參考圖片,維持多主體一致性)與 start-end-to-video。分為 Standard、Pro、Turbo 三個等級,部分版本最長可輸出 16 秒。您可以透過程式呼叫,也可以在上方連結的 Playground 試用。
如何呼叫 Vidu Q3 API?
註冊 WaveSpeedAI 帳號,從 /accesskey 複製您的 API 金鑰,然後將您的輸入以 JSON 格式 POST 到 https://api.wavespeed.ai/api/v3/vidu/q3/text-to-video。端點會回傳 prediction id。請從約每 2 秒輪詢一次結果端點開始,長時間任務可拉長間隔,並在任何終止狀態時停止。上方提供適用於生產環境的 Python / Node.js / cURL 範例。
Vidu Q3 API 的費用是多少?
Vidu Q3 每次執行 $0.15 起。實際費用會依您設定的參數(解析度、時長、輸出數量、參考素材)而調整。Playground 中「生成」按鈕旁的即時費用預覽會顯示您目前輸入的確切價格。
有哪些 Vidu Q3 變體可用?
WaveSpeedAI 提供 13 個已上線的 Vidu Q3 端點:vidu/q3-pro/text-to-video, vidu/q3-pro/start-end-to-video, vidu/q3-turbo/start-end-to-video, vidu/q3/start-end-to-video, vidu/q3/reference-to-video, vidu/q3/image-to-video-pro, vidu/q3-pro/image-to-video, vidu/q3-turbo/image-to-video等。每個變體都有各自的 Playground 頁面與定價。
Vidu Q3 的輸出可以商用嗎?
商業使用權依 Shengshu 的模型授權而定。多數 Shengshu 模型允許商用輸出;請參閱各模型 Playground 頁面中的授權摘要,以及 WaveSpeedAI 的服務條款了解平台層級的條件。
為什麼要在 WaveSpeedAI 上使用 Vidu Q3,而不是直接使用?
一組 API 金鑰、一個帳單帳戶,即可使用 Vidu Q3 以及來自其他供應商的 1,000+ 個 AI 模型。無需設定各家 SDK、無需處理各自獨立的速率限制、無需為每個供應商重寫整合程式碼。價格通常與 Shengshu 的直接 API 相當或更低。
提供者
關於 Shengshu
Vidu Q3 背後的團隊,以及 Shengshu 在 WaveSpeedAI 上的完整模型陣容。
生數科技是源自清華大學的中國 AI 實驗室,也是 Vidu 影片生成模型系列背後的團隊。Vidu Q3 在 Standard、Pro 與 Turbo 等級中提供文字轉影片、圖片轉影片、參考轉影片(1 至 4 張參考圖片,維持多主體一致性)與 start-end-to-video(在兩張靜態圖之間進行關鍵幀插值)。部分版本最長可輸出 16 秒。
在 WaveSpeedAI 上開始使用 Vidu Q3 打造應用
註冊即贈免費入門點數。一組 API 金鑰,涵蓋來自 Shengshu 與所有其他供應商的 1,000+ 個 AI 模型。