Seedream 5.0 Pro अब लाइव है | Image Generator में आज़माएं →
साइन इन

Wan 2.6 Reference to Video Flash | Fast Image-to-Video

alibaba/

WAN 2.6 Reference-to-Video Flash turns character, prop, or scene references from images or videos into new video shots with preserved identity, style, and layout plus smooth, coherent motion. Flash version with faster generation speed. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

image-to-video
इनपुट

निष्क्रिय

$0.125प्रति रन·~80 / $10

आगे:

उदाहरणसभी देखें

The man in the image is walking on the moon. He says, "The earth is so beautiful!".

The woman in picture 1 and the man in picture 2 are sitting in the school cafeteria eating. The girl asks the boy, "Why are you looking at me?" Then, both of them laugh. With natural background music added.

Make the woman in the reference video put on a jacket and sunglasses. And say, "Oh, it's cold!"

संबंधित मॉडल

README

Wan 2.6 Reference-to-Video Flash

Wan 2.6 Reference-to-Video Flash is fast reference-driven video generation model. Upload up to 5 reference images and describe the scene — the model generates high-quality video that preserves character identity and appearance, with optional audio generation and multi-shot support.

Why Choose This?

  • Multi-reference input Upload up to 5 reference images for precise character and scene guidance.

  • Identity preservation Maintains character appearance and identity across generated video frames.

  • Audio generation Optional synchronized audio for complete video output.

  • Shot type control Choose between single continuous shot or multi-shot composition.

  • Multiple resolutions Support for 720p and 1080p in both landscape and portrait orientations.

  • Prompt Enhancer Built-in tool to automatically improve your video descriptions.

Parameters

ParameterRequiredDescription
reference_urlsYesReference images (1-5, click "+ Add Item" for multiple)
promptYesText description of the video scene and motion
audioNoCustom audio track (URL or upload)
negative_promptNoElements to exclude from generation
sizeNoOutput size: 1280720, 7201280, 19201080, 10801920
durationNoVideo length: 5 or 10 seconds (default: 5)
shot_typeNoShot composition: single, multi (default: multi)
enable_audioNoGenerate synchronized audio (default: enabled)
enable_prompt_expansionNoEnable prompt optimizer (default: disabled)
seedNoRandom seed for reproducibility (-1 for random)

How to Use

  1. Upload reference images — add 1-5 character or scene references.
  2. Write your prompt — describe the scene, motion, and camera work.
  3. Upload audio (optional) — provide a custom audio track.
  4. Set size — choose resolution and orientation.
  5. Set duration — 5 or 10 seconds.
  6. Choose shot type — single for one continuous shot, multi for varied compositions.
  7. Configure audio — enable/disable audio generation.
  8. Run — submit and download your video.

Pricing

Pricing depends on resolution, duration, and audio settings.

SizeDurationAudio OffAudio On
720p5s$0.25$0.50
720p10s$0.375$0.75
1080p5s$0.40$0.80
1080p10s$0.60$1.20

Billing Rules

  • Resolution multiplier: 720p (1280720 / 7201280) = 1×, 1080p (19201080 / 10801920) = 1.6×
  • Audio multiplier: disabled = 1×, enabled = 2×

Best Use Cases

  • Character Animation — Generate videos that preserve character identity from reference photos.
  • Social Media Content — Create engaging videos featuring consistent characters.
  • Storytelling — Produce narrative scenes with identity-consistent characters.
  • Marketing & Ads — Generate promotional videos featuring specific people or characters.
  • Multi-shot Production — Create videos with varied camera angles and compositions.

Pro Tips

  • Use multiple reference images from different angles for better identity preservation.
  • Use "multi" shot type for more dynamic, cinematic compositions.
  • Disable enable_audio for faster processing when audio is not needed.
  • Add negative prompts to avoid common issues (e.g., "blurry, distorted").
  • Enable prompt expansion for automatic prompt optimization.
  • Use 720p for drafts and testing, 1080p for final production.

Notes

  • Both reference_urls and prompt are required fields.
  • Maximum 5 reference images per generation.
  • Duration options are 5 or 10 seconds only.
  • Ensure uploaded image and audio URLs are publicly accessible.
  • Seed value -1 generates a random seed each time.
  • If your result don't have sound, please add prompt like "Add background sound".

More Models to Try

नोट:यह वेबसाइट तृतीय पक्षों द्वारा प्रदान किए गए AI मॉडलों का उपयोग करती है।

Wan 2.6 Reference To Video Flash API — Quick start

Grab a WaveSpeedAI API key, then call POST https://api.wavespeed.ai/api/v3/alibaba/wan-2.6/reference-to-video-flash with your input as JSON. The endpoint returns a prediction id. Start polling the result endpoint around every 2 seconds, increase the interval for long-running tasks, and stop on any terminal status. On completed, read output values from data.outputs. Examples for Wan 2.6 Reference To Video Flash below.

HTTP example
set -euo pipefail

: "${WAVESPEED_API_KEY:?Set WAVESPEED_API_KEY}"

REQUEST_BODY=$(cat <<'JSON'
{
    "reference_urls": [
        "https://interactive-examples.mdn.mozilla.net/media/cc0-images/painted-hand-298-332.jpg"
    ],
    "prompt": "A cinematic shot of a city at sunset, soft golden light",
    "size": "1280*720",
    "duration": 5,
    "shot_type": "single",
    "enable_audio": true,
    "enable_prompt_expansion": false,
    "seed": -1
}
JSON
)

# 1. Submit the prediction.
SUBMIT_RESPONSE=$(curl --silent --show-error --fail-with-body \
  -X POST "https://api.wavespeed.ai/api/v3/alibaba/wan-2.6/reference-to-video-flash" \
  -H "Content-Type: application/json" \
  -H "Authorization: Bearer $WAVESPEED_API_KEY" \
  -d "$REQUEST_BODY")

TASK=$(printf '%s' "$SUBMIT_RESPONSE" | jq 'if has("data") then .data else . end')
PREDICTION_ID=$(printf '%s' "$TASK" | jq -r '.id')
if [ -z "$PREDICTION_ID" ] || [ "$PREDICTION_ID" = "null" ]; then
  printf 'Submission response did not contain a prediction id
' >&2
  exit 1
fi
RESULT_URL=$(printf '%s' "$TASK" | jq -r '.urls.get // empty')
if [ -z "$RESULT_URL" ]; then
  RESULT_URL="https://api.wavespeed.ai/api/v3/predictions/$PREDICTION_ID/result"
fi

# 2. Poll until the prediction finishes.
while true; do
  RESPONSE=$(curl --silent --show-error --fail-with-body "$RESULT_URL" \
    -H "Authorization: Bearer $WAVESPEED_API_KEY")
  RESULT=$(printf '%s' "$RESPONSE" | jq 'if has("data") then .data else . end')
  STATUS=$(printf '%s' "$RESULT" | jq -r '.status')
  case "$STATUS" in
    completed) printf '%s\n' "$RESULT" | jq '.outputs'; break ;;
    failed|cancelled|timeout) printf '%s\n' "$RESULT" | jq . >&2; exit 1 ;;
    created|processing) sleep 2 ;;
    *) printf 'Unexpected status: %s
' "$STATUS" >&2; exit 1 ;;
  esac
done
Node.js example
const submitUrl = "https://api.wavespeed.ai/api/v3/alibaba/wan-2.6/reference-to-video-flash";
const apiKey = process.env.WAVESPEED_API_KEY;
if (!apiKey) throw new Error('Set WAVESPEED_API_KEY');

async function requestJson(url, options = {}) {
  const response = await fetch(url, options);
  if (!response.ok) throw new Error(await response.text());
  return response.json();
}

// 1. Submit the prediction.
const body = await requestJson(submitUrl, {
  method: "POST",
  headers: {
    "Authorization": `Bearer ${apiKey}`,
    "Content-Type": "application/json",
  },
  body: JSON.stringify({
        "reference_urls": [
                "https://interactive-examples.mdn.mozilla.net/media/cc0-images/painted-hand-298-332.jpg"
        ],
        "prompt": "A cinematic shot of a city at sunset, soft golden light",
        "size": "1280*720",
        "duration": 5,
        "shot_type": "single",
        "enable_audio": true,
        "enable_prompt_expansion": false,
        "seed": -1
}),
});
const task = body.data ?? body;
if (!task.id) throw new Error("Submission response did not contain a prediction id");
const resultUrl = task.urls?.get ||
  `https://api.wavespeed.ai/api/v3/predictions/${task.id}/result`;

// 2. Poll until the prediction finishes.
while (true) {
  const resultBody = await requestJson(resultUrl, {
    headers: { "Authorization": `Bearer ${apiKey}` },
  });
  const result = resultBody.data ?? resultBody;
  if (result.status === "completed") {
    console.log(result.outputs);
    break;
  }
  if (["failed", "cancelled", "timeout"].includes(result.status)) throw new Error(JSON.stringify(result));
  if (!["created", "processing"].includes(result.status)) throw new Error("Unexpected status: " + result.status);
  await new Promise(resolve => setTimeout(resolve, 2000));
}
Python example
import json
import os
import time
from urllib.request import Request, urlopen

api_key = os.environ["WAVESPEED_API_KEY"]
headers = {"Authorization": f"Bearer {api_key}", "Content-Type": "application/json"}
payload = {
    "reference_urls": [
        "https://interactive-examples.mdn.mozilla.net/media/cc0-images/painted-hand-298-332.jpg"
    ],
    "prompt": "A cinematic shot of a city at sunset, soft golden light",
    "size": "1280*720",
    "duration": 5,
    "shot_type": "single",
    "enable_audio": True,
    "enable_prompt_expansion": False,
    "seed": -1
}

def request_json(url, data=None):
    request = Request(url, data=data, headers=headers, method="POST" if data else "GET")
    with urlopen(request) as response:
        return json.load(response)

# 1. Submit the prediction.
body = request_json("https://api.wavespeed.ai/api/v3/alibaba/wan-2.6/reference-to-video-flash", json.dumps(payload).encode())
task = body.get("data", body)
if not task.get("id"):
    raise RuntimeError("Submission response did not contain a prediction id")
result_url = task.get("urls", {}).get("get") or f"https://api.wavespeed.ai/api/v3/predictions/{task['id']}/result"

# 2. Poll until the prediction finishes.
while True:
    result_body = request_json(result_url)
    result = result_body.get("data", result_body)
    status = result.get("status")
    if status == "completed":
        print(result.get("outputs", []))
        break
    if status in {"failed", "cancelled", "timeout"}:
        raise RuntimeError(result)
    if status not in {"created", "processing"}:
        raise RuntimeError(f"Unexpected status: {status}")
    time.sleep(2)

Wan 2.6 Reference To Video Flash API — Frequently asked questions

What is the Wan 2.6 Reference To Video Flash API?

Wan 2.6 Reference To Video Flash is a Alibaba model for video generation from images, exposed as a REST API on WaveSpeedAI. WAN 2.6 Reference-to-Video Flash turns character, prop, or scene references from images or videos into new video shots with preserved identity, style, and layout plus smooth, coherent motion. Flash version with faster generation speed. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing. You can call it programmatically or try it from the playground above.

How do I call the Wan 2.6 Reference To Video Flash API?

POST your input parameters to the model's REST endpoint (shown in the API tab of this playground) with your WaveSpeedAI API key in the Authorization header. Submission returns a prediction ID. Poll the result endpoint starting around every 2 seconds, increase the interval for long-running tasks, and stop on any terminal status. The playground generates production-oriented Python, JavaScript, and cURL examples with timeouts, transient-error handling, and safe GET retries. Full request/response shape is documented at https://wavespeed.ai/docs/docs-api/alibaba/alibaba-wan-2.6-reference-to-video-flash.

How much does Wan 2.6 Reference To Video Flash cost per run?

Wan 2.6 Reference To Video Flash starts at $0.13 per run. That figure is the base price — the final charge scales with the parameters you set in the form (output size, length, count, references, or whatever knobs this model exposes), so a higher-quality or larger output costs more than a minimal one. The exact cost for your current input is shown live next to the Generate button before you submit, and the actual per-call charge is recorded on the prediction afterwards.

What inputs does Wan 2.6 Reference To Video Flash accept?

Key inputs: `prompt`, `audio`, `duration`, `size`, `seed`, `negative_prompt`. The full JSON schema (types, defaults, allowed values) is rendered above the Generate button and mirrored in the API reference at https://wavespeed.ai/docs/docs-api/alibaba/alibaba-wan-2.6-reference-to-video-flash.

How long does Wan 2.6 Reference To Video Flash take to generate?

Median end-to-end generation time on WaveSpeedAI is around 68 seconds per request, based on recent successful runs. Queue time varies with global demand; live status is visible in the prediction record.

Can I use Wan 2.6 Reference To Video Flash outputs commercially?

Commercial usage rights depend on the model's license, set by its provider (Alibaba). The license summary appears on the model card above; see WaveSpeedAI's Terms of Service for platform-level conditions.

Wan 2.6 Reference to Video Flash | Fast Image-to-Video API on WaveSpeedAI