Alibaba Wan 3.0 Prime Image To Video Spicy API Documentation
Playground
Try it on WaveSpeedAI!Wan 3.0 Prime Spicy Image to Video is an accelerated variant that animates a first-frame image into a cinematic video, with optional last-frame guidance, flexible 2-30 second duration, aspect ratio control, optional audio, and deep-thinking controls for high-quality motion and scene continuity. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
Features
Wan 3.0 Prime Spicy Image-to-Video animates a first-frame image into a coherent video with optional last-frame guidance and audio generation. It uses the input image as the visual starting point while adding prompt-guided motion, camera movement, and scene progression.
Why Choose This?
-
Image-to-video generation
Animate a first-frame image into a complete video. -
Optional last-frame guidance
Uselast_imageto guide the ending frame and create more controlled start-to-end motion. -
Prompt-guided motion
Describe subject action, camera movement, scene progression, lighting, and mood. -
Flexible resolution options
Choose480p,720p, or1080pdepending on cost and output quality needs. -
Audio generation
Keepenable_audioenabled when audio should be included in the generated video. -
Prompt expansion
Enableenable_prompt_expansionwhen you want the prompt optimizer to expand the original instruction.
Parameters
| Parameter | Required | Description |
|---|---|---|
| image | Yes | First-frame image URL or Base64-encoded data used to guide the video generation. |
| prompt | No | Optional prompt describing the scene, action, camera movement, and mood. |
| last_image | No | Optional last-frame image URL or Base64-encoded data used to guide the ending of the video. |
| resolution | No | Output resolution: 480p, 720p, or 1080p. Default: 720p. |
| aspect_ratio | No | Output aspect ratio: 16:9, 9:16, 1:1, 4:3, or 3:4. If omitted, the output adapts to the input image. |
| duration | No | Video duration in seconds. Range: 2–30. Default: 5. |
| enable_prompt_expansion | No | Enable prompt optimization and expansion. Default: false. |
| enable_audio | No | Include audio in the generated video. Default: true. |
| seed | No | Random seed for generation. Use -1 for a random seed. |
How to Use
- Upload a first-frame image — Provide the image that should define the starting frame and composition.
- Write a prompt optional — Describe the desired action, motion, camera movement, lighting, mood, and scene progression.
- Add a last image optional — Use
last_imagewhen you want stronger control over the ending frame. - Choose resolution — Select
480p,720p, or1080p. - Choose aspect ratio optional — Select a supported ratio or leave it empty to adapt to the input image.
- Set duration — Choose a video length from
2to30seconds. - Configure prompt expansion optional — Enable
enable_prompt_expansionwhen you want the prompt optimized automatically. - Configure audio optional — Keep
enable_audioenabled when audio is needed. - Set seed optional — Use a fixed seed for reproducible results, or
-1for random generation. - Submit — Generate the final image-to-video output.
Pricing
Pricing is based on selected resolution and billed duration.
Billed duration is rounded up to the next whole second and clamped to the 2–30s range.
| Resolution | Per 5s | Per second |
|---|---|---|
| 480p | $0.375 | $0.075 |
| 720p | $0.750 | $0.150 |
| 1080p | $1.500 | $0.300 |
Example Costs
| Resolution | 2s | 5s | 10s | 30s |
|---|---|---|---|---|
| 480p | $0.15 | $0.375 | $0.75 | $2.25 |
| 720p | $0.30 | $0.75 | $1.50 | $4.50 |
| 1080p | $0.60 | $1.50 | $3.00 | $9.00 |
last_image, aspect_ratio, enable_prompt_expansion, enable_audio, and seed do not add separate charges.
Best Use Cases
- Image animation — Turn still images into short motion clips.
- Character motion — Animate portraits, characters, or illustrated subjects.
- Start-and-end-frame generation — Use
imageandlast_imageto guide both the beginning and ending of the video. - Product and marketing videos — Animate product shots, campaign visuals, and promotional images.
- Creative prototyping — Test motion, camera direction, and scene progression from a fixed first frame.
- Social video content — Create short-form videos from static images for different layouts.
Pro Tips
- Use a clear, high-quality first-frame image for stronger subject and composition consistency.
- Use
last_imagewhen the final pose, framing, or scene state matters. - Describe visible action clearly instead of only describing visual style.
- Include camera movement, lighting, mood, and scene progression when they matter.
- Use
480pfor lower-cost testing and1080pfor higher-resolution output. - Enable
enable_prompt_expansionwhen the original prompt is short or needs more detail. - Set a fixed
seedwhen comparing prompt or parameter changes.
Notes
imageis the only required parameter.durationsupports values from2to30seconds.- If
aspect_ratiois omitted, the output adapts to the input image. last_imageis optional and provides ending-frame guidance.enable_audiodefaults totrue.enable_prompt_expansiondefaults tofalse.
Related Models
- Wan 3.0 Prime Image-to-Video — Animate a first-frame image with the standard Prime image-to-video workflow.
- Wan 3.0 Image-to-Video — Generate videos from a first-frame image with optional last-frame guidance.
- Wan 3.0 Reference-to-Video — Generate videos using reference images, videos, or audio for multimodal guidance.
Authentication
For authentication details, please refer to the Authentication Guide.
API Endpoints
Submit Task & Query Result
set -euo pipefail
export WAVESPEED_API_KEY="your-api-key"
REQUEST_BODY=$(cat <<'JSON'
{
"image": "https://interactive-examples.mdn.mozilla.net/media/cc0-images/painted-hand-298-332.jpg",
"resolution": "720p",
"aspect_ratio": "16:9",
"duration": 5,
"enable_prompt_expansion": false,
"enable_audio": true
}
JSON
)
# 1. Submit the prediction.
SUBMIT_RESPONSE=$(curl --silent --show-error --fail-with-body \
-X POST "https://api.wavespeed.ai/api/v3/alibaba/wan-3.0-prime/image-to-video-spicy" \
-H "Authorization: Bearer ${WAVESPEED_API_KEY}" \
-H "Content-Type: application/json" \
-d "${REQUEST_BODY}")
TASK=$(printf '%s' "${SUBMIT_RESPONSE}" | jq 'if type == "object" and has("data") then .data else . end')
PREDICTION_ID=$(printf '%s' "${TASK}" | jq -r '.id // empty')
if [ -z "${PREDICTION_ID}" ]; then
printf 'Submission response did not contain a prediction id
' >&2
exit 1
fi
RESULT_URL="https://api.wavespeed.ai/api/v3/predictions/${PREDICTION_ID}/result"
# 2. Poll until the prediction finishes.
while true; do
RESPONSE=$(curl --silent --show-error --fail-with-body \
"${RESULT_URL}" \
-H "Authorization: Bearer ${WAVESPEED_API_KEY}")
RESULT=$(printf '%s' "${RESPONSE}" | jq 'if type == "object" and has("data") then .data else . end')
STATUS=$(printf '%s' "${RESULT}" | jq -r '.status // empty')
case "${STATUS}" in
completed) printf '%s\n' "${RESULT}" | jq '.outputs'; break ;;
failed|cancelled|timeout|deleted) printf '%s\n' "${RESULT}" | jq . >&2; exit 1 ;;
*) sleep 2 ;;
esac
doneParameters
Task Submission Parameters
Request Parameters
| Parameter | Type | Required | Default | Range | Description |
|---|---|---|---|---|---|
| image | string | Yes | - | The first frame image for generating the video. Supports URL or Base64-encoded data. | |
| prompt | string | No | - | Optional prompt describing the scene, action, camera movement, and mood for the video. | |
| last_image | string | No | - | - | The last frame image for generating the video (optional). Supports URL or Base64-encoded data. |
| resolution | string | No | 720p | 480p, 720p, 1080p | The resolution of the generated video. |
| aspect_ratio | string | No | - | 16:9, 9:16, 1:1, 4:3, 3:4 | The aspect ratio of the generated video. If not specified, it is adapted to the input image. |
| duration | integer | No | 5 | 2 ~ 30 | The duration of the generated media in seconds (2-30s). |
| enable_prompt_expansion | boolean | No | false | - | If set to true, the prompt optimizer will be enabled. |
| enable_audio | boolean | No | true | - | Whether to include audio in the generated video. |
| seed | integer | No | - | - | The random seed to use for the generation. -1 means a random seed will be used. |
Response Parameters
| Parameter | Type | Description |
|---|---|---|
| code | integer | HTTP status code (e.g., 200 for success) |
| message | string | Status message (e.g., “success”) |
| data.id | string | Unique identifier for the prediction, Task Id |
| data.model | string | Model ID used for the prediction |
| data.outputs | array | Output values, usually URL strings; some models return text strings or structured result objects (empty when status is not completed) |
| data.urls | object | Object containing related API endpoints |
| data.status | string | Task status. completed is successful; failed, cancelled, timeout, and deleted are failure terminal statuses. |
| data.created_at | string | ISO timestamp of when the request was created (e.g., “2023-04-01T12:34:56.789Z”) |
| data.error | string | Error message (empty if no error occurred) |
| data.timings | object | Object containing timing details |
| data.timings.inference | integer | Inference time in milliseconds |
Result Request Parameters
| Parameter | Type | Required | Default | Description |
|---|---|---|---|---|
| id | string | Yes | - | Task ID |
Result Response Parameters
| Parameter | Type | Description |
|---|---|---|
| code | integer | HTTP status code (e.g., 200 for success) |
| message | string | Status message (e.g., “success”) |
| data | object | The prediction data object containing all details |
| data.id | string | Unique identifier for the prediction |
| data.model | string | Model ID used for the prediction |
| data.outputs | array<string | object> | Array of generated outputs (empty when status is not completed). Items are usually URL strings, but may be text strings or structured result objects, depending on the model. |
| data.urls | object | Object containing related API endpoints |
| data.status | string | Status: completed is successful; failed, cancelled, timeout, and deleted are failure terminal statuses |
| data.created_at | string | ISO timestamp of when the request was created |
| data.error | string | Error message (empty if no error occurred) |
| data.timings | object | Object containing timing details |
| data.timings.inference | integer | Inference time in milliseconds |