Black Forest Labs Flux 3 Image To Video
Playground
Try it on WaveSpeedAI!FLUX 3 Image-to-Video animates a required reference image into a 5-20 second video, preserving subject appearance while using prompt control to guide motion, camera movement, visual style, and optional synchronized audio at 720P / 1080P output. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
Features
FLUX 3 Image-to-Video transforms a required reference image into a short video with coherent motion, prompt-guided action, and optional synchronized audio. Use the image to anchor the subject, character, product, scene, or visual style, then describe how it should move and evolve.
Key Capabilities
- Animate a reference image into a 5-20 second video.
- Preserve the main subject’s appearance while adding prompt-controlled motion and camera direction.
- Generate optional synchronized audio for ambience, sound effects, music, or dialogue.
- Choose 720p or 1080p output for social content, advertisements, product showcases, and creative production.
- Support landscape, portrait, square, and cinematic aspect ratios.
Parameters
| Parameter | Required | Description |
|---|---|---|
| prompt | Yes | Describe the action, motion, environment, camera movement, timing, mood, lighting, and visual style. |
| image | Yes | Reference image used as the visual starting point. Supports PNG, JPEG, or WebP URLs. |
| aspect_ratio | No | Output aspect ratio. Options: 21:9, 2:1, 16:9, 4:3, 1:1, 3:4, or 9:16. |
| resolution | No | Output resolution. Options: 720p or 1080p. Default: 720p. |
| duration | No | Video duration in seconds. Choose an integer from 5 to 20. Default: 5. |
| generate_audio | No | Generate synchronized audio for the video. Default: true. |
How to Use
- Provide a clear reference image with the subject or composition you want to animate.
- Describe the desired action, movement, camera path, environment, and visual style in the prompt.
- Choose an aspect ratio and resolution that match the intended destination.
- Set a duration from 5 to 20 seconds and enable generate_audio when sound should be created with the scene.
- Submit the request and download the generated video after processing is complete.
Prompting Tips
- Explain what should stay consistent from the image and what should change through motion.
- Use concrete actions such as hair moving in the wind, a product rotating, a character turning, or a camera pushing in.
- Add camera language such as tracking shot, slow dolly-in, orbit, handheld movement, or locked-off frame.
- Mention lighting, atmosphere, and sound when they are important to the final result.
Pricing
Pricing is calculated by billed video duration and output resolution.
| Resolution | Price |
|---|---|
| 720p | $0.17/s |
| 1080p | $0.29/s |
Billing Rules
- The billed duration is rounded up to the next whole second.
- The minimum billed duration is 5 seconds.
- The maximum billed duration is 20 seconds.
- Audio generation does not add a separate charge in the current pricing formula.
- Example: a 10-second 720p video costs $1.70.
- Example: a 10-second 1080p video costs $2.90.
Best Use Cases
- Bringing character portraits, product photos, and illustrations to life.
- Product advertisements and e-commerce motion creatives.
- Social media animations and short promotional clips.
- Character performance tests, visual storytelling, and concept development.
Notes
- A reference image and prompt are both required.
- Use a clear image with the subject visible and enough visual detail for motion planning.
- Generated motion and audio may vary with the image content and prompt specificity.
Related Models
- FLUX 3 Start-End-to-Video — Generate video using both a starting frame and an ending frame.
- FLUX 3 Image-to-Video — Animate a reference image into a video.
- FLUX 3 Video Extend — Extend an existing video with new generated footage.
- FLUX 3 Text-to-Video — Generate video directly from a text prompt.
Authentication
For authentication details, please refer to the Authentication Guide.
API Endpoints
Submit Task & Query Result
set -euo pipefail
export WAVESPEED_API_KEY="your-api-key"
REQUEST_BODY=$(cat <<'JSON'
{
"prompt": "A cinematic ocean wave at sunrise, highly detailed",
"image": "https://interactive-examples.mdn.mozilla.net/media/cc0-images/painted-hand-298-332.jpg",
"aspect_ratio": "21:9",
"resolution": "720p",
"duration": 5,
"generate_audio": true
}
JSON
)
# 1. Submit the prediction.
SUBMIT_RESPONSE=$(curl --silent --show-error --fail-with-body \
-X POST "https://api.wavespeed.ai/api/v3/black-forest-labs/flux-3/image-to-video" \
-H "Authorization: Bearer ${WAVESPEED_API_KEY}" \
-H "Content-Type: application/json" \
-d "${REQUEST_BODY}")
TASK=$(printf '%s' "${SUBMIT_RESPONSE}" | jq 'if type == "object" and has("data") then .data else . end')
PREDICTION_ID=$(printf '%s' "${TASK}" | jq -r '.id // empty')
if [ -z "${PREDICTION_ID}" ]; then
printf 'Submission response did not contain a prediction id
' >&2
exit 1
fi
RESULT_URL=$(printf '%s' "${TASK}" | jq -r '.urls.get // empty')
if [ -z "${RESULT_URL}" ]; then RESULT_URL="https://api.wavespeed.ai/api/v3/predictions/${PREDICTION_ID}/result"; fi
# 2. Poll until the prediction finishes.
while true; do
RESPONSE=$(curl --silent --show-error --fail-with-body \
"${RESULT_URL}" \
-H "Authorization: Bearer ${WAVESPEED_API_KEY}")
RESULT=$(printf '%s' "${RESPONSE}" | jq 'if type == "object" and has("data") then .data else . end')
STATUS=$(printf '%s' "${RESULT}" | jq -r '.status // empty')
case "${STATUS}" in
completed) printf '%s\n' "${RESULT}" | jq '.outputs'; break ;;
failed|cancelled|timeout) printf '%s\n' "${RESULT}" | jq . >&2; exit 1 ;;
created|processing) sleep 2 ;;
*) printf 'Unexpected status: %s
' "${STATUS}" >&2; exit 1 ;;
esac
doneParameters
Task Submission Parameters
Request Parameters
| Parameter | Type | Required | Default | Range | Description |
|---|---|---|---|---|---|
| prompt | string | Yes | - | The text prompt describing the video you want to generate. | |
| image | string | Yes | - | URL of the image the video starts from (PNG, JPEG, or WebP).. | |
| aspect_ratio | string | No | - | 21:9, 2:1, 16:9, 4:3, 1:1, 3:4, 9:16 | The aspect ratio of the generated media. |
| resolution | string | No | 720p | 720p, 1080p | Resolution of the generated video. |
| duration | integer | No | 5 | 5 ~ 20 | Video length in seconds (5-20). |
| generate_audio | boolean | No | true | - | Whether to generate audio for the video. |
Response Parameters
| Parameter | Type | Description |
|---|---|---|
| code | integer | HTTP status code (e.g., 200 for success) |
| message | string | Status message (e.g., “success”) |
| data.id | string | Unique identifier for the prediction, Task Id |
| data.model | string | Model ID used for the prediction |
| data.outputs | array | Output values, usually URL strings; some models return text strings or structured result objects (empty when status is not completed) |
| data.urls | object | Object containing related API endpoints |
| data.urls.get | string | URL to retrieve the prediction result |
| data.status | string | Status of the task: created, processing, completed, or failed |
| data.created_at | string | ISO timestamp of when the request was created (e.g., “2023-04-01T12:34:56.789Z”) |
| data.error | string | Error message (empty if no error occurred) |
| data.timings | object | Object containing timing details |
| data.timings.inference | integer | Inference time in milliseconds |
Result Request Parameters
| Parameter | Type | Required | Default | Description |
|---|---|---|---|---|
| id | string | Yes | - | Task ID |
Result Response Parameters
| Parameter | Type | Description |
|---|---|---|
| code | integer | HTTP status code (e.g., 200 for success) |
| message | string | Status message (e.g., “success”) |
| data | object | The prediction data object containing all details |
| data.id | string | Unique identifier for the prediction |
| data.model | string | Model ID used for the prediction |
| data.outputs | array<string | object> | Array of generated outputs (empty when status is not completed). Items are usually URL strings, but may be text strings or structured result objects, depending on the model. |
| data.urls | object | Object containing related API endpoints |
| data.urls.get | string | URL to poll for the prediction result |
| data.status | string | Status: created, processing, completed, or failed |
| data.created_at | string | ISO timestamp of when the request was created |
| data.error | string | Error message (empty if no error occurred) |
| data.timings | object | Object containing timing details |
| data.timings.inference | integer | Inference time in milliseconds |