Black Forest Labs Flux 3 Start End To Video
Playground
Try it on WaveSpeedAI!FLUX 3 Start-End-to-Video generates 5-20 second transition videos between required start and end images, using prompt guidance to control action, camera movement, visual style, motion continuity, and optional synchronized audio at 720P / 1080P output. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
Features
FLUX 3 Start-End-to-Video generates a video transition between a required starting image and a required ending image. Use the prompt to define the action, transformation, camera movement, pacing, and visual style that connect the two keyframes, with optional synchronized audio.
Key Capabilities
- Generate a controlled transition between a start image and an end image.
- Use prompts to define the subject movement, transformation, camera path, environment, and visual style between keyframes.
- Generate optional synchronized audio for the transition.
- Choose 720p or 1080p output and a duration from 5 to 20 seconds.
- Support landscape, portrait, square, and cinematic aspect ratios.
Parameters
| Parameter | Required | Description |
|---|---|---|
| prompt | Yes | Describe the action or transformation connecting the start and end images, including camera movement, timing, mood, lighting, and style. |
| start_image | Yes | Image used as the first visual keyframe. |
| end_image | Yes | Image used as the final visual keyframe. |
| aspect_ratio | No | Output aspect ratio. Options: 21:9, 2:1, 16:9, 4:3, 1:1, 3:4, or 9:16. |
| resolution | No | Output resolution. Options: 720p or 1080p. Default: 720p. |
| duration | No | Video duration in seconds. Choose an integer from 5 to 20. Default: 5. |
| generate_audio | No | Generate synchronized audio for the video. Default: true. |
How to Use
- Provide the image that should appear at the beginning of the video.
- Provide the image that should appear at the end of the video.
- Describe the movement, transformation, camera path, and pacing that should connect the two keyframes.
- Choose the aspect ratio, resolution, duration, and whether synchronized audio should be generated.
- Submit the request and download the generated transition after processing is complete.
Prompting Tips
- Describe the intended visual path from the start image to the end image rather than listing unrelated scenes.
- Explain what moves, transforms, appears, disappears, or changes between the keyframes.
- Use camera directions such as dolly-in, pull-back, orbit, tracking shot, or crane movement to shape the transition.
- Specify how quickly the transition should happen and which visual details must remain stable.
Pricing
Pricing is calculated by billed video duration and output resolution.
| Resolution | Price |
|---|---|
| 720p | $0.17/s |
| 1080p | $0.29/s |
Billing Rules
- The billed duration is rounded up to the next whole second.
- The minimum billed duration is 5 seconds.
- The maximum billed duration is 20 seconds.
- Audio generation does not add a separate charge in the current pricing formula.
- Example: a 10-second 720p video costs $1.70.
- Example: a 10-second 1080p video costs $2.90.
Best Use Cases
- Product reveals, logo transitions, and before-and-after presentations.
- Character transformations and outfit or environment changes.
- Keyframe-based story moments, advertisements, and social media transitions.
- Creative experiments that need a defined beginning and ending state.
Notes
- Prompt, start_image, and end_image are required.
- Use visually compatible start and end images when you want a smooth transition.
- Generated motion and audio may vary depending on the distance between the two keyframes and the prompt detail.
Related Models
- FLUX 3 Start-End-to-Video — Generate video using both a starting frame and an ending frame.
- FLUX 3 Image-to-Video — Animate a reference image into a video.
- FLUX 3 Video Extend — Extend an existing video with new generated footage.
- FLUX 3 Text-to-Video — Generate video directly from a text prompt.
Authentication
For authentication details, please refer to the Authentication Guide.
API Endpoints
Submit Task & Query Result
set -euo pipefail
export WAVESPEED_API_KEY="your-api-key"
REQUEST_BODY=$(cat <<'JSON'
{
"prompt": "A cinematic ocean wave at sunrise, highly detailed",
"start_image": "https://interactive-examples.mdn.mozilla.net/media/cc0-images/painted-hand-298-332.jpg",
"end_image": "https://interactive-examples.mdn.mozilla.net/media/cc0-images/painted-hand-298-332.jpg",
"aspect_ratio": "21:9",
"resolution": "720p",
"duration": 5,
"generate_audio": true
}
JSON
)
# 1. Submit the prediction.
SUBMIT_RESPONSE=$(curl --silent --show-error --fail-with-body \
-X POST "https://api.wavespeed.ai/api/v3/black-forest-labs/flux-3/start-end-to-video" \
-H "Authorization: Bearer ${WAVESPEED_API_KEY}" \
-H "Content-Type: application/json" \
-d "${REQUEST_BODY}")
TASK=$(printf '%s' "${SUBMIT_RESPONSE}" | jq 'if type == "object" and has("data") then .data else . end')
PREDICTION_ID=$(printf '%s' "${TASK}" | jq -r '.id // empty')
if [ -z "${PREDICTION_ID}" ]; then
printf 'Submission response did not contain a prediction id
' >&2
exit 1
fi
RESULT_URL=$(printf '%s' "${TASK}" | jq -r '.urls.get // empty')
if [ -z "${RESULT_URL}" ]; then RESULT_URL="https://api.wavespeed.ai/api/v3/predictions/${PREDICTION_ID}/result"; fi
# 2. Poll until the prediction finishes.
while true; do
RESPONSE=$(curl --silent --show-error --fail-with-body \
"${RESULT_URL}" \
-H "Authorization: Bearer ${WAVESPEED_API_KEY}")
RESULT=$(printf '%s' "${RESPONSE}" | jq 'if type == "object" and has("data") then .data else . end')
STATUS=$(printf '%s' "${RESULT}" | jq -r '.status // empty')
case "${STATUS}" in
completed) printf '%s\n' "${RESULT}" | jq '.outputs'; break ;;
failed|cancelled|timeout) printf '%s\n' "${RESULT}" | jq . >&2; exit 1 ;;
created|processing) sleep 2 ;;
*) printf 'Unexpected status: %s
' "${STATUS}" >&2; exit 1 ;;
esac
doneParameters
Task Submission Parameters
Request Parameters
| Parameter | Type | Required | Default | Range | Description |
|---|---|---|---|---|---|
| prompt | string | Yes | - | The text prompt describing the video you want to generate. | |
| start_image | string | Yes | - | - | Optional image used as the edited video's first frame. |
| end_image | string | Yes | - | - | Optional image used as the edited video's first frame. |
| aspect_ratio | string | No | - | 21:9, 2:1, 16:9, 4:3, 1:1, 3:4, 9:16 | The aspect ratio of the generated media. |
| resolution | string | No | 720p | 720p, 1080p | Resolution of the generated video. |
| duration | integer | No | 5 | 5 ~ 20 | Video length in seconds (5-20). |
| generate_audio | boolean | No | true | - | Whether to generate audio for the video. |
Response Parameters
| Parameter | Type | Description |
|---|---|---|
| code | integer | HTTP status code (e.g., 200 for success) |
| message | string | Status message (e.g., “success”) |
| data.id | string | Unique identifier for the prediction, Task Id |
| data.model | string | Model ID used for the prediction |
| data.outputs | array | Output values, usually URL strings; some models return text strings or structured result objects (empty when status is not completed) |
| data.urls | object | Object containing related API endpoints |
| data.urls.get | string | URL to retrieve the prediction result |
| data.status | string | Status of the task: created, processing, completed, or failed |
| data.created_at | string | ISO timestamp of when the request was created (e.g., “2023-04-01T12:34:56.789Z”) |
| data.error | string | Error message (empty if no error occurred) |
| data.timings | object | Object containing timing details |
| data.timings.inference | integer | Inference time in milliseconds |
Result Request Parameters
| Parameter | Type | Required | Default | Description |
|---|---|---|---|---|
| id | string | Yes | - | Task ID |
Result Response Parameters
| Parameter | Type | Description |
|---|---|---|
| code | integer | HTTP status code (e.g., 200 for success) |
| message | string | Status message (e.g., “success”) |
| data | object | The prediction data object containing all details |
| data.id | string | Unique identifier for the prediction |
| data.model | string | Model ID used for the prediction |
| data.outputs | array<string | object> | Array of generated outputs (empty when status is not completed). Items are usually URL strings, but may be text strings or structured result objects, depending on the model. |
| data.urls | object | Object containing related API endpoints |
| data.urls.get | string | URL to poll for the prediction result |
| data.status | string | Status: created, processing, completed, or failed |
| data.created_at | string | ISO timestamp of when the request was created |
| data.error | string | Error message (empty if no error occurred) |
| data.timings | object | Object containing timing details |
| data.timings.inference | integer | Inference time in milliseconds |