Black Forest Labs Flux 3 Text To Video Draft API Documentation

Black Forest Labs Flux 3 Text To Video Draft API Documentation

Playground

Try it on WaveSpeedAI!

FLUX 3 Text-to-Video Draft generates short draft videos from text prompts with coherent motion, prompt-guided camera direction, optional synchronized audio, and pricing optimized for rapid iteration.

Features

FLUX 3 Draft Text-to-Video turns a written prompt into a fast draft video preview with coherent motion and prompt-guided scene composition. It is designed for rapid ideation, storyboarding, prompt iteration, and testing creative directions before moving to a final-quality workflow.


Why Choose This?

  • Text-to-video draft generation
    Generate draft videos directly from natural-language prompts.

  • Fast creative testing
    Explore subject behavior, scene structure, camera movement, lighting, pacing, and visual style quickly.

  • Prompt-guided motion
    Describe action, timing, camera direction, environment, mood, and style in one prompt.

  • Optional synchronized audio
    Generate matching audio for early ambience, effects, music, or dialogue tests.

  • Flexible duration
    Choose a draft video duration from 5 to 20 seconds.

  • Multiple aspect ratios
    Supports landscape, portrait, square, and cinematic formats.


Parameters

ParameterRequiredDescription
promptYesDescribe the subject, action, environment, camera movement, timing, mood, lighting, and visual style of the draft video.
aspect_ratioNoOutput aspect ratio. Options: 21:9, 2:1, 16:9, 4:3, 1:1, 3:4, or 9:16. Default: 9:16.
durationNoDraft video duration in seconds. Range: 5–20. Default: 5.
generate_audioNoGenerate synchronized audio for the draft video. Default: true.

How to Use

  1. Write the prompt — Describe the subject, action, setting, camera movement, timing, lighting, mood, and visual style.
  2. Choose aspect ratio — Select the output format that matches your target use case.
  3. Set duration — Choose a draft duration from 5 to 20 seconds.
  4. Configure audio optional — Enable generate_audio when you want to test ambience, effects, music, or dialogue.
  5. Generate variations — Compare motion, composition, and pacing across multiple drafts.
  6. Refine the prompt — Adjust the scene direction before moving to a final-quality workflow.

Pricing

Pricing is based on billed video duration.

Billed duration is rounded up to the next whole second and capped at 20 seconds. The API accepts durations from 5 to 20 seconds. Audio generation does not add a separate charge.

Billing UnitPrice
Per 5s$0.30
Per second$0.06

Example Costs

DurationBilled DurationCost
5s5s$0.30
10s10s$0.60
15s15s$0.90
20s20s$1.20

Best Use Cases

  • Rapid prompt testing — Test different scene ideas, actions, and visual styles quickly.
  • Storyboarding — Turn written concepts into motion previews for planning and review.
  • Pre-visualization — Explore camera movement, pacing, and scene composition before production.
  • Social video concepts — Draft vertical, square, landscape, or cinematic content ideas.
  • Advertising drafts — Test early product-video, brand, or campaign directions.
  • Motion comparison — Compare multiple camera paths or timing choices from the same concept.

Pro Tips

  • Start with one clear action so the draft has a focused motion target.
  • Include the subject, setting, action, camera movement, lighting, mood, and style.
  • Add camera instructions such as tracking shot, dolly-in, orbit, zoom, tilt, or locked-off frame.
  • Use short timing cues when the scene has multiple beats.
  • Keep the first prompt simple, then add more secondary details after the motion direction is stable.
  • Use shorter durations for quick iteration and longer durations when the scene needs more time to unfold.

Authentication

For authentication details, please refer to the Authentication Guide.

API Endpoints

Submit Task & Query Result

set -euo pipefail

export WAVESPEED_API_KEY="your-api-key"

REQUEST_BODY=$(cat <<'JSON'
{
  "prompt": "A cinematic ocean wave at sunrise, highly detailed",
  "aspect_ratio": "16:9",
  "duration": 5,
  "generate_audio": true
}
JSON
)

# 1. Submit the prediction.
SUBMIT_RESPONSE=$(curl --silent --show-error --fail-with-body \
  -X POST "https://api.wavespeed.ai/api/v3/black-forest-labs/flux-3/text-to-video-draft" \
  -H "Authorization: Bearer ${WAVESPEED_API_KEY}" \
  -H "Content-Type: application/json" \
  -d "${REQUEST_BODY}")

TASK=$(printf '%s' "${SUBMIT_RESPONSE}" | jq 'if type == "object" and has("data") then .data else . end')
PREDICTION_ID=$(printf '%s' "${TASK}" | jq -r '.id // empty')
if [ -z "${PREDICTION_ID}" ]; then
  printf 'Submission response did not contain a prediction id
' >&2
  exit 1
fi
RESULT_URL="https://api.wavespeed.ai/api/v3/predictions/${PREDICTION_ID}/result"

# 2. Poll until the prediction finishes.
while true; do
  RESPONSE=$(curl --silent --show-error --fail-with-body \
    "${RESULT_URL}" \
    -H "Authorization: Bearer ${WAVESPEED_API_KEY}")
  RESULT=$(printf '%s' "${RESPONSE}" | jq 'if type == "object" and has("data") then .data else . end')
  STATUS=$(printf '%s' "${RESULT}" | jq -r '.status // empty')

  case "${STATUS}" in
    completed) printf '%s\n' "${RESULT}" | jq '.outputs'; break ;;
    failed|cancelled|timeout|deleted) printf '%s\n' "${RESULT}" | jq . >&2; exit 1 ;;
    *) sleep 2 ;;
  esac
done

Parameters

Task Submission Parameters

Request Parameters

ParameterTypeRequiredDefaultRangeDescription
promptstringYes-The text prompt describing the video you want to generate.
aspect_ratiostringNo16:921:9, 2:1, 16:9, 4:3, 1:1, 3:4, 9:16The aspect ratio of the generated media.
durationintegerNo55 ~ 20Video length in seconds (5-20).
generate_audiobooleanNotrue-Whether to generate audio for the video.

Response Parameters

ParameterTypeDescription
codeintegerHTTP status code (e.g., 200 for success)
messagestringStatus message (e.g., “success”)
data.idstringUnique identifier for the prediction, Task Id
data.modelstringModel ID used for the prediction
data.outputsarrayOutput values, usually URL strings; some models return text strings or structured result objects (empty when status is not completed)
data.urlsobjectObject containing related API endpoints
data.statusstringTask status. completed is successful; failed, cancelled, timeout, and deleted are failure terminal statuses.
data.created_atstringISO timestamp of when the request was created (e.g., “2023-04-01T12:34:56.789Z”)
data.errorstringError message (empty if no error occurred)
data.timingsobjectObject containing timing details
data.timings.inferenceintegerInference time in milliseconds

Result Request Parameters

ParameterTypeRequiredDefaultDescription
idstringYes-Task ID

Result Response Parameters

ParameterTypeDescription
codeintegerHTTP status code (e.g., 200 for success)
messagestringStatus message (e.g., “success”)
dataobjectThe prediction data object containing all details
data.idstringUnique identifier for the prediction
data.modelstringModel ID used for the prediction
data.outputsarray<string | object>Array of generated outputs (empty when status is not completed). Items are usually URL strings, but may be text strings or structured result objects, depending on the model.
data.urlsobjectObject containing related API endpoints
data.statusstringStatus: completed is successful; failed, cancelled, timeout, and deleted are failure terminal statuses
data.created_atstringISO timestamp of when the request was created
data.errorstringError message (empty if no error occurred)
data.timingsobjectObject containing timing details
data.timings.inferenceintegerInference time in milliseconds
© 2026 WaveSpeedAI. All rights reserved.