Google Gemini Omni 1.1 Flash Text To Video API Documentation
Playground
Try it on WaveSpeedAI!Gemini Omni 1.1 Flash Text to Video creates short AI videos with synchronized audio from text prompts at resolutions from 360p to 4K. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
Features
Turn a prompt into a polished video with synchronized audio and stronger creative control. Gemini Omni 1.1 Flash supports a fast 360p draft workflow for prompt exploration, standard 720p generation, and production-ready 1080p or 4K delivery.
Why Choose This?
- Fast creative iteration — Use 360p previews to test scenes, camera direction, and prompt variations quickly and economically.
- Professional delivery — Move the chosen concept to 1080p or crisp 4K output.
- Audiovisual generation — Create motion and synchronized sound together from one prompt.
- Directable results — Describe framing, camera movement, scene progression, dialogue, music, and ambience in natural language.
Parameters
| Parameter | Required | Description |
|---|---|---|
prompt | Yes | Scene, motion, camera, style, dialogue, and audio instructions. |
aspect_ratio | No | 16:9 or 9:16. Default: 16:9. |
resolution | No | 360p, 720p, 1080p, or 4k. Default: 720p. |
duration | No | 3–10 seconds. Default: 8 seconds. |
Pricing
| Resolution | Price per second |
|---|---|
| 360p | $0.039 |
| 720p | $0.13 |
| 1080p | $0.195 |
| 4K | $0.39 |
Billing is based on generated duration and the selected resolution. Start at 360p for drafts, then render the selected direction at a higher resolution.
Authentication
For authentication details, please refer to the Authentication Guide.
API Endpoints
Submit Task & Query Result
set -euo pipefail
export WAVESPEED_API_KEY="your-api-key"
REQUEST_BODY=$(cat <<'JSON'
{
"prompt": "A cinematic ocean wave at sunrise, highly detailed",
"aspect_ratio": "16:9",
"resolution": "720p",
"duration": 8
}
JSON
)
# 1. Submit the prediction.
SUBMIT_RESPONSE=$(curl --silent --show-error --fail-with-body \
-X POST "https://api.wavespeed.ai/api/v3/google/gemini-omni-1.1-flash/text-to-video" \
-H "Authorization: Bearer ${WAVESPEED_API_KEY}" \
-H "Content-Type: application/json" \
-d "${REQUEST_BODY}")
TASK=$(printf '%s' "${SUBMIT_RESPONSE}" | jq 'if type == "object" and has("data") then .data else . end')
PREDICTION_ID=$(printf '%s' "${TASK}" | jq -r '.id // empty')
if [ -z "${PREDICTION_ID}" ]; then
printf 'Submission response did not contain a prediction id
' >&2
exit 1
fi
RESULT_URL="https://api.wavespeed.ai/api/v3/predictions/${PREDICTION_ID}/result"
# 2. Poll until the prediction finishes.
while true; do
RESPONSE=$(curl --silent --show-error --fail-with-body \
"${RESULT_URL}" \
-H "Authorization: Bearer ${WAVESPEED_API_KEY}")
RESULT=$(printf '%s' "${RESPONSE}" | jq 'if type == "object" and has("data") then .data else . end')
STATUS=$(printf '%s' "${RESULT}" | jq -r '.status // empty')
case "${STATUS}" in
completed) printf '%s\n' "${RESULT}" | jq '.outputs'; break ;;
failed|cancelled|timeout|deleted) printf '%s\n' "${RESULT}" | jq . >&2; exit 1 ;;
*) sleep 2 ;;
esac
doneParameters
Task Submission Parameters
Request Parameters
| Parameter | Type | Required | Default | Range | Description |
|---|---|---|---|---|---|
| prompt | string | Yes | - | Text prompt describing the video to generate. | |
| aspect_ratio | string | No | 16:9 | 16:9, 9:16 | Aspect ratio of the generated video. |
| resolution | string | No | 720p | 360p, 720p, 1080p, 4k | Resolution of the generated video. |
| duration | integer | No | 8 | 3, 4, 5, 6, 7, 8, 9, 10 | Duration of the generated video in seconds. |
Response Parameters
| Parameter | Type | Description |
|---|---|---|
| code | integer | HTTP status code (e.g., 200 for success) |
| message | string | Status message (e.g., “success”) |
| data.id | string | Unique identifier for the prediction, Task Id |
| data.model | string | Model ID used for the prediction |
| data.outputs | array | Output values, usually URL strings; some models return text strings or structured result objects (empty when status is not completed) |
| data.urls | object | Object containing related API endpoints |
| data.status | string | Task status. completed is successful; failed, cancelled, timeout, and deleted are failure terminal statuses. |
| data.created_at | string | ISO timestamp of when the request was created (e.g., “2023-04-01T12:34:56.789Z”) |
| data.error | string | Error message (empty if no error occurred) |
| data.timings | object | Object containing timing details |
| data.timings.inference | integer | Inference time in milliseconds |
Result Request Parameters
| Parameter | Type | Required | Default | Description |
|---|---|---|---|---|
| id | string | Yes | - | Task ID |
Result Response Parameters
| Parameter | Type | Description |
|---|---|---|
| code | integer | HTTP status code (e.g., 200 for success) |
| message | string | Status message (e.g., “success”) |
| data | object | The prediction data object containing all details |
| data.id | string | Unique identifier for the prediction |
| data.model | string | Model ID used for the prediction |
| data.outputs | array<string | object> | Array of generated outputs (empty when status is not completed). Items are usually URL strings, but may be text strings or structured result objects, depending on the model. |
| data.urls | object | Object containing related API endpoints |
| data.status | string | Status: completed is successful; failed, cancelled, timeout, and deleted are failure terminal statuses |
| data.created_at | string | ISO timestamp of when the request was created |
| data.error | string | Error message (empty if no error occurred) |
| data.timings | object | Object containing timing details |
| data.timings.inference | integer | Inference time in milliseconds |