Bytedance Seedance 2.0 Mini Text To Video Turbo
Playground
Try it on WaveSpeedAI!Seedance 2.0 Mini Text to Video Turbo is ByteDance’s faster, lower-cost text-to-video model for cinematic multi-shot videos. It generates narrative sequences from text prompts with AI camera control, consistent characters, 720P / 1080P output, 5-12s duration, and flexible aspect ratios. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
Features
ByteDance Seedance 2.0 Mini Text-to-Video Turbo generates videos from natural-language prompts with optional reference images, videos, and audio. Describe the scene, action, camera movement, and mood, then choose aspect ratio, resolution, duration, and audio generation settings.
- Need the standard text-to-video version instead? Try ByteDance Seedance 2.0 Mini Text-to-Video.
- Need to generate from a start image instead? Try ByteDance Seedance 2.0 Mini Image-to-Video.
Why Choose This?
-
Turbo text-to-video generation
Generate videos directly from text prompts with a faster turbo workflow. -
Reference-guided generation
Add reference images, videos, or audio to guide visual style, characters, scene composition, motion, or audio direction. -
Native audio generation
Generate synchronized audio together with the output video usinggenerate_audio. -
Flexible aspect ratios
Supports16:9,9:16,4:3,3:4,1:1, and21:9for different video layouts. -
Resolution options
Generate videos in720por1080p. -
Web search option
Enableenable_web_searchwhen the prompt needs real-time information.
Parameters
| Parameter | Required | Description |
|---|---|---|
| prompt | Yes | Describe the scene, action, camera movement, and mood for the video. |
| reference_images | No | Reference image URLs to guide visual style, characters, or scene composition. Supports up to 9 images. |
| reference_videos | No | Reference video URLs. Supports up to 3 videos, with total reference video length no more than 15 seconds. |
| reference_audios | No | Reference audio URLs. Supports up to 3 audio files, with total reference audio length no more than 15 seconds. |
| aspect_ratio | No | Output aspect ratio: 16:9, 9:16, 4:3, 3:4, 1:1, or 21:9. Default: 16:9. |
| resolution | No | Output video resolution: 720p or 1080p. Default: 720p. |
| duration | No | Duration of the generated video in seconds. Range: 4–15. Default: 5. |
| enable_web_search | No | Enable web search for real-time information. Default: false. |
| generate_audio | No | Whether to generate native audio synchronized with the output video. Default: true. |
How to Use
- Write your prompt — Describe the scene, subject, action, camera movement, lighting, and mood.
- Add references optional — Upload reference images, videos, or audio when you want stronger guidance.
- Choose aspect ratio — Select the output format, such as
16:9,9:16, or1:1. - Choose resolution — Select
720por1080p. - Set duration — Choose a duration between
4and15seconds. - Configure audio optional — Keep
generate_audioenabled for synchronized native audio, or disable it if audio is not needed. - Submit — Generate the final video.
Example Prompt
A cinematic night market scene, warm lantern light, people walking through a narrow street, soft camera tracking movement, steam rising from food stalls, realistic atmosphere, gentle ambient motion, rich colors, detailed urban background.
Pricing
Per 5 Seconds
| Resolution | Without Reference Videos | With Reference Videos |
|---|---|---|
| 720p | $0.35 | $0.65 |
| 1080p | $0.375 | $0.675 |
Per Second
| Resolution | Without Reference Videos | With Reference Videos |
|---|---|---|
| 720p | $0.07 | $0.13 |
| 1080p | $0.075 | $0.135 |
Example Costs
| Resolution | Reference Videos | 4s | 5s | 10s | 15s |
|---|---|---|---|---|---|
| 720p | No | $0.28 | $0.35 | $0.70 | $1.05 |
| 1080p | No | $0.30 | $0.375 | $0.75 | $1.125 |
| 720p | Yes | $0.52 | $0.65 | $1.30 | $1.95 |
| 1080p | Yes | $0.54 | $0.675 | $1.35 | $2.025 |
Best Use Cases
- Turbo text-to-video generation — Create short videos directly from natural-language prompts.
- Reference-to-video workflows — Use reference images, videos, or audio to guide the generated video.
- Cinematic short clips — Generate videos with camera movement, atmosphere, and synchronized audio.
- Character and style consistency — Use reference images to guide characters, visual style, or scene composition.
- Motion and audio guidance — Use reference videos or audio to guide motion direction or sound style.
- Social media content — Generate vertical, square, landscape, or ultra-wide video formats.
- Creative prototyping — Test scene ideas, visual concepts, and motion directions quickly.
Pro Tips
- Be specific about scene, action, camera movement, lighting, and mood.
- Use reference images when character, style, or composition consistency matters.
- Use reference videos when motion direction is important.
- Use reference audio when the generated video should follow a specific audio style or sound direction.
- Keep
generate_audioenabled when you want the output video to include synchronized native audio. - Choose an aspect ratio based on the final format, such as
9:16for vertical content or16:9for widescreen video.
Related Models
- ByteDance Seedance 2.0 Mini Text-to-Video — Generate video directly from text prompts.
- ByteDance Seedance 2.0 Mini Image-to-Video — Generate video from a start image and prompt.
---
**Notice:** Use `@image1`, `@image2`, `@audio1`, etc. to reference your uploaded assets. The references will stay as plain text—don't worry.
## Authentication
For authentication details, please refer to the [Authentication Guide](/api-authentication).
## API Endpoints
### Submit Task & Query Result
<ApiTabs submitUrl={model.submitUrl} resultUrl={model.resultUrl} payload={model.defaultValues} />
## Parameters
### Task Submission Parameters
#### Request Parameters
<RequestParams params={model.params} />
#### Response Parameters
<SubmitResponse />
#### Result Request Parameters
| Parameter | Type | Required | Default | Description |
|-----------|------|----------|---------|-------------|
| id | string | Yes | - | Task ID |
#### Result Response Parameters
| Parameter | Type | Description |
|-----------|------|-------------|
| code | integer | HTTP status code (e.g., 200 for success) |
| message | string | Status message (e.g., "success") |
| data | object | The prediction data object containing all details |
| data.id | string | Unique identifier for the prediction |
| data.model | string | Model ID used for the prediction |
| data.outputs | array<string \| object> | Array of generated outputs (empty when status is not completed). Items are usually URL strings, but may be text strings or structured result objects, depending on the model. |
| data.urls | object | Object containing related API endpoints |
| data.urls.get | string | URL to poll for the prediction result |
| data.status | string | Status: `created`, `processing`, `completed`, or `failed` |
| data.created_at | string | ISO timestamp of when the request was created |
| data.error | string | Error message (empty if no error occurred) |
| data.timings | object | Object containing timing details |
| data.timings.inference | integer | Inference time in milliseconds |