Pruna AI P Video 2 Pro Text To Video API Documentation
Playground
Try it on WaveSpeedAI!Pruna P-Video-2-Pro Text-to-Video generates videos from text prompts at 480P / 768P output with generated audio, supporting prompt-driven video creation for creative clips, social content, ads, and production workflows. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
Features
Pruna P-Video-2 Pro Text-to-Video generates videos directly from text prompts with generated audio. Describe the scene, subject, motion, camera movement, and visual style, then choose duration, aspect ratio, resolution, turbo mode, and prompt upsampling settings.
Why Choose This?
-
Text-to-video generation
Generate videos directly from natural-language prompts. -
Generated audio included
Output video includes generated audio without requiring separate audio input. -
480p and 768p output
Use480pfor lower-cost generation or768pfor higher-resolution output. -
Speed and quality modes
Choosemode=speedfor faster generation ormode=qualityfor higher-quality generation. Default:speed. -
Prompt upsampling control
Chooseoff,turbo, ormaxprompt expansion depending on how much prompt enhancement you need. -
Multiple aspect ratios
Supports landscape, vertical, square, classic, and portrait video formats.
Parameters
| Parameter | Required | Description |
|---|---|---|
| prompt | Yes | Text prompt describing the video to generate, including subject, action, camera movement, scene, mood, and visual style. |
| duration | No | Video duration in seconds. Range: 5–15. Default: 5. |
| aspect_ratio | No | Output aspect ratio: 16:9, 9:16, 4:3, 3:4, 3:2, 2:3, or 1:1. Default: 16:9. |
| resolution | No | Output video resolution: 480p or 768p. Default: 768p. |
| mode | No | Generation mode: speed for faster generation or quality for higher-quality generation. Default: speed. Independent of prompt_upsampler. |
| prompt_upsampler | No | Prompt expansion mode: off, turbo, or max. Default: turbo. Independent of mode; does not change the price. |
| seed | No | Optional random seed for reproducible generation. Omit for an upstream-selected random seed. |
How to Use
- Write a prompt — Describe the scene, subject, action, camera movement, lighting, mood, and visual style.
- Set duration — Choose a video length from
5to15seconds. - Choose aspect ratio — Select the layout that matches your target format.
- Choose resolution — Use
480pfor lower-cost generation or768pfor higher-resolution output. - Choose generation mode — Select
speedfor faster generation orqualityfor higher-quality generation. Pricing depends on the selected mode and resolution. - Set prompt upsampling — Use
off,turbo, ormaxdepending on how much prompt expansion you want. - Set seed optional — Use a fixed seed for reproducible results, or omit it for random generation.
- Submit — Generate the final video with audio.
Pricing
Pricing is based on the requested duration, selected resolution, and generation mode.
| Resolution | Speed / second | Quality / second |
|---|---|---|
| 480p | $0.02 | $0.04 |
| 768p | $0.035 | $0.075 |
Example Costs
| Duration | 480p Speed | 480p Quality | 768p Speed | 768p Quality |
|---|---|---|---|---|
| 5s | $0.10 | $0.20 | $0.175 | $0.375 |
| 10s | $0.20 | $0.40 | $0.35 | $0.75 |
| 15s | $0.30 | $0.60 | $0.525 | $1.125 |
The default settings are 768p, speed, and 5 seconds, costing $0.175 per request.
Billing uses the requested duration, not the measured duration of the generated output. Generated audio is included. Prompt upsampling, seed selection, and supported image-conditioning inputs do not add separate charges.
Best Use Cases
- Text-to-video generation — Create videos directly from written scene descriptions.
- Cinematic concepts — Generate short scenes with camera movement, lighting, and visual style.
- Social video content — Create vertical, square, landscape, or portrait-format clips.
- Marketing and product videos — Generate short-form promotional clips from text prompts.
- Creative prototyping — Test motion, pacing, style, and audio direction before final production.
- Prompt iteration — Compare
speedandqualitygeneration modes and independently adjust prompt upsampling.
Pro Tips
- Write prompts that describe visible motion, not just the visual style.
- Include subject, action, environment, camera movement, lighting, mood, and scene progression.
- Keep the prompt focused on one main scene or action for stronger motion coherence.
- Use
mode=speedfor faster, lower-cost iteration, ormode=qualitywhen prioritizing output quality. - Use
prompt_upsampler=offwhen you want the model to follow your original prompt more directly. - Use
prompt_upsampler=maxwhen the prompt is short and needs stronger expansion. - Use
480pfor lower-cost tests and768pfor higher-resolution output. - Set a fixed
seedwhen comparing prompt or parameter changes.
Notes
promptis required.durationsupports values from5to15seconds.- Generated audio is included in the output.
- Image conditioning and last-frame guidance are not exposed on this text-to-video product.
Related Models
- Pruna P-Video-2 Pro Text-to-Video — Generate Pro text-to-video outputs with generated audio.
- Pruna P-Video-2 Pro Image-to-Video — Generate Pro videos from an input image and motion prompt.
- Pruna P-Video-2 Text-to-Video — Generate videos directly from text prompts.
- Pruna P-Video-2 Image-to-Video — Generate videos from an input image and motion prompt.
Authentication
For authentication details, please refer to the Authentication Guide.
API Endpoints
Submit Task & Query Result
set -euo pipefail
export WAVESPEED_API_KEY="your-api-key"
REQUEST_BODY=$(cat <<'JSON'
{
"prompt": "A cinematic ocean wave at sunrise, highly detailed",
"duration": 5,
"aspect_ratio": "16:9",
"resolution": "768p",
"mode": "speed",
"prompt_upsampler": "turbo"
}
JSON
)
# 1. Submit the prediction.
SUBMIT_RESPONSE=$(curl --silent --show-error --fail-with-body \
-X POST "https://api.wavespeed.ai/api/v3/pruna-ai/p-video-2-pro/text-to-video" \
-H "Authorization: Bearer ${WAVESPEED_API_KEY}" \
-H "Content-Type: application/json" \
-d "${REQUEST_BODY}")
TASK=$(printf '%s' "${SUBMIT_RESPONSE}" | jq 'if type == "object" and has("data") then .data else . end')
PREDICTION_ID=$(printf '%s' "${TASK}" | jq -r '.id // empty')
if [ -z "${PREDICTION_ID}" ]; then
printf 'Submission response did not contain a prediction id
' >&2
exit 1
fi
RESULT_URL="https://api.wavespeed.ai/api/v3/predictions/${PREDICTION_ID}/result"
# 2. Poll until the prediction finishes.
while true; do
RESPONSE=$(curl --silent --show-error --fail-with-body \
"${RESULT_URL}" \
-H "Authorization: Bearer ${WAVESPEED_API_KEY}")
RESULT=$(printf '%s' "${RESPONSE}" | jq 'if type == "object" and has("data") then .data else . end')
STATUS=$(printf '%s' "${RESULT}" | jq -r '.status // empty')
case "${STATUS}" in
completed) printf '%s\n' "${RESULT}" | jq '.outputs'; break ;;
failed|cancelled|timeout|deleted) printf '%s\n' "${RESULT}" | jq . >&2; exit 1 ;;
*) sleep 2 ;;
esac
doneParameters
Task Submission Parameters
Request Parameters
| Parameter | Type | Required | Default | Range | Description |
|---|---|---|---|---|---|
| prompt | string | Yes | - | Text prompt describing the video to generate. | |
| duration | integer | No | 5 | 5 ~ 15 | Video duration in seconds, from 5 to 15. |
| aspect_ratio | string | No | 16:9 | 16:9, 9:16, 4:3, 3:4, 3:2, 2:3, 1:1 | Aspect ratio of the generated video. |
| resolution | string | No | 768p | 480p, 768p | Output video resolution. |
| mode | string | No | speed | speed, quality | Generation mode: speed for faster generation, or quality for higher-quality generation. Independent of prompt_upsampler. |
| prompt_upsampler | string | No | turbo | off, turbo, max | Prompt expansion mode. Independent of the speed or quality generation mode. |
| seed | integer | No | - | - | Optional random seed for reproducible generation. Omit for an upstream-selected random seed. |
Response Parameters
| Parameter | Type | Description |
|---|---|---|
| code | integer | HTTP status code (e.g., 200 for success) |
| message | string | Status message (e.g., “success”) |
| data.id | string | Unique identifier for the prediction, Task Id |
| data.model | string | Model ID used for the prediction |
| data.outputs | array | Output values, usually URL strings; some models return text strings or structured result objects (empty when status is not completed) |
| data.urls | object | Object containing related API endpoints |
| data.status | string | Task status. completed is successful; failed, cancelled, timeout, and deleted are failure terminal statuses. |
| data.created_at | string | ISO timestamp of when the request was created (e.g., “2023-04-01T12:34:56.789Z”) |
| data.error | string | Error message (empty if no error occurred) |
| data.timings | object | Object containing timing details |
| data.timings.inference | integer | Inference time in milliseconds |
Result Request Parameters
| Parameter | Type | Required | Default | Description |
|---|---|---|---|---|
| id | string | Yes | - | Task ID |
Result Response Parameters
| Parameter | Type | Description |
|---|---|---|
| code | integer | HTTP status code (e.g., 200 for success) |
| message | string | Status message (e.g., “success”) |
| data | object | The prediction data object containing all details |
| data.id | string | Unique identifier for the prediction |
| data.model | string | Model ID used for the prediction |
| data.outputs | array<string | object> | Array of generated outputs (empty when status is not completed). Items are usually URL strings, but may be text strings or structured result objects, depending on the model. |
| data.urls | object | Object containing related API endpoints |
| data.status | string | Status: completed is successful; failed, cancelled, timeout, and deleted are failure terminal statuses |
| data.created_at | string | ISO timestamp of when the request was created |
| data.error | string | Error message (empty if no error occurred) |
| data.timings | object | Object containing timing details |
| data.timings.inference | integer | Inference time in milliseconds |