Pruna AI P Video 2 Pro Text To Video API Documentation

Pruna AI P Video 2 Pro Text To Video API Documentation

Playground

Try it on WaveSpeedAI!

Pruna P-Video-2-Pro Text-to-Video generates videos from text prompts at 480P / 768P output with generated audio, supporting prompt-driven video creation for creative clips, social content, ads, and production workflows. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

Features

Pruna P-Video-2 Pro Text-to-Video generates videos directly from text prompts with generated audio. Describe the scene, subject, motion, camera movement, and visual style, then choose duration, aspect ratio, resolution, turbo mode, and prompt upsampling settings.


Why Choose This?

  • Text-to-video generation
    Generate videos directly from natural-language prompts.

  • Generated audio included
    Output video includes generated audio without requiring separate audio input.

  • 480p and 768p output
    Use 480p for lower-cost generation or 768p for higher-resolution output.

  • Speed and quality modes
    Choose mode=speed for faster generation or mode=quality for higher-quality generation. Default: speed.

  • Prompt upsampling control
    Choose off, turbo, or max prompt expansion depending on how much prompt enhancement you need.

  • Multiple aspect ratios
    Supports landscape, vertical, square, classic, and portrait video formats.


Parameters

ParameterRequiredDescription
promptYesText prompt describing the video to generate, including subject, action, camera movement, scene, mood, and visual style.
durationNoVideo duration in seconds. Range: 5–15. Default: 5.
aspect_ratioNoOutput aspect ratio: 16:9, 9:16, 4:3, 3:4, 3:2, 2:3, or 1:1. Default: 16:9.
resolutionNoOutput video resolution: 480p or 768p. Default: 768p.
modeNoGeneration mode: speed for faster generation or quality for higher-quality generation. Default: speed. Independent of prompt_upsampler.
prompt_upsamplerNoPrompt expansion mode: off, turbo, or max. Default: turbo. Independent of mode; does not change the price.
seedNoOptional random seed for reproducible generation. Omit for an upstream-selected random seed.

How to Use

  1. Write a prompt — Describe the scene, subject, action, camera movement, lighting, mood, and visual style.
  2. Set duration — Choose a video length from 5 to 15 seconds.
  3. Choose aspect ratio — Select the layout that matches your target format.
  4. Choose resolution — Use 480p for lower-cost generation or 768p for higher-resolution output.
  5. Choose generation mode — Select speed for faster generation or quality for higher-quality generation. Pricing depends on the selected mode and resolution.
  6. Set prompt upsampling — Use off, turbo, or max depending on how much prompt expansion you want.
  7. Set seed optional — Use a fixed seed for reproducible results, or omit it for random generation.
  8. Submit — Generate the final video with audio.

Pricing

Pricing is based on the requested duration, selected resolution, and generation mode.

ResolutionSpeed / secondQuality / second
480p$0.02$0.04
768p$0.035$0.075

Example Costs

Duration480p Speed480p Quality768p Speed768p Quality
5s$0.10$0.20$0.175$0.375
10s$0.20$0.40$0.35$0.75
15s$0.30$0.60$0.525$1.125

The default settings are 768p, speed, and 5 seconds, costing $0.175 per request.

Billing uses the requested duration, not the measured duration of the generated output. Generated audio is included. Prompt upsampling, seed selection, and supported image-conditioning inputs do not add separate charges.


Best Use Cases

  • Text-to-video generation — Create videos directly from written scene descriptions.
  • Cinematic concepts — Generate short scenes with camera movement, lighting, and visual style.
  • Social video content — Create vertical, square, landscape, or portrait-format clips.
  • Marketing and product videos — Generate short-form promotional clips from text prompts.
  • Creative prototyping — Test motion, pacing, style, and audio direction before final production.
  • Prompt iteration — Compare speed and quality generation modes and independently adjust prompt upsampling.

Pro Tips

  • Write prompts that describe visible motion, not just the visual style.
  • Include subject, action, environment, camera movement, lighting, mood, and scene progression.
  • Keep the prompt focused on one main scene or action for stronger motion coherence.
  • Use mode=speed for faster, lower-cost iteration, or mode=quality when prioritizing output quality.
  • Use prompt_upsampler=off when you want the model to follow your original prompt more directly.
  • Use prompt_upsampler=max when the prompt is short and needs stronger expansion.
  • Use 480p for lower-cost tests and 768p for higher-resolution output.
  • Set a fixed seed when comparing prompt or parameter changes.

Notes

  • prompt is required.
  • duration supports values from 5 to 15 seconds.
  • Generated audio is included in the output.
  • Image conditioning and last-frame guidance are not exposed on this text-to-video product.

Authentication

For authentication details, please refer to the Authentication Guide.

API Endpoints

Submit Task & Query Result

set -euo pipefail

export WAVESPEED_API_KEY="your-api-key"

REQUEST_BODY=$(cat <<'JSON'
{
  "prompt": "A cinematic ocean wave at sunrise, highly detailed",
  "duration": 5,
  "aspect_ratio": "16:9",
  "resolution": "768p",
  "mode": "speed",
  "prompt_upsampler": "turbo"
}
JSON
)

# 1. Submit the prediction.
SUBMIT_RESPONSE=$(curl --silent --show-error --fail-with-body \
  -X POST "https://api.wavespeed.ai/api/v3/pruna-ai/p-video-2-pro/text-to-video" \
  -H "Authorization: Bearer ${WAVESPEED_API_KEY}" \
  -H "Content-Type: application/json" \
  -d "${REQUEST_BODY}")

TASK=$(printf '%s' "${SUBMIT_RESPONSE}" | jq 'if type == "object" and has("data") then .data else . end')
PREDICTION_ID=$(printf '%s' "${TASK}" | jq -r '.id // empty')
if [ -z "${PREDICTION_ID}" ]; then
  printf 'Submission response did not contain a prediction id
' >&2
  exit 1
fi
RESULT_URL="https://api.wavespeed.ai/api/v3/predictions/${PREDICTION_ID}/result"

# 2. Poll until the prediction finishes.
while true; do
  RESPONSE=$(curl --silent --show-error --fail-with-body \
    "${RESULT_URL}" \
    -H "Authorization: Bearer ${WAVESPEED_API_KEY}")
  RESULT=$(printf '%s' "${RESPONSE}" | jq 'if type == "object" and has("data") then .data else . end')
  STATUS=$(printf '%s' "${RESULT}" | jq -r '.status // empty')

  case "${STATUS}" in
    completed) printf '%s\n' "${RESULT}" | jq '.outputs'; break ;;
    failed|cancelled|timeout|deleted) printf '%s\n' "${RESULT}" | jq . >&2; exit 1 ;;
    *) sleep 2 ;;
  esac
done

Parameters

Task Submission Parameters

Request Parameters

ParameterTypeRequiredDefaultRangeDescription
promptstringYes-Text prompt describing the video to generate.
durationintegerNo55 ~ 15Video duration in seconds, from 5 to 15.
aspect_ratiostringNo16:916:9, 9:16, 4:3, 3:4, 3:2, 2:3, 1:1Aspect ratio of the generated video.
resolutionstringNo768p480p, 768pOutput video resolution.
modestringNospeedspeed, qualityGeneration mode: speed for faster generation, or quality for higher-quality generation. Independent of prompt_upsampler.
prompt_upsamplerstringNoturbooff, turbo, maxPrompt expansion mode. Independent of the speed or quality generation mode.
seedintegerNo--Optional random seed for reproducible generation. Omit for an upstream-selected random seed.

Response Parameters

ParameterTypeDescription
codeintegerHTTP status code (e.g., 200 for success)
messagestringStatus message (e.g., “success”)
data.idstringUnique identifier for the prediction, Task Id
data.modelstringModel ID used for the prediction
data.outputsarrayOutput values, usually URL strings; some models return text strings or structured result objects (empty when status is not completed)
data.urlsobjectObject containing related API endpoints
data.statusstringTask status. completed is successful; failed, cancelled, timeout, and deleted are failure terminal statuses.
data.created_atstringISO timestamp of when the request was created (e.g., “2023-04-01T12:34:56.789Z”)
data.errorstringError message (empty if no error occurred)
data.timingsobjectObject containing timing details
data.timings.inferenceintegerInference time in milliseconds

Result Request Parameters

ParameterTypeRequiredDefaultDescription
idstringYes-Task ID

Result Response Parameters

ParameterTypeDescription
codeintegerHTTP status code (e.g., 200 for success)
messagestringStatus message (e.g., “success”)
dataobjectThe prediction data object containing all details
data.idstringUnique identifier for the prediction
data.modelstringModel ID used for the prediction
data.outputsarray<string | object>Array of generated outputs (empty when status is not completed). Items are usually URL strings, but may be text strings or structured result objects, depending on the model.
data.urlsobjectObject containing related API endpoints
data.statusstringStatus: completed is successful; failed, cancelled, timeout, and deleted are failure terminal statuses
data.created_atstringISO timestamp of when the request was created
data.errorstringError message (empty if no error occurred)
data.timingsobjectObject containing timing details
data.timings.inferenceintegerInference time in milliseconds
© 2026 WaveSpeedAI. All rights reserved.