VEED Subtitles is a fast AI video subtitle generation model that adds styled captions to videos using automatic transcription or supplied SRT files and subtitle content. Ready-to-use REST inference API for subtitled MP4 generation, social media videos, creator content, marketing clips, accessibility workflows, video localization, and professional captioning with simple integration, no coldstarts, and affordable pricing.
Idle
$0.11per run·~90 / $10
VEED Subtitles adds styled subtitles to videos using automatic transcription or an optional SRT subtitle file. Choose subtitle presets, positions, shadow styles, languages, and output resolution tiers.
Automatic transcription
Upload a video and automatically generate subtitles.
Optional SRT subtitles
Upload an SRT file or provide raw SRT content when exact subtitle text and timing are required.
Multiple subtitle styles
Choose from standard and dynamic subtitle presets.
Position and shadow controls
Adjust subtitle placement and shadow strength for better readability.
Multiple resolution tiers
Generate subtitled videos in 480p, 720p, or 1080p.
| Parameter | Required | Description |
|---|---|---|
| video | Yes | Input video. |
| target_resolution | No | Output resolution tier: 480p, 720p, or 1080p. Default: 720p. |
| preset | No | Subtitle visual preset. Default: simple. Standard and dynamic presets are supported. Dynamic presets use a higher pricing tier. |
| position | No | Subtitle position: top, center, or bottom. Default: bottom. |
| shadow | No | Text shadow strength: none, min, mid, or max. Default: mid. |
| language | No | Optional transcription language locale, such as en-US or zh-CN. Leave empty to automatically detect the language. |
| file | No | Optional SRT subtitle file. Do not use together with srt_content. |
| srt_content | No | Optional raw SRT subtitle content. Do not use together with file. |
480p, 720p, or 1080p.file or enter raw srt_content.Pricing depends on input video duration, selected resolution, and whether the selected subtitle preset is standard or dynamic.
480p and 720p use the base resolution multiplier1080p uses a 2× resolution multiplierposition, shadow, language, file, and srt_content do not affect pricing| Billed Duration | 480p / 720p Standard | 480p / 720p Dynamic | 1080p Standard | 1080p Dynamic |
|---|---|---|---|---|
| 60 seconds | $0.11 | $0.22 | $0.22 | $0.44 |
| 90 seconds | $0.165 | $0.33 | $0.33 | $0.66 |
| 120 seconds | $0.22 | $0.44 | $0.44 | $0.88 |
language empty when you want automatic language detection.file or srt_content, not both.1080p doubles the resolution price.1080p with a dynamic preset results in a 4× total multiplier.video is required.target_resolution defaults to 720p.preset defaults to simple.position defaults to bottom.shadow defaults to mid.file and srt_content in the same request is not supported.Grab a WaveSpeedAI API key, then call POST https://api.wavespeed.ai/api/v3/veed/subtitles with your input as JSON. The endpoint returns a prediction id. Start polling the result endpoint around every 2 seconds, increase the interval for long-running tasks, and stop on any terminal status. On completed, read output values from data.outputs. Examples for Subtitles below.
set -euo pipefail
: "${WAVESPEED_API_KEY:?Set WAVESPEED_API_KEY}"
REQUEST_BODY=$(cat <<'JSON'
{
"video": "https://interactive-examples.mdn.mozilla.net/media/cc0-videos/flower.mp4",
"target_resolution": "720p",
"preset": "simple",
"position": "bottom",
"shadow": "mid",
"language": "af-ZA"
}
JSON
)
# 1. Submit the prediction.
SUBMIT_RESPONSE=$(curl --silent --show-error --fail-with-body \
-X POST "https://api.wavespeed.ai/api/v3/veed/subtitles" \
-H "Content-Type: application/json" \
-H "Authorization: Bearer $WAVESPEED_API_KEY" \
-d "$REQUEST_BODY")
TASK=$(printf '%s' "$SUBMIT_RESPONSE" | jq 'if has("data") then .data else . end')
PREDICTION_ID=$(printf '%s' "$TASK" | jq -r '.id')
if [ -z "$PREDICTION_ID" ] || [ "$PREDICTION_ID" = "null" ]; then
printf 'Submission response did not contain a prediction id
' >&2
exit 1
fi
RESULT_URL="https://api.wavespeed.ai/api/v3/predictions/$PREDICTION_ID/result"
# 2. Poll until the prediction finishes.
while true; do
RESPONSE=$(curl --silent --show-error --fail-with-body "$RESULT_URL" \
-H "Authorization: Bearer $WAVESPEED_API_KEY")
RESULT=$(printf '%s' "$RESPONSE" | jq 'if has("data") then .data else . end')
STATUS=$(printf '%s' "$RESULT" | jq -r '.status')
case "$STATUS" in
completed) printf '%s\n' "$RESULT" | jq '.outputs'; break ;;
failed|cancelled|timeout|deleted) printf '%s\n' "$RESULT" | jq . >&2; exit 1 ;;
*) sleep 2 ;;
esac
doneconst submitUrl = "https://api.wavespeed.ai/api/v3/veed/subtitles";
const apiKey = process.env.WAVESPEED_API_KEY;
if (!apiKey) throw new Error('Set WAVESPEED_API_KEY');
async function requestJson(url, options = {}) {
const response = await fetch(url, options);
if (!response.ok) throw new Error(await response.text());
return response.json();
}
// 1. Submit the prediction.
const body = await requestJson(submitUrl, {
method: "POST",
headers: {
"Authorization": `Bearer ${apiKey}`,
"Content-Type": "application/json",
},
body: JSON.stringify({
"video": "https://interactive-examples.mdn.mozilla.net/media/cc0-videos/flower.mp4",
"target_resolution": "720p",
"preset": "simple",
"position": "bottom",
"shadow": "mid",
"language": "af-ZA"
}),
});
const task = body.data ?? body;
if (!task.id) throw new Error("Submission response did not contain a prediction id");
const resultUrl = `https://api.wavespeed.ai/api/v3/predictions/${task.id}/result`;
// 2. Poll until the prediction finishes.
while (true) {
const resultBody = await requestJson(resultUrl, {
headers: { "Authorization": `Bearer ${apiKey}` },
});
const result = resultBody.data ?? resultBody;
if (result.status === "completed") {
console.log(result.outputs);
break;
}
if (["failed", "cancelled", "timeout", "deleted"].includes(result.status)) throw new Error(JSON.stringify(result));
await new Promise(resolve => setTimeout(resolve, 2000));
}import json
import os
import time
from urllib.request import Request, urlopen
api_key = os.environ["WAVESPEED_API_KEY"]
headers = {"Authorization": f"Bearer {api_key}", "Content-Type": "application/json"}
payload = {
"video": "https://interactive-examples.mdn.mozilla.net/media/cc0-videos/flower.mp4",
"target_resolution": "720p",
"preset": "simple",
"position": "bottom",
"shadow": "mid",
"language": "af-ZA"
}
def request_json(url, data=None):
request = Request(url, data=data, headers=headers, method="POST" if data else "GET")
with urlopen(request) as response:
return json.load(response)
# 1. Submit the prediction.
body = request_json("https://api.wavespeed.ai/api/v3/veed/subtitles", json.dumps(payload).encode())
task = body.get("data", body)
if not task.get("id"):
raise RuntimeError("Submission response did not contain a prediction id")
result_url = f"https://api.wavespeed.ai/api/v3/predictions/{task['id']}/result"
# 2. Poll until the prediction finishes.
while True:
result_body = request_json(result_url)
result = result_body.get("data", result_body)
status = result.get("status")
if status == "completed":
print(result.get("outputs", []))
break
if status in {"failed", "cancelled", "timeout", "deleted"}:
raise RuntimeError(result)
time.sleep(2)Subtitles is a Veed model for video editing, exposed as a REST API on WaveSpeedAI. VEED Subtitles is a fast AI video subtitle generation model that adds styled captions to videos using automatic transcription or supplied SRT files and subtitle content. Ready-to-use REST inference API for subtitled MP4 generation, social media videos, creator content, marketing clips, accessibility workflows, video localization, and professional captioning with simple integration, no coldstarts, and affordable pricing. You can call it programmatically or try it from the playground above.
POST your input parameters to the model's REST endpoint (shown in the API tab of this playground) with your WaveSpeedAI API key in the Authorization header. Submission returns a prediction ID. Poll the result endpoint starting around every 2 seconds, increase the interval for long-running tasks, and stop on any terminal status. The playground generates Python, JavaScript, and cURL examples for submitting requests and polling results. Full request/response shape is documented at https://wavespeed.ai/docs/docs-api/veed/veed-subtitles.
Subtitles starts at $0.11 per run. That figure is the base price — the final charge scales with the parameters you set in the form (output size, length, count, references, or whatever knobs this model exposes), so a higher-quality or larger output costs more than a minimal one. The exact cost for your current input is shown live next to the Generate button before you submit, and the actual per-call charge is recorded on the prediction afterwards.
Key inputs: `video`, `file`, `language`, `position`, `preset`, `shadow`. The full JSON schema (types, defaults, allowed values) is rendered above the Generate button and mirrored in the API reference at https://wavespeed.ai/docs/docs-api/veed/veed-subtitles.
Reported generation time on WaveSpeedAI is around 45 seconds per request. This is an estimate, not a latency guarantee; queue time and input settings can change the total wait. live status is visible in the prediction record.
Commercial usage rights depend on the model's license, set by its provider (Veed). Check the provider's applicable terms and WaveSpeedAI's Terms of Service before commercial use.