Elevenlabs Multilingual V2 API Documentation

Elevenlabs Multilingual V2 API Documentation

Playground

Try it on WaveSpeedAI!

ElevenLabs Multilingual V2 is a multilingual text-to-speech model; cost $0.1 per 1000 characters. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

Features

Multilingual V2 converts written text into natural, expressive speech across multiple languages. It delivers clear pronunciation, smooth pacing, and lifelike tone—ideal for voiceovers, narration, learning content, product videos, and global customer support. See the list here.


Key Features

  • High naturalness with humanlike intonation and timing
  • Strong multilingual support and improved accent handling
  • Tunable delivery via similarity and stability
  • Speaker Boost — enhances similarity to the original speaker’s voice.

Pricing

Billed by the exact character count of the input text, prorated — no rounding up to 1,000.

  • Rate: $0.20 per 1,000 characters ($200 per 1M characters)
  • No minimum charge; a 100-character request costs $0.020
CharactersCost
100$0.020
500$0.10
1,000$0.20
10,000$2.00

How to Use

  1. Enter your script in the text field.
  2. Choose a voice_id from the built-in catalog or your custom voices. See the voice list for options.
  3. Optional controls • similarity: 0–1 (higher = closer to the base voice timbre) • stability: 0–1 (higher = more consistent delivery) • use_speaker_boost: enhances similarity to the original speaker’s voice
  4. Click Run to synthesize and preview your audio.

Notes

  • Use clear punctuation and split very long text into shorter segments for the most stable prosody.

  • voice_id must be valid; if you see a voice-ID error, pick one from the official list linked above.

  • Speaker Boost may slightly increase generation latency.

  • Each request accepts 1–10,000 characters. Split longer text into separate requests.

  • Use <break time="1.5s" /> to add a pause of up to 3 seconds.

Authentication

For authentication details, please refer to the Authentication Guide.

API Endpoints

Submit Task & Query Result

set -euo pipefail

export WAVESPEED_API_KEY="your-api-key"

REQUEST_BODY=$(cat <<'JSON'
{
  "text": "A clear example input",
  "voice_id": "Alicia",
  "similarity": 1,
  "stability": 0.5,
  "use_speaker_boost": true
}
JSON
)

# 1. Submit the prediction.
SUBMIT_RESPONSE=$(curl --silent --show-error --fail-with-body \
  -X POST "https://api.wavespeed.ai/api/v3/elevenlabs/multilingual-v2" \
  -H "Authorization: Bearer ${WAVESPEED_API_KEY}" \
  -H "Content-Type: application/json" \
  -d "${REQUEST_BODY}")

TASK=$(printf '%s' "${SUBMIT_RESPONSE}" | jq 'if type == "object" and has("data") then .data else . end')
PREDICTION_ID=$(printf '%s' "${TASK}" | jq -r '.id // empty')
if [ -z "${PREDICTION_ID}" ]; then
  printf 'Submission response did not contain a prediction id
' >&2
  exit 1
fi
RESULT_URL="https://api.wavespeed.ai/api/v3/predictions/${PREDICTION_ID}/result"

# 2. Poll until the prediction finishes.
while true; do
  RESPONSE=$(curl --silent --show-error --fail-with-body \
    "${RESULT_URL}" \
    -H "Authorization: Bearer ${WAVESPEED_API_KEY}")
  RESULT=$(printf '%s' "${RESPONSE}" | jq 'if type == "object" and has("data") then .data else . end')
  STATUS=$(printf '%s' "${RESULT}" | jq -r '.status // empty')

  case "${STATUS}" in
    completed) printf '%s\n' "${RESULT}" | jq '.outputs'; break ;;
    failed|cancelled|timeout|deleted) printf '%s\n' "${RESULT}" | jq . >&2; exit 1 ;;
    *) sleep 2 ;;
  esac
done

Parameters

Task Submission Parameters

Request Parameters

ParameterTypeRequiredDefaultRangeDescription
textstringYes-1 ~ 10000 charactersText to convert to speech. Maximum 10,000 characters per request. Use <break time="1.5s" /> to add a pause of up to 3 seconds.
voice_idstringYesAlicia-The voice to use for speech generation. Choose a preset voice name, or enter any ElevenLabs voice ID.
similaritynumberNo10 ~ 1High enhancement boosts overall voice clarity and target speaker similarity. Very high values can cause artifacts, so adjusting this setting to find the optimal value is encouraged.
stabilitynumberNo0.50 ~ 1Voice stability (0-1) Default value: 0.5
use_speaker_boostbooleanNotrue-Enhances similarity to the original speaker's voice. May slightly increase generation latency.

Response Parameters

ParameterTypeDescription
codeintegerHTTP status code (e.g., 200 for success)
messagestringStatus message (e.g., “success”)
data.idstringUnique identifier for the prediction, Task Id
data.modelstringModel ID used for the prediction
data.outputsarrayOutput values, usually URL strings; some models return text strings or structured result objects (empty when status is not completed)
data.urlsobjectObject containing related API endpoints
data.statusstringTask status. completed is successful; failed, cancelled, timeout, and deleted are failure terminal statuses.
data.created_atstringISO timestamp of when the request was created (e.g., “2023-04-01T12:34:56.789Z”)
data.errorstringError message (empty if no error occurred)
data.timingsobjectObject containing timing details
data.timings.inferenceintegerInference time in milliseconds

Result Request Parameters

ParameterTypeRequiredDefaultDescription
idstringYes-Task ID

Result Response Parameters

ParameterTypeDescription
codeintegerHTTP status code (e.g., 200 for success)
messagestringStatus message (e.g., “success”)
dataobjectThe prediction data object containing all details
data.idstringUnique identifier for the prediction
data.modelstringModel ID used for the prediction
data.outputsarray<string | object>Array of generated outputs (empty when status is not completed). Items are usually URL strings, but may be text strings or structured result objects, depending on the model.
data.urlsobjectObject containing related API endpoints
data.statusstringStatus: completed is successful; failed, cancelled, timeout, and deleted are failure terminal statuses
data.created_atstringISO timestamp of when the request was created
data.errorstringError message (empty if no error occurred)
data.timingsobjectObject containing timing details
data.timings.inferenceintegerInference time in milliseconds
© 2026 WaveSpeedAI. All rights reserved.