Elevenlabs Multilingual V2 API Documentation
Playground
Try it on WaveSpeedAI!ElevenLabs Multilingual V2 is a multilingual text-to-speech model; cost $0.1 per 1000 characters. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
Features
Multilingual V2 converts written text into natural, expressive speech across multiple languages. It delivers clear pronunciation, smooth pacing, and lifelike tone—ideal for voiceovers, narration, learning content, product videos, and global customer support. See the list here.
Key Features
- High naturalness with humanlike intonation and timing
- Strong multilingual support and improved accent handling
- Tunable delivery via similarity and stability
- Speaker Boost — enhances similarity to the original speaker’s voice.
Pricing
Billed by the exact character count of the input text, prorated — no rounding up to 1,000.
- Rate: $0.20 per 1,000 characters ($200 per 1M characters)
- No minimum charge; a 100-character request costs $0.020
| Characters | Cost |
|---|---|
| 100 | $0.020 |
| 500 | $0.10 |
| 1,000 | $0.20 |
| 10,000 | $2.00 |
How to Use
- Enter your script in the text field.
- Choose a voice_id from the built-in catalog or your custom voices. See the voice list for options.
- Optional controls • similarity: 0–1 (higher = closer to the base voice timbre) • stability: 0–1 (higher = more consistent delivery) • use_speaker_boost: enhances similarity to the original speaker’s voice
- Click Run to synthesize and preview your audio.
Notes
-
Use clear punctuation and split very long text into shorter segments for the most stable prosody.
-
voice_id must be valid; if you see a voice-ID error, pick one from the official list linked above.
-
Speaker Boost may slightly increase generation latency.
-
Each request accepts 1–10,000 characters. Split longer text into separate requests.
-
Use
<break time="1.5s" />to add a pause of up to 3 seconds.
Authentication
For authentication details, please refer to the Authentication Guide.
API Endpoints
Submit Task & Query Result
set -euo pipefail
export WAVESPEED_API_KEY="your-api-key"
REQUEST_BODY=$(cat <<'JSON'
{
"text": "A clear example input",
"voice_id": "Alicia",
"similarity": 1,
"stability": 0.5,
"use_speaker_boost": true
}
JSON
)
# 1. Submit the prediction.
SUBMIT_RESPONSE=$(curl --silent --show-error --fail-with-body \
-X POST "https://api.wavespeed.ai/api/v3/elevenlabs/multilingual-v2" \
-H "Authorization: Bearer ${WAVESPEED_API_KEY}" \
-H "Content-Type: application/json" \
-d "${REQUEST_BODY}")
TASK=$(printf '%s' "${SUBMIT_RESPONSE}" | jq 'if type == "object" and has("data") then .data else . end')
PREDICTION_ID=$(printf '%s' "${TASK}" | jq -r '.id // empty')
if [ -z "${PREDICTION_ID}" ]; then
printf 'Submission response did not contain a prediction id
' >&2
exit 1
fi
RESULT_URL="https://api.wavespeed.ai/api/v3/predictions/${PREDICTION_ID}/result"
# 2. Poll until the prediction finishes.
while true; do
RESPONSE=$(curl --silent --show-error --fail-with-body \
"${RESULT_URL}" \
-H "Authorization: Bearer ${WAVESPEED_API_KEY}")
RESULT=$(printf '%s' "${RESPONSE}" | jq 'if type == "object" and has("data") then .data else . end')
STATUS=$(printf '%s' "${RESULT}" | jq -r '.status // empty')
case "${STATUS}" in
completed) printf '%s\n' "${RESULT}" | jq '.outputs'; break ;;
failed|cancelled|timeout|deleted) printf '%s\n' "${RESULT}" | jq . >&2; exit 1 ;;
*) sleep 2 ;;
esac
doneParameters
Task Submission Parameters
Request Parameters
| Parameter | Type | Required | Default | Range | Description |
|---|---|---|---|---|---|
| text | string | Yes | - | 1 ~ 10000 characters | Text to convert to speech. Maximum 10,000 characters per request. Use <break time="1.5s" /> to add a pause of up to 3 seconds. |
| voice_id | string | Yes | Alicia | - | The voice to use for speech generation. Choose a preset voice name, or enter any ElevenLabs voice ID. |
| similarity | number | No | 1 | 0 ~ 1 | High enhancement boosts overall voice clarity and target speaker similarity. Very high values can cause artifacts, so adjusting this setting to find the optimal value is encouraged. |
| stability | number | No | 0.5 | 0 ~ 1 | Voice stability (0-1) Default value: 0.5 |
| use_speaker_boost | boolean | No | true | - | Enhances similarity to the original speaker's voice. May slightly increase generation latency. |
Response Parameters
| Parameter | Type | Description |
|---|---|---|
| code | integer | HTTP status code (e.g., 200 for success) |
| message | string | Status message (e.g., “success”) |
| data.id | string | Unique identifier for the prediction, Task Id |
| data.model | string | Model ID used for the prediction |
| data.outputs | array | Output values, usually URL strings; some models return text strings or structured result objects (empty when status is not completed) |
| data.urls | object | Object containing related API endpoints |
| data.status | string | Task status. completed is successful; failed, cancelled, timeout, and deleted are failure terminal statuses. |
| data.created_at | string | ISO timestamp of when the request was created (e.g., “2023-04-01T12:34:56.789Z”) |
| data.error | string | Error message (empty if no error occurred) |
| data.timings | object | Object containing timing details |
| data.timings.inference | integer | Inference time in milliseconds |
Result Request Parameters
| Parameter | Type | Required | Default | Description |
|---|---|---|---|---|
| id | string | Yes | - | Task ID |
Result Response Parameters
| Parameter | Type | Description |
|---|---|---|
| code | integer | HTTP status code (e.g., 200 for success) |
| message | string | Status message (e.g., “success”) |
| data | object | The prediction data object containing all details |
| data.id | string | Unique identifier for the prediction |
| data.model | string | Model ID used for the prediction |
| data.outputs | array<string | object> | Array of generated outputs (empty when status is not completed). Items are usually URL strings, but may be text strings or structured result objects, depending on the model. |
| data.urls | object | Object containing related API endpoints |
| data.status | string | Status: completed is successful; failed, cancelled, timeout, and deleted are failure terminal statuses |
| data.created_at | string | ISO timestamp of when the request was created |
| data.error | string | Error message (empty if no error occurred) |
| data.timings | object | Object containing timing details |
| data.timings.inference | integer | Inference time in milliseconds |