50% di sconto sui modelli Vidu Q3 e Q3 Pro · Solo su WaveSpeedAI | 20 maggio – 2 giugno

Voice Changer

elevenlabs /

ElevenLabs Voice Changer transforms any audio into speech with a different voice while preserving the original speech patterns and timing. Ready-to-use REST inference API, best performance, no cold starts, affordable pricing.

audio-to-audio
Input

Trascina e rilascia o clicca per caricare

Remove background noise from the audio

Inattivo

$0.025per esecuzione·~40 / $1

EsempiVedi tutto

Modelli correlati

README

ElevenLabs Voice Changer

ElevenLabs Voice Changer transforms any audio into speech with a different voice. Upload your audio and select a target voice — the model converts the speech while preserving the original timing, emotion, and delivery. Built on ElevenLabs' industry-leading voice AI with best-in-class quality.

REST inference API, best performance, no cold starts, affordable pricing.

Why Choose This?

  • High-quality voice conversion Industry-leading voice transformation that maintains natural speech patterns and emotional delivery.

  • Multiple voice options Choose from a variety of pre-built voices to match your content needs.

  • Background noise removal Optional noise reduction to clean up audio before conversion.

  • Fast processing Optimized for quick turnaround with no cold starts.

  • Production-ready API Reliable REST endpoint with predictable per-minute pricing.

Parameters

ParameterRequiredDescription
audioYesSource audio file to transform (upload or URL)
voice_idNoTarget voice for conversion (default: Alice)
remove_background_noiseNoRemove background noise from the audio

How to Use

  1. Upload your audio — drag and drop, paste a URL, or record directly.
  2. Select voice — choose the target voice for conversion.
  3. Enable noise removal (optional) — check to clean up background noise.
  4. Run — submit and download the converted audio.

Pricing

DurationCost
Per minute$0.30
30 seconds$0.15
5 minutes$1.50

Best Use Cases

  • Content Creation — Change voices for podcasts, videos, or audiobooks.
  • Dubbing — Convert speech to different voices for localization.
  • Privacy — Anonymize voice recordings while preserving content.
  • Character Voices — Create distinct character voices for storytelling.
  • Accessibility — Convert speech to preferred voice styles.

Pro Tips

  • Use clean, high-quality source audio for best results.
  • Enable background noise removal if your source has ambient sounds.
  • Shorter clips process faster — split long audio for parallel processing.
  • Test with different voices to find the best match for your content.

Notes

  • Maximum audio duration is 10 minutes per job.
  • For longer content, split into segments and process separately.
  • Supported audio formats include MP3, WAV, and other common formats.

Related Models

  • ElevenLabs V3 — Generate speech from text with natural-sounding voices.
  • OpenAI Whisper — Transcribe audio to text with high accuracy.
Accessibilità:Questo sito web utilizza modelli di intelligenza artificiale forniti da terze parti.

Voice Changer API — Quick start

Grab a WaveSpeedAI API key, then call POST https://api.wavespeed.ai/api/v3/elevenlabs/voice-changer with your input as JSON. The endpoint returns a prediction id; poll the prediction endpoint until status flips to completed, then read the output URL from data.outputs[0]. Examples for Voice Changer below.

HTTP example
# Submit the prediction
curl -X POST "https://api.wavespeed.ai/api/v3/elevenlabs/voice-changer" \
  -H "Content-Type: application/json" \
  -H "Authorization: Bearer $WAVESPEED_API_KEY" \
  -d '{
    "audio": "https://example.com/your-audio.mp3",
    "voice_id": "Alice",
    "remove_background_noise": false
}'

# Response includes a prediction id. Poll for the result:
curl -X GET "https://api.wavespeed.ai/api/v3/predictions/{request_id}/result" \
  -H "Authorization: Bearer $WAVESPEED_API_KEY"

# When status is "completed", read the output from data.outputs[0].
Node.js example
// npm install wavespeed
const WaveSpeed = require('wavespeed');

const client = new WaveSpeed(); // reads WAVESPEED_API_KEY from env

const result = await client.run("elevenlabs/voice-changer", {
        "audio": "https://example.com/your-audio.mp3",
        "voice_id": "Alice",
        "remove_background_noise": false
});

console.log(result.outputs[0]); // → URL of the generated output
Python example
# pip install wavespeed
import wavespeed

output = wavespeed.run(
    "elevenlabs/voice-changer",
    {
    "audio": "https://example.com/your-audio.mp3",
    "voice_id": "Alice",
    "remove_background_noise": false
}
)

print(output["outputs"][0])  # → URL of the generated output

Voice Changer API — Frequently asked questions

What is the Voice Changer API?

Voice Changer is a ElevenLabs model for AI inference, exposed as a REST API on WaveSpeedAI. ElevenLabs Voice Changer transforms any audio into speech with a different voice while preserving the original speech patterns and timing. Ready-to-use REST inference API, best performance, no cold starts, affordable pricing. You can call it programmatically or try it from the playground above.

How do I call the Voice Changer API?

POST your input parameters to the model's REST endpoint (shown in the API tab of this playground) with your WaveSpeedAI API key in the Authorization header. Submission returns a prediction ID; poll the prediction endpoint until status flips to "completed", then read the output URL from the result. The playground generates a ready-to-paste code sample in Python, JavaScript, or cURL for whatever inputs you've set. Full request/response shape is documented at https://wavespeed.ai/docs/docs-api/elevenlabs/elevenlabs-voice-changer.

How much does Voice Changer cost per run?

Voice Changer starts at $0.025 per run. That figure is the base price — the final charge scales with the parameters you set in the form (output size, length, count, references, or whatever knobs this model exposes), so a higher-quality or larger output costs more than a minimal one. The exact cost for your current input is shown live next to the Generate button before you submit, and the actual per-call charge is recorded on the prediction afterwards.

What inputs does Voice Changer accept?

Key inputs: `audio`, `remove_background_noise`, `voice_id`. The full JSON schema (types, defaults, allowed values) is rendered above the Generate button and mirrored in the API reference at https://wavespeed.ai/docs/docs-api/elevenlabs/elevenlabs-voice-changer.

How long does Voice Changer take to generate?

Average end-to-end generation time on WaveSpeedAI is around 14 seconds per request — measured across recent runs. Queue time scales with global demand; live status is visible in the prediction record.

Can I use Voice Changer outputs commercially?

Commercial usage rights depend on the model's license, set by its provider (ElevenLabs). The license summary appears on the model card above; see WaveSpeedAI's Terms of Service for platform-level conditions.