WAN 3.0 is LIVE — 30s in one shot | Try in Video Generator →
ElevenLabs AI Models

ElevenLabs AI Models

ElevenLabs delivers lifelike AI voice, music, sound effects, dubbing, transcription, and audio generation workflows

ElevenLabs delivers lifelike AI voice, music, sound effects, dubbing, transcription, and audio generation workflows

All models

11 models
elevenlabs/turbo-v2.5
text-to-audio

elevenlabs/turbo-v2.5

ElevenLabs Turbo V2.5 is a text-to-speech model available via WaveSpeedAI, billed at $0.05 per 1000 characters for TTS requests. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

elevenlabs/multilingual-v2
text-to-audio

elevenlabs/multilingual-v2

ElevenLabs Multilingual V2 is a multilingual text-to-speech model; cost $0.1 per 1000 characters. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

elevenlabs/turbo-v2
text-to-audio

elevenlabs/turbo-v2

ElevenLabs Turbo V2 is a Text-To-Speech model available via WaveSpeedAI, billed at $0.05 per 1000 characters for API requests. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

elevenlabs/flash-v2
text-to-audio

elevenlabs/flash-v2

ElevenLabs Flash V2 is a Text-to-Speech model that converts text into spoken audio using the ElevenLabs Flash V2 engine. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

elevenlabs/flash-v2.5
text-to-audio

elevenlabs/flash-v2.5

ElevenLabs Flash v2.5 is a text-to-speech model on WaveSpeedAI, billed at $0.05 per 1000 characters for generated speech. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

elevenlabs/multilingual-v1
text-to-audio

elevenlabs/multilingual-v1

ElevenLabs Multilingual V1 provides natural-sounding multilingual text-to-speech across many languages. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

elevenlabs/eleven-v3
text-to-audio

elevenlabs/eleven-v3

ElevenLabs eleven-v3 is a text-to-speech model available as a hosted endpoint; requests cost $0.1 per 1000 characters. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

elevenlabs/eleven-v3/timing
text-to-audio

elevenlabs/eleven-v3/timing

ElevenLabs Eleven-V3 Timing converts text to natural speech and returns alignment metadata—character/word timestamps in JSON—for precise subtitles, karaoke effects, and lip-sync. Supports voice_id, similarity/stability, and optional Speaker Boost. Priced at $0.10 per 1,000 characters. Ready-to-use REST inference API, best performance, no cold starts, affordable pricing.

elevenlabs/voice-changer
audio-to-audio

elevenlabs/voice-changer

ElevenLabs Voice Changer transforms any audio into speech with a different voice while preserving the original speech patterns and timing. Ready-to-use REST inference API, best performance, no cold starts, affordable pricing.

elevenlabs/dubbing
video-dubbing

elevenlabs/dubbing

ElevenLabs Dubbing automatically translates and dubs video/audio content into different languages while preserving the original speakers' voices. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

elevenlabs/music
text-to-audio

elevenlabs/music

ElevenLabs Music generates original songs from text descriptions. Create instrumentals or full compositions with customizable duration. Ready-to-use REST inference API, best performance, no cold starts, affordable pricing.

ElevenLabs AI Models

ElevenLabs provides a professional AI audio generation and voice intelligence model suite for text-to-speech, speech-to-text, voice changing, dubbing, music generation, sound effects generation, text-to-dialogue, and voice isolation workflows. The collection is designed for creators, developers, media teams, educators, game studios, and AI applications that need natural voice synthesis, multilingual audio, cinematic sound design, and scalable audio production.

Built for high-quality speech and audio creation, ElevenLabs helps users generate lifelike voiceovers, transcribe spoken audio, create sound effects from text prompts, produce music, transform voices, dub content across languages, and clean voice audio from noisy recordings. It is suitable for videos, podcasts, audiobooks, games, ads, product explainers, education, localization, digital humans, and real-time voice applications.

Core Model Capabilities

Text-to-Speech Generation:

Create natural, expressive voiceovers from text with realistic pacing, intonation, emotion, and multilingual support for narration, dialogue, explainers, ads, and creator content.

Speech-to-Text Transcription:

Convert spoken audio into accurate text for subtitles, transcripts, meeting notes, podcasts, interviews, media indexing, and automated content workflows.

Voice Changer:

Transform uploaded or recorded speech into a different voice while preserving performance, timing, emotion, and delivery style.

AI Dubbing:

Translate and dub audio or video content into other languages while maintaining natural speech flow, speaker style, and localization quality.

Text-to-Music Generation:

Generate music from text prompts for background scoring, social media content, games, ads, podcasts, trailers, and creative audio production.

Production-Ready Audio API:

Access ElevenLabs models through scalable APIs for automated voice generation, transcription, localization, sound design, music creation, and AI-powered audio applications.

ElevenLabs AI Models on WaveSpeedAI give creators and developers fast access to professional audio, voice, music, dubbing, and transcription tools with flexible pricing, scalable API access, and production-ready audio quality.

ElevenLabs AI Models API — pricing & performance

Run any model in the ElevenLabs AI Models collection through a single REST API. Pay per generation — no subscriptions, no minimums — with industry-leading latency on a 99.9% uptime infrastructure.

Why run ElevenLabs AI Models on WaveSpeedAI

Transparent pricing

Per-call pricing for every ElevenLabs AI Models model. The price is listed on each model page — no platform fees on top.

Optimized for low latency

Most ElevenLabs AI Models image models complete in under 2 seconds. Video and 3D models run several times faster than self-hosted alternatives.

99.9% uptime

Multi-region failover and automatic retries keep your production traffic online — even during provider outages.

Frequently asked questions

How much does the ElevenLabs AI Models API cost?+

Each model has its own per-call price listed on the model page. We bill per successful generation, with no subscription fees or minimums.

How fast are ElevenLabs AI Models models on WaveSpeedAI?+

Image models in this collection typically complete in under 2 seconds. Video and 3D models depend on duration and resolution but are usually several times faster than self-hosted runs.

Can I try the API without a credit card?+

Yes — every account gets $1 in free credits on signup, enough to try most ElevenLabs AI Models models without a credit card.

Are there rate limits?+

Standard accounts have generous concurrent-job limits. Enterprise plans offer custom RPM, higher concurrency, and dedicated capacity — contact sales for details.