Nano Banana 2.1 公開中 — Google 最新モデル | 今すぐ試す →
ElevenLabs AI Models

ElevenLabs AI Models

ElevenLabs delivers lifelike AI voice, music, sound effects, dubbing, transcription, and audio generation workflows

ElevenLabs delivers lifelike AI voice, music, sound effects, dubbing, transcription, and audio generation workflows

すべてのモデル

18 モデル
elevenlabs/music-v2
ai-music$0.7500

elevenlabs/music-v2

ElevenLabs Music v2 Text-to-Music generates songs with vocals or instrumental music from text prompts, with configurable duration and MP3 output for songwriting demos, background music, social content, ads, and creative audio production workflows. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

elevenlabs/music-v2.5
ai-music$0.7500

elevenlabs/music-v2.5

ElevenLabs Music v2.5 Text-to-Music generates songs with vocals or instrumental music from text prompts, with configurable duration and MP3 output for songwriting demos, background music, social content, ads, and creative audio production workflows. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

elevenlabs/flash-v2.5
text-to-speech$0.1000

elevenlabs/flash-v2.5

ElevenLabs Flash v2.5 is a text-to-speech model on WaveSpeedAI, billed at $0.05 per 1000 characters for generated speech. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

elevenlabs/music
ai-music$0.1000

elevenlabs/music

ElevenLabs Music Text-to-Music generates original songs from text descriptions, supporting instrumental tracks, full compositions, customizable duration, background music, social content, ads, and creative audio production workflows. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

elevenlabs/turbo-v2.5
text-to-speech$0.1000

elevenlabs/turbo-v2.5

ElevenLabs Turbo V2.5 is a text-to-speech model available via WaveSpeedAI, billed at $0.05 per 1000 characters for TTS requests. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

elevenlabs/eleven-v3
text-to-speech$0.2000

elevenlabs/eleven-v3

ElevenLabs eleven-v3 is a text-to-speech model available as a hosted endpoint. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

elevenlabs/eleven-v3/timing
text-to-speech$0.2000

elevenlabs/eleven-v3/timing

ElevenLabs Eleven V3 Timing converts text to expressive speech and returns character-level alignment timestamps for subtitles, highlighting, and audio editing. Supports preset or custom voice IDs and adjustable stability.

elevenlabs/voice-changer
speech-to-speech$0.0040

elevenlabs/voice-changer

ElevenLabs Voice Changer transforms any audio into speech with a different voice while preserving the original speech patterns and timing. Ready-to-use REST inference API, best performance, no cold starts, affordable pricing.

elevenlabs/dubbing
video-dubbing$0.0100

elevenlabs/dubbing

ElevenLabs Dubbing automatically translates and dubs video/audio content into different languages while preserving the original speakers' voices. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

elevenlabs/flash-v2
text-to-speech$0.1000

elevenlabs/flash-v2

ElevenLabs Flash V2 is a Text-to-Speech model that converts text into spoken audio using the ElevenLabs Flash V2 engine. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

elevenlabs/forced-alignment
audio-tools$0.3000

elevenlabs/forced-alignment

ElevenLabs Forced Alignment aligns an existing transcript with an audio recording and returns character-level and word-level timestamps as structured JSON for subtitles, captions, dubbing, localization, and transcript synchronization workflows. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

elevenlabs/turbo-v2
text-to-speech$0.1000

elevenlabs/turbo-v2

ElevenLabs Turbo V2 is a Text-To-Speech model available via WaveSpeedAI, billed at $0.05 per 1000 characters for API requests. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

elevenlabs/multilingual-v2
text-to-speech$0.2000

elevenlabs/multilingual-v2

ElevenLabs Multilingual V2 is a multilingual text-to-speech model; cost $0.1 per 1000 characters. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

elevenlabs/eleven-v4
text-to-speech$0.0800

elevenlabs/eleven-v4

Eleven V4 Text-to-Speech converts text into expressive multilingual speech with configurable voice selection, stability, and similarity controls for narration, dialogue, localization, virtual assistants, and voice production workflows. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

elevenlabs/eleven-v4-turbo
text-to-speech$0.0400

elevenlabs/eleven-v4-turbo

Eleven V4 Turbo Text-to-Speech converts text into expressive multilingual speech with configurable voice selection, stability, and similarity controls for narration, dialogue, localization, virtual assistants, and scalable voice production workflows. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

elevenlabs/audio-isolation
audio-to-audio$0.1100

elevenlabs/audio-isolation

ElevenLabs Audio Isolation isolates speech from background noise to produce clearer dialogue, interviews, voice recordings, podcasts, and other speech-focused audio workflows. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

elevenlabs/scribe-v2
speech-to-text$0.0100

elevenlabs/scribe-v2

ElevenLabs Scribe V2 Speech-to-Text transcribes audio with automatic language detection, speaker labels, and word-level timestamps for transcription, subtitles, captions, meetings, and audio processing workflows. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

elevenlabs/sound-effects-v2
sound-effects$0.0022

elevenlabs/sound-effects-v2

ElevenLabs Sound Effects V2 Text-to-SFX generates high-quality sound effects and seamlessly looping ambience from text descriptions for video, games, ads, social content, and sound design workflows. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

ElevenLabs AI Models

ElevenLabs provides a professional AI audio generation and voice intelligence model suite for text-to-speech, speech-to-text, voice changing, dubbing, music generation, sound effects generation, text-to-dialogue, audio isolation, and speech cleanup workflows. The collection is designed for creators, developers, media teams, educators, game studios, and AI applications that need natural voice synthesis, multilingual transcription, cinematic sound design, localization, and scalable audio production.

Built for high-quality speech and audio creation, ElevenLabs helps users generate lifelike voiceovers, transcribe spoken audio, create sound effects from text prompts, produce music, transform voices, dub content across languages, generate multi-speaker dialogue, and isolate clean speech from noisy recordings. It is suitable for videos, podcasts, audiobooks, games, ads, product explainers, education, localization, digital humans, and real-time voice applications.

Core Model Capabilities

Text-to-Speech Generation:

Create natural and expressive voiceovers from text with realistic pacing, intonation, emotion, and multilingual support for narration, dialogue, explainers, ads, audiobooks, and creator content.

Speech-to-Text Transcription:

Convert spoken audio into accurate text for subtitles, transcripts, meeting notes, podcasts, interviews, localization, media indexing, and automated content workflows.

Voice Changer:

Transform uploaded or recorded speech into a different voice while preserving timing, performance, emotion, and delivery style.

AI Dubbing:

Translate and dub audio or video content into other languages while maintaining natural speech flow, speaker characteristics, and localization quality.

AI Music Generation:

Generate music from text prompts for background scoring, social media content, games, ads, podcasts, trailers, and creative audio production.

Sound Effects Generation:

Create sound effects from natural-language prompts for cinematic impacts, ambience, Foley, environmental audio, transitions, UI sounds, action effects, and other custom sound design workflows.

Text-to-Dialogue:

Generate natural multi-speaker dialogue from text for character scenes, audio dramas, training content, game dialogue, conversational media, and narrative audio production.

Audio Isolation:

Separate speech from background noise or mixed audio while preserving clear voice content. This is useful for podcasts, interviews, dubbing, transcription, video production, localization, and post-production cleanup.

Speech and Audio Enhancement:

Improve voice clarity and overall audio usability for production workflows that need cleaner speech, stronger intelligibility, and more polished audio output.

Production-Ready Audio API:

Access ElevenLabs models through scalable APIs for automated voice generation, transcription, localization, dubbing, sound design, music creation, dialogue generation, speech cleanup, and AI-powered audio applications.

ElevenLabs AI Models on WaveSpeedAI give creators and developers fast access to professional voice, transcription, music, sound effects, dubbing, dialogue, and audio isolation tools with flexible pricing, scalable API access, and production-ready audio quality.

ElevenLabs AI Models API — 料金とパフォーマンス

ElevenLabs AI Models コレクションのすべてのモデルを単一の REST API で実行できます。生成ごとに課金 — サブスクなし、最低料金なし — で、稼働率 99.9% のインフラ上の業界トップクラスのレイテンシを提供します。

WaveSpeedAI で ElevenLabs AI Models を使う理由

透明な料金体系

各 ElevenLabs AI Models モデルにコールごとの料金が設定されています。料金は各モデルのページに表示され、プラットフォーム手数料はかかりません。

低レイテンシに最適化

ほとんどの ElevenLabs AI Models 画像モデルは 2 秒以内に完了します。動画や 3D モデルはセルフホスト構成より数倍高速です。

稼働率 99.9%

マルチリージョンのフェイルオーバーと自動リトライで、プロバイダー障害時にも本番トラフィックを維持します。

よくある質問

ElevenLabs AI Models API の料金はいくらですか?+

各モデルにはモデルページ上にコール単価が記載されています。成功した生成ごとに課金され、サブスクリプション料金や最低料金はありません。

WaveSpeedAI 上の ElevenLabs AI Models モデルはどのくらい高速ですか?+

このコレクションの画像モデルは通常 2 秒以内に完了します。動画や 3D モデルは長さや解像度に依存しますが、セルフホスト実行より数倍高速なことが多いです。

クレジットカードなしで API を試せますか?+

条件を満たす新規アカウントは、クレジットカードなしで ElevenLabs AI Models モデルを試すためのプロモーションクレジット 1 ドル分を受け取れる場合があります。すべての登録で付与されるわけではありません。生成前にアカウント残高をご確認ください。

レート制限はありますか?+

標準アカウントには十分な同時実行ジョブ枠があります。Enterprise プランではカスタム RPM、より高い同時実行性、専用キャパシティを提供します — 詳細は営業へお問い合わせください。