一文字も書く前に聴いてみる
このページのモデルによる実際の生成結果。各カードにはそれを生み出した台本やプロンプトが表示されます。
Hey, take a deep breath. You’ve been doing great — even when it doesn’t feel like it. Sometimes, the smallest step forward is still progress.
Hello,welcome to WaveSpeedAI!
High-energy cyberpunk synthwave, driving analog bassline, retro futuristic synthesizers, punchy electronic drums, neon noir atmosphere, 120 BPM, mechanical textures.
I don’t need luck. I need focus. Every move, every second — it all leads to one goal.
Hello, little friend! I am your learning partner, Dr. Bot. Today, we will explore the amazing world of numbers together. Are you ready for a fun adventure? Let's go!
Modern deep house fashion runway music, stylish and elegant, groovy bass, rhythmic hi-hats, vocal chops, luxury brand advertisement vibe, confident and cool.
Meet the future of creativity — powered by AI, designed for humans. This is innovation that speaks in your voice.
Some things define eternity without a single word. They are not a product of time, but a witness to it. This is a classic, born for you—a perfect fusion of craftsmanship and art.
Steampunk ambient atmosphere, ticking clock sounds, mechanical gears clicking, steam hissing, soft acoustic guitar in the background, mysterious and studious.
In a world built on shadows and steel, only one voice can bring the story to life — yours.
Look at this! A golden, crispy crust wrapped around a tender, juicy center—the aroma is simply irresistible. Just one bite, and that rich flavor will melt on your tongue. This is absolutely the best dish I have ever tasted!
Once upon a time, in a forest painted with golden leaves, a little fox discovered that courage often comes wrapped in kindness.
最初の一行から完成したサウンドトラックまで、1つのスタジオで
Welcome to our advanced text-to-speech system! Experience high-quality voice synthesis with natural pronunciation and clear articulation.
2. Multilingual v2でより多くの言語を話す
多言語モデルは同じ音声とコントロールのまま、英語以外の台本も読み上げます。
生成Oh, brilliant. Another 'world-changing' idea, I'm sure. I can barely contain my excitement, really. So, I guess we're all just supposed to start working on this right now, right?
3. Eleven Musicでその瞬間を彩る
ジャンル、雰囲気、楽器、テンポを説明。長さを最大5分まで設定してインストゥルメンタルのままにするか、プロンプトに歌詞を書けばボーカル入りになります。
生成Hard Trap, Hip Hop, 808 bass, energetic flow, confident male rap vocals, brass stabs, hype music, motivational sports anthem. [Verse] Laced up tight, ready to go Putting on a major show Sweat and tears on the floor Coming back to get some more [Chorus] Eyes on the prize, reach for the net The greatest game you ever met I take the shot, I make the play Doing this every single day [Outro] Nothing but net. Yeah, we winning.
4. 読み上げ方を調整する
StabilityとSimilarityのスライダー、Speaker boostのスイッチで、台本を書き換えずに各行の読み上げ方を変えられます。
5. または自社製品に組み込む
各モデルにはここで見るのと同じ入力を受け付ける専用のAPIエンドポイントがあり、アプリからの台本が1回のリクエストで音声になります。
APIを見るPOST /api/v3/elevenlabs/eleven-v3 { "text": "Every setback is fuel.", "voice_id": "Talia", "stability": 0.5 }
モデルと料金
料金は生成1回あたりで、リアルタイムに更新されます。解像度・長さ・品質などの設定で最終料金が変わる場合があり、実行前にジェネレーターで確認できます。
| モデル | 入力 | 料金 | |
|---|---|---|---|
| Eleven v3 Text to Speech | テキスト | $0.20 | 生成 |
| Eleven Multilingual v2 Text to Speech | テキスト | $0.20 | 生成 |
| Eleven Music | プロンプト | $0.10 | 生成 |
よくある質問
どのElevenLabs音声モデルから始めるべきですか?
ここで最も新しい音声モデルであるEleven v3から始めましょう。別の言語の台本で音声を比較したい場合はMultilingual v2に切り替えてください。どちらも同じテキスト、音声、設定を使用します。
音声はどうやって選びますか?
音声生成ツールのVoiceフィールドに、Talia、Darian、Elowenなどの音声名を入力します。APIを呼び出す際は同じ名前をvoice_idフィールドに指定します。
Stability、Similarity、Speaker boostは何をするものですか?
いずれもElevenLabsの音声設定です。StabilityとSimilarityはそれぞれ0から1の範囲で、読み上げの一貫性と選んだ音声への忠実度を変化させます。Speaker boostはオン/オフのスイッチです。1行で複数の値を試して違いを聴いてみてください。
Eleven Musicの楽曲はどのくらいの長さにできますか?
長さは生成ツールで5秒から5分まで設定できます。楽曲はデフォルトでインストゥルメンタルです。ボーカルが欲しい場合はそれをオフにし、プロンプトに歌詞を入れてください。
これらのモデルを自分のコードから利用できますか?
はい。表の各モデルにはリクエスト例付きのAPIページがあり、ドキュメントでは認証と結果の取得方法も説明しています。