MiniMax H3 API
MiniMax H3 API は、テキストから動画、画像から動画、リファレンスから動画のシネマティックな生成に対応し、ネイティブステレオオーディオ、優れたモーション品質、被写体の一貫性、シーンのまとまりを備えています。オープンウェイト版を WaveSpeed のインフラで、または公式の MiniMax エンドポイントを、1つの WaveSpeedAI API から実行できます。
テキストプロンプト、静止画、ビジュアルリファレンスからシネマティックなAI動画を生成します。WaveSpeed のインフラ上で推奨されるオープンウェイト版(wavespeed-ai/minimax-h3)は、480p/540p/768p/1080pでまとまりのある動画を、ネイティブステレオオーディオ付き・3〜15秒・リーズナブルな秒単位料金で生成します。公式の MiniMax エンドポイントも同じファミリーで利用できます。
概要
MiniMax H3 APIについて
MiniMax H3でできること、MiniMaxのモデルラインナップにおける位置づけ、そして多くのチームに選ばれる理由。
MiniMax H3はMiniMaxの動画生成モデルで、WaveSpeedAIのREST APIから利用できます。MiniMax H3 API は、テキストから動画、画像から動画、リファレンスから動画のシネマティックな生成に対応し、ネイティブステレオオーディオ、優れたモーション品質、被写体の一貫性、シーンのまとまりを備えています。オープンウェイト版を WaveSpeed のインフラで、または公式の MiniMax エンドポイントを、1つの WaveSpeedAI API から実行できます。
テキストプロンプト、静止画、ビジュアルリファレンスからシネマティックなAI動画を生成します。WaveSpeed のインフラ上で推奨されるオープンウェイト版(wavespeed-ai/minimax-h3)は、480p/540p/768p/1080pでまとまりのある動画を、ネイティブステレオオーディオ付き・3〜15秒・リーズナブルな秒単位料金で生成します。公式の MiniMax エンドポイントも同じファミリーで利用できます。
WaveSpeedAIのMiniMax H3ファミリーには、Motion-Control, Image-To-Video, Reference-To-Video, Image-To-Image, Text-To-Image, Video-To-Video, Video-Extend, Text-To-Videoのワークフローをカバーする20個のRESTエンドポイントがあります。各バリアントには、それぞれ独自の料金、パラメーター、作例があります。入力の種類と本番環境の制約に合うものを選ぶか、同じAPIキーで複数を呼び出して、多段階のパイプラインを組み立ててください。
WaveSpeedAIの他の1,000以上のAIモデルと同じAPIキー、請求アカウント、レート制限の枠組みでMiniMax H3を実行できます。ベンダーごとの設定も、プロバイダーごとのSDKも、ベンダーごとのレート制限も不要です。1つの連携で、テキストから画像、テキストから動画から、音声合成、3D生成、高画質化、編集まで、すべてカバーできます。
仕様
MiniMax H3 APIの機能とリリース状況
APIを選ぶ前に開発者が調べる、モデル固有の詳細情報:提供状況、想定される出力の長さ、参照入力への対応、そして現在利用できる代替エンドポイント。
テキストから動画
プロンプト駆動の生成
wavespeed-ai/minimax-h3/text-to-video で、文章のクリエイティブブリーフから、480p/540p/768p/1080pのまとまりある動画を生成します。ネイティブステレオオーディオ、3〜15秒の長さ、柔軟なアスペクト比に対応しています。
画像から動画
静止画をアニメーション化
wavespeed-ai/minimax-h3/image-to-video で、最初のフレームの画像(必要に応じて最後のフレームも)をアニメーション化し、被写体、フレーミング、アートディレクションをモーションに引き継いだ、ネイティブステレオオーディオ付きの動画を生成します。
リファレンスから動画
アイデンティティとスタイルをガイド
wavespeed-ai/minimax-h3/reference-to-video で、最大9枚の参照画像、3本の参照動画、3本の参照オーディオを使って生成をガイドします。被写体、キャラクターのアイデンティティ、スタイルを一貫させたいときに使います。
API連携
1つの WaveSpeedAI ワークフロー
WaveSpeedAI のAPIキーで H3 の予測を送信し、標準の予測ライフサイクルで各リクエストを追跡して、結果レスポンスから生成された動画を取得します。WaveSpeed ホストの版も公式の MiniMax エンドポイントも、同じワークフローで扱えます。
エンドポイント
MiniMax H3のAPIエンドポイント一覧
WaveSpeedAIで現在利用可能なMiniMax H3のエンドポイントは20個です。ワークフローに合ったバリアントを選んでください。
/filters:quality(82)/media/images/1790573118566544886_jzIR0oxG.webp)
Minimax H3 Controlnet Union
MiniMax H3 Open Weights ControlNet Union generates a new video that follows the motion and composition of a source video. Pose, depth, edges, lines, scribble or grayscale structure is extracted from the source automatically and guides the output, optionally with reference images for the subject or style, with native stereo audio generated in the same pass. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/1789888783447157992_hG8iqAJS.webp)
Minimax H3 Singularity Image To Video Lora
MiniMax H3 Singularity Open Weights Image to Video animates a first-frame image, optionally with last-frame guidance, into coherent 480P / 540P / 768P videos with native stereo audio, 5-15 second duration, flexible aspect ratios, and per-second billing. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/1789888766253281059_pb1bluDL.webp)
Minimax H3 Singularity Image To Video
MiniMax H3 Singularity Open Weights Image-to-Video animates a first-frame image, with optional last-frame guidance, into coherent 480P / 540P / 768P videos with native stereo audio, 5-15 second duration, flexible aspect ratios, and per-second billing. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/1789742735070040548_71aJS2bl.webp)
Minimax H3 Singularity Reference To Video Lora
MiniMax H3 Singularity Open Weights Reference to Video with custom LoRA support generates coherent 480P / 540P / 768P videos from prompts and multimodal references, guided by up to 9 reference images, 3 reference videos, and 3 reference audios, with native stereo audio and flexible reference-based video generation. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/1789742804979721138_b4T1aktE.webp)
Minimax H3 Singularity Reference To Video
MiniMax H3 Singularity Open Weights Reference to Video generates coherent 480P / 540P / 768P videos from prompts and multimodal references, guided by up to 9 reference images, 3 reference videos, and 3 reference audios, with native stereo audio and flexible reference-based video generation on WaveSpeedAI infrastructure. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/1788684887040987486_ILX6fU3e.webp)
Minimax H3 Image Edit Lora
MiniMax H3 Open Weights Image Edit re-renders a subject from up to 9 reference images into a new scene, outfit or style from a text instruction, preserving identity at 1K or 2K resolution. Loads up to 3 custom LoRA weights per request. Ready-to-use REST inference API, best performance, no cold starts, affordable pricing.
/filters:quality(82)/media/images/1788684121915845410_PlPY7gqA.webp)
Minimax H3 Text To Image Lora
MiniMax H3 Open Weights Text-to-Image generates photorealistic, cinema-grade stills from a text prompt at 1K or 2K resolution with fifteen aspect ratios. Loads up to 3 custom LoRA weights per request. Ready-to-use REST inference API, best performance, no cold starts, affordable pricing.
/filters:quality(82)/media/images/1788684872529921426_LGPY7gpy.webp)
Minimax H3 Image Edit
MiniMax H3 Open Weights Image Edit re-renders a subject from up to 9 reference images into a new scene, outfit or style from a text instruction, preserving identity at 1K or 2K resolution. Ready-to-use REST inference API, best performance, no cold starts, affordable pricing.
/filters:quality(82)/media/images/1788684072554630514_Lg4clvEO.webp)
Minimax H3 Text To Image
MiniMax H3 Open Weights Text-to-Image generates photorealistic, cinema-grade stills from a text prompt at 1K or 2K resolution with fifteen aspect ratios. Ready-to-use REST inference API, best performance, no cold starts, affordable pricing.
/filters:quality(82)/media/images/1788329903545692055_lEMW5fpy.webp)
Minimax H3 Video Edit
MiniMax H3 Open Weights Video-Edit edits an input video from a natural-language prompt. The input video drives subject identity, composition, and motion while the model rewrites lighting, style, weather, environment, or specific elements as instructed, with native stereo audio generated in the same pass. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/1788333421627652695_cNUHuh5T.webp)
Minimax H3 Video Extend
MiniMax H3 Open Weights Video-Extend appends a new cinematic continuation to an existing video. A fresh segment with native stereo audio is generated from the input video's last frame and a natural-language prompt, then the original and new segment are concatenated into a single output. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/1786952873106530692_cwjsCLT3.webp)
Minimax H3 Reference To Video Lora
MiniMax H3 Open Weights Reference to Video with custom LoRA support generates coherent 480P / 540P / 768P videos from prompts and multimodal references, guided by up to 9 reference images, 3 reference videos, and 3 reference audios, with native stereo audio and flexible reference-based video generation. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/1786952854061501695_t8heoxGQ.webp)
Minimax H3 Image To Video Lora
MiniMax H3 Open Weights Image to Video animates a first-frame image, optionally with last-frame guidance, into coherent 480P / 540P / 768P videos with native stereo audio, 5-15 second duration, flexible aspect ratios, and per-second billing. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/1788760867264586097_7BZ9isBK.webp)
Minimax H3 Text To Video Lora
MiniMax H3 Open Weights Text to Video with custom LoRA support generates coherent videos from text prompts, with 480P / 540P / 768P output, native stereo audio, 5-15 second duration, flexible aspect ratios, and per-second billing. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/1785814033019871155_vhrAKT2c.webp)
Minimax H3 Reference To Video
MiniMax H3 Open Weights Reference to Video generates coherent 480P / 540P / 768P videos from prompts and multimodal references, guided by up to 9 reference images, 3 reference videos, and 3 reference audios, with native stereo audio and flexible reference-based video generation on WaveSpeedAI infrastructure. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/1785815150161566057_2VF6wZoP.webp)
Minimax H3 Image To Video
MiniMax H3 Open Weights Image to Video animates a first-frame image, optionally with last-frame guidance, into coherent 480P / 540P / 768P videos with native stereo audio, 5-15 second duration, flexible aspect ratios, and per-second billing. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/1785814561183432100_tW5eoxGP.webp)
Minimax H3 Text To Video
MiniMax H3 Open Weights Text to Video generates coherent videos from text prompts, with 480P / 540P / 768P output, native stereo audio, 5-15 second duration, flexible aspect ratios, and per-second billing. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/1785415007635242852_UoxHQZ9H.webp)
H3 Reference To Video
MiniMax H3 Reference to Video generates coherent 2K videos from natural-language prompts and multimodal references, including images, videos, and audio, guiding subject consistency, motion, timing, visual style, and scene continuity. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/1785414895750297795_1JajtDMV.webp)
H3 Image To Video
MiniMax H3 Image to Video animates a first-frame image into a coherent 2K video, with natural-language motion instructions and optional last-frame control for consistent motion, scene continuity, and cinematic video generation. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/1785412796956877340_AktCLU4J.webp)
H3 Text To Video
MiniMax H3 Text to Video generates coherent 2K videos from text prompts, with flexible 5-15 second duration and adaptive or custom aspect ratios for cinematic scenes, creative videos, and production workflows. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
作例
MiniMax H3の実力を見る
MiniMax H3 APIで実際に生成された出力です。動画にカーソルを合わせるとプレビュー、クリックすると原寸のビューアで開きます。
使い方
MiniMax H3 APIの使い方
登録から生成完了まで4ステップ。Python、Node.js、cURLの完全なサンプルは、下のAPIセクションにあります。
- 01
APIキーを取得
WaveSpeedAIのアカウントに登録し、ダッシュボードからAPIキーをコピーします。新規アカウントには無料のスタータークレジットが付くので、課金が始まる前にプレイグラウンドを数十回実行できます。
- 02
予測を送信
入力をJSONとしてhttps://api.wavespeed.ai/api/v3/wavespeed-ai/minimax-h3/text-to-videoにPOSTします。エンドポイントはすぐに予測IDを返します。生成は非同期なので、推論中に接続を開いたままにする必要はありません。
- 03
完了までポーリング
https://api.wavespeed.ai/api/v3/predictions/{request_id}/resultにGETします。completedなら出力を返し、failed、cancelled、timeout、deletedならエラーで停止し、それ以外のステータスの間はポーリングを続けます。
- 04
出力URLを読み取る
ステータスが"completed"になったら、data.outputs[0]からURLを読み取ります。URLは、WaveSpeedAIのCDN上にある生成されたメディアを指します。呼び出したMiniMax H3のバリアントに応じて、画像、動画、音声、3Dファイルのいずれかです。
活用例
MiniMax H3で作れるもの
開発者やクリエイターがMiniMax H3 APIでよく使うワークフロー。
オリジナルのコンセプト向けテキストから動画
wavespeed-ai/minimax-h3/text-to-video で、プロンプトからオリジナルのシネマティックなシーンを生成。480p/540p/768p/1080pのまとまりある出力、ネイティブステレオオーディオ、柔軟なアスペクト比に対応します。コンセプトの可視化、ストーリーテリング、キャンペーンのアイデア、既存素材から始める必要のないショットに最適です。
商品紹介向けの画像から動画
商品写真、キャンペーンの静止画、キービジュアル、キャラクター画像を、元のビジュアルの方向性を保ったままアニメーション化します。商品のお披露目、SNS広告、モーション中心のランディングコンテンツに便利です。
一貫した被写体のためのリファレンスから動画
最大9枚の参照画像、3本の参照動画、3本の参照オーディオで、キャラクターのアイデンティティ、被写体の見た目、シーンデザイン、スタイルをガイドできます。リファレンスから動画は、繰り返し登場するキャラクター、ブランドビジュアル、関連するクリエイティブセットのための H3 ワークフローです。
SNS・広告クリエイティブ
キャンペーン、ローンチ、SNSフィード、パフォーマンスマーケティング向けに、ショートフォームのクリエイティブを生成します。テキスト、完成した静止画、承認済みのビジュアルリファレンスのどこから始めても、APIプラットフォームを変える必要はありません。
キャラクターシーンとビジュアルストーリーテリング
キャラクター主体のシーン、ストーリーのビート、ムード映像、プリビズを制作できます。素材に合わせてエンドポイントを選びます。新しいシーンにはテキスト、アニメーション化には画像、より強い連続性にはリファレンスです。
スケーラブルなAI動画生成パイプライン
WaveSpeedAI を通じて MiniMax H3 をアプリや自動化されたクリエイティブワークフローに組み込めます。H3 の3つの生成モードすべてで、1つのAPIキーと共通の予測ライフサイクルを使えます。
コツ
MiniMax H3のプロンプトのコツ
MiniMax H3からより良い出力を得るための実践的なアドバイス。本番のパイプラインで動画モデル全般に通用するパターンに基づいています。
- 01
素材に合わせてエンドポイントを選ぶ
オリジナルのプロンプトにはテキストから動画、1枚の元画像のアニメーション化には画像から動画、ビジュアルの参照でアイデンティティ、被写体の見た目、シーンデザイン、スタイルをガイドしたい場合にはリファレンスから動画を使います。
- 02
モーションを時系列で記述する
開始時の状態、アクション、環境の反応、カメラの動き、終了時の状態を、時系列で書いてください。つながりのない視覚的な形容詞を並べるより、明確な流れのほうが、H3 により強いモーションの計画を与えます。
- 03
構図が承認済みなら画像から動画を使う
被写体、商品、フレーミング、アートディレクションが静止画で固まっているなら、テキストからシーンを作り直さず、画像から動画から始めてください。プロンプトは動きとカメラの挙動に集中させます。
- 04
参照には、はっきりした役割を1つ与える
リファレンスから動画では、参照が何を制御するのか(キャラクターのアイデンティティ、商品の見た目、ビジュアルスタイル、環境、構図)を明記します。同じ参照に、ショットの複数の相反する要素を定義させないでください。
- 05
短いシーンでは、主なアクションを1つに絞る
関連のない出来事を並べるより、1つのアクションに絞るほうが、まとまりよくレンダリングできます。ストーリーの区切りごとに別々のショットを生成し、コンセプトにより広い物語の範囲が必要なら、編集でつなげてください。
料金
MiniMax H3 APIの料金
料金は出力ごとです。最終的な請求額は、各バリアントのプレイグラウンドで設定するパラメーター(解像度、長さ、出力数、参照入力)に応じて変わります。
API
MiniMax H3 APIを呼び出す
wavespeed.ai/accesskeyでAPIキーに登録し、RESTで予測を送信します。プレイグラウンドでは、入力の組み合わせに応じて、そのまま貼り付けられるサンプルが生成されます。
POSThttps://api.wavespeed.ai/api/v3/wavespeed-ai/minimax-h3/text-to-video
# 1. Submit the prediction.
SUBMIT_RESPONSE=$(curl --silent --show-error --fail-with-body \
-X POST "https://api.wavespeed.ai/api/v3/wavespeed-ai/minimax-h3/text-to-video" \
-H "Content-Type: application/json" \
-H "Authorization: Bearer $WAVESPEED_API_KEY" \
-d '{
"prompt": "A cinematic shot of a city at sunset, soft golden light",
"aspect_ratio": "16:9",
"resolution": "480p",
"duration": 5
}')
TASK=$(printf '%s' "$SUBMIT_RESPONSE" | jq 'if has("data") then .data else . end')
PREDICTION_ID=$(printf '%s' "$TASK" | jq -r '.id')
if [ -z "$PREDICTION_ID" ] || [ "$PREDICTION_ID" = "null" ]; then
printf 'Submission response did not contain a prediction id
' >&2
exit 1
fi
RESULT_URL="https://api.wavespeed.ai/api/v3/predictions/$PREDICTION_ID/result"
# 2. Poll until the prediction finishes.
while true; do
RESPONSE=$(curl --silent --show-error --fail-with-body "$RESULT_URL" \
-H "Authorization: Bearer $WAVESPEED_API_KEY")
RESULT=$(printf '%s' "$RESPONSE" | jq 'if has("data") then .data else . end')
STATUS=$(printf '%s' "$RESULT" | jq -r '.status')
case "$STATUS" in
completed) printf '%s\n' "$RESULT" | jq '.outputs'; break ;;
failed|cancelled|timeout|deleted) printf '%s\n' "$RESULT" | jq . >&2; exit 1 ;;
*) sleep 2 ;;
esac
doneconst submitUrl = "https://api.wavespeed.ai/api/v3/wavespeed-ai/minimax-h3/text-to-video";
const apiKey = process.env.WAVESPEED_API_KEY;
if (!apiKey) throw new Error('Set WAVESPEED_API_KEY');
async function requestJson(url, options = {}) {
const response = await fetch(url, options);
if (!response.ok) throw new Error(await response.text());
return response.json();
}
// 1. Submit the prediction.
const body = await requestJson(submitUrl, {
method: "POST",
headers: {
"Authorization": `Bearer ${apiKey}`,
"Content-Type": "application/json",
},
body: JSON.stringify({
"prompt": "A cinematic shot of a city at sunset, soft golden light",
"aspect_ratio": "16:9",
"resolution": "480p",
"duration": 5
}),
});
const task = body.data ?? body;
const resultUrl = `https://api.wavespeed.ai/api/v3/predictions/${task.id}/result`;
// 2. Poll until the prediction finishes.
while (true) {
const resultBody = await requestJson(resultUrl, {
headers: { "Authorization": `Bearer ${apiKey}` },
});
const result = resultBody.data ?? resultBody;
if (result.status === "completed") {
console.log(result.outputs);
break;
}
if (["failed", "cancelled", "timeout", "deleted"].includes(result.status)) throw new Error(JSON.stringify(result));
await new Promise(resolve => setTimeout(resolve, 2000));
}import json
import os
import time
from urllib.request import Request, urlopen
api_key = os.environ["WAVESPEED_API_KEY"]
headers = {"Authorization": f"Bearer {api_key}", "Content-Type": "application/json"}
payload = {
"prompt": "A cinematic shot of a city at sunset, soft golden light",
"aspect_ratio": "16:9",
"resolution": "480p",
"duration": 5
}
def request_json(url, data=None):
request = Request(url, data=data, headers=headers, method="POST" if data else "GET")
with urlopen(request) as response:
return json.load(response)
# 1. Submit the prediction.
body = request_json("https://api.wavespeed.ai/api/v3/wavespeed-ai/minimax-h3/text-to-video", json.dumps(payload).encode())
task = body.get("data", body)
result_url = f"https://api.wavespeed.ai/api/v3/predictions/{task['id']}/result"
# 2. Poll until the prediction finishes.
while True:
result_body = request_json(result_url)
result = result_body.get("data", result_body)
status = result.get("status")
if status == "completed":
print(result.get("outputs", []))
break
if status in {"failed", "cancelled", "timeout", "deleted"}:
raise RuntimeError(result)
time.sleep(2)比較
MiniMax H3と他モデルの比較
WaveSpeedAI上の類似モデルではなくMiniMax H3を選ぶべきケース。
MiniMax H3 vs Hailuo 2.3
Hailuo 2.3 は、Standard、Pro、Fast、Fast Pro のバリアントにまたがる階層型のテキストから動画・画像から動画エンドポイントを提供します。MiniMax H3 は3つのエンドポイントに絞ったスイートで、ビジュアルガイドと被写体の一貫性のためのリファレンスから動画を、主要なワークフローとして備えています。
MiniMax H3 vs Seedance 2.0
Seedance 2.0 は、テキストから動画、画像から動画、video-edit、video-extend、ネイティブオーディオ、複数のパフォーマンス階層を備えた幅広い制作向けファミリーです。テキスト、元画像、ビジュアルリファレンスからの生成が中心のワークフローなら、MiniMax H3 のほうがシンプルな選択肢です。
MiniMax H3 vs Kling 3.0
Kling 3.0 は、階層化された出力品質と専用のモーションコントロールエンドポイントに力を入れています。MiniMax H3 は素材の種類に沿ってAPIを構成しており、アイデンティティ、被写体、スタイルをガイドする生成のための専用のリファレンスから動画ルートがあります。
FAQ
MiniMax H3 API — よくある質問
料金、ライセンス、連携など、WaveSpeedAIでMiniMax H3を実行する際のよくある質問。
MiniMax H3 API とは何ですか?
MiniMax H3 API は、WaveSpeedAI で利用できる3モデル構成のAI動画生成スイートです。プロンプト駆動のシーン向けのテキストから動画、静止画をアニメーション化する画像から動画、ビジュアルリファレンスでガイドされる生成のためのリファレンスから動画に対応しています。
MiniMax H3 API ではどのエンドポイントを使えばよいですか?
まずは WaveSpeed のインフラ上で推奨されるオープンウェイト版から始めてください。プロンプトから始めるなら wavespeed-ai/minimax-h3/text-to-video、元画像をアニメーション化するなら wavespeed-ai/minimax-h3/image-to-video、ビジュアルリファレンスで被写体のアイデンティティ、見た目、スタイルをガイドしたいなら wavespeed-ai/minimax-h3/reference-to-video です。公式の minimax/h3 エンドポイントも同じファミリーで利用できます。
MiniMax H3 は画像から動画の生成に対応していますか?
はい。wavespeed-ai/minimax-h3/image-to-video エンドポイントは、最初のフレームの画像(必要に応じて最後のフレームも)を、ネイティブステレオオーディオ付きのまとまりある動画にアニメーション化します。商品写真、キャンペーンアート、キャラクター画像、コンセプトフレームなど、既存のビジュアルのアニメーション化に適しています。
MiniMax H3 のリファレンスから動画とは何ですか?
MiniMax H3 のリファレンスから動画は、最大9枚の参照画像、3本の参照動画、3本の参照オーディオを使い、被写体、キャラクターのアイデンティティ、スタイル、シーンの語り口をガイドして新しい動画を生成します。繰り返し登場するキャラクターや、ブランドに一貫したクリエイティブワークフローに最適な H3 エンドポイントです。
WaveSpeedAI で MiniMax H3 API を使うにはどうすればよいですか?
素材に合った H3 エンドポイントを選び、WaveSpeedAI のAPIキーで認証して、モデル入力をJSONで送信します。返された予測IDまたは結果URLで、生成された動画を取得できます。このページのエンドポイントカードとコードサンプルに、現在のリクエストスキーマが表示されています。
MiniMax H3 APIとは何ですか?
MiniMax H3は、MiniMaxの動画生成モデルで、WaveSpeedAI上でREST APIとして提供されています。MiniMax H3 API は、テキストから動画、画像から動画、リファレンスから動画のシネマティックな生成に対応し、ネイティブステレオオーディオ、優れたモーション品質、被写体の一貫性、シーンのまとまりを備えています。オープンウェイト版を WaveSpeed のインフラで、または公式の MiniMax エンドポイントを、1つの WaveSpeedAI API から実行できます。プログラムから呼び出すことも、上にリンクされたプレイグラウンドで試すこともできます。
MiniMax H3 APIはどう呼び出しますか?
WaveSpeedAIのアカウントに登録し、/accesskeyからAPIキーをコピーして、入力をJSONとしてhttps://api.wavespeed.ai/api/v3/wavespeed-ai/minimax-h3/text-to-videoにPOSTします。エンドポイントは予測IDを返します。結果のエンドポイントを約2秒ごとにポーリングし、長時間かかるタスクでは間隔を広げ、終端ステータスになったら停止してください。本番運用向けのPython / Node.js / cURLのサンプルは上にあります。
MiniMax H3 APIの料金はいくらですか?
MiniMax H3は1回あたり$0.02からです。実際の料金は、設定するパラメーター(解像度、長さ、出力数、参照入力)に応じて変わります。プレイグラウンドの「生成」ボタンの横に表示されるリアルタイムの料金プレビューで、現在の入力での正確な料金を確認できます。
MiniMax H3にはどのバリアントがありますか?
WaveSpeedAIでは、20個のMiniMax H3エンドポイントが利用可能です:wavespeed-ai/minimax-h3/controlnet-union, wavespeed-ai/minimax-h3-singularity/image-to-video-lora, wavespeed-ai/minimax-h3-singularity/image-to-video, wavespeed-ai/minimax-h3-singularity/reference-to-video-lora, wavespeed-ai/minimax-h3-singularity/reference-to-video, wavespeed-ai/minimax-h3/image-edit-lora, wavespeed-ai/minimax-h3/text-to-image-lora, wavespeed-ai/minimax-h3/image-editほか。各バリアントには、専用のプレイグラウンドページと料金があります。
MiniMax H3の出力を商用利用できますか?
商用利用の権利は、MiniMaxのモデルライセンスに従います。MiniMaxのほとんどのモデルは出力の商用利用を認めています。具体的なライセンスの概要は各モデルのプレイグラウンドページで、プラットフォーム全体の条件はWaveSpeedAIの利用規約でご確認ください。
なぜ直接ではなくWaveSpeedAIでMiniMax H3を使うのですか?
MiniMax H3と、他のプロバイダーの1,000以上のAIモデルを、1つのAPIキー、1つの請求アカウントで利用できます。ベンダーごとのSDK設定も、個別のレート制限も、ベンダーごとの連携コードの書き直しも不要です。料金は、通常MiniMaxの直接のAPIと同等かそれ以下です。
提供元
MiniMaxについて
WaveSpeedAIにおけるMiniMax H3と、MiniMaxのモデルラインナップ全体を手がけるチーム。
MiniMax は、Hailuo の動画生成と音声モデルで知られる、中国のAIラボです。Hailuo 2.3 は、Standard、Pro、Fast の階層で、物理を考慮したテキストから動画と画像から動画を提供し、クリエイターやマーケティングのワークフロー向けに、2.5倍の効率と、複雑な指示への高い正確さを打ち出しています。
WaveSpeedAIでMiniMax H3を使った開発を始めよう
登録時に無料のスタータークレジットを進呈。MiniMaxをはじめ、あらゆるプロバイダーの1,000以上のAIモデルを、1つのAPIキーで。