Nano Banana 2.1 公開中 — Google 最新モデル | 今すぐ試す →
MiniMax動画 API1回あたり$0.02から

MiniMax H3 API

MiniMax H3 API は、テキストから動画、画像から動画、リファレンスから動画のシネマティックな生成に対応し、ネイティブステレオオーディオ、優れたモーション品質、被写体の一貫性、シーンのまとまりを備えています。オープンウェイト版を WaveSpeed のインフラで、または公式の MiniMax エンドポイントを、1つの WaveSpeedAI API から実行できます。

テキストプロンプト、静止画、ビジュアルリファレンスからシネマティックなAI動画を生成します。WaveSpeed のインフラ上で推奨されるオープンウェイト版(wavespeed-ai/minimax-h3)は、480p/540p/768p/1080pでまとまりのある動画を、ネイティブステレオオーディオ付き・3〜15秒・リーズナブルな秒単位料金で生成します。公式の MiniMax エンドポイントも同じファミリーで利用できます。

概要

MiniMax H3 APIについて

MiniMax H3でできること、MiniMaxのモデルラインナップにおける位置づけ、そして多くのチームに選ばれる理由。

MiniMax H3はMiniMaxの動画生成モデルで、WaveSpeedAIのREST APIから利用できます。MiniMax H3 API は、テキストから動画、画像から動画、リファレンスから動画のシネマティックな生成に対応し、ネイティブステレオオーディオ、優れたモーション品質、被写体の一貫性、シーンのまとまりを備えています。オープンウェイト版を WaveSpeed のインフラで、または公式の MiniMax エンドポイントを、1つの WaveSpeedAI API から実行できます。

テキストプロンプト、静止画、ビジュアルリファレンスからシネマティックなAI動画を生成します。WaveSpeed のインフラ上で推奨されるオープンウェイト版(wavespeed-ai/minimax-h3)は、480p/540p/768p/1080pでまとまりのある動画を、ネイティブステレオオーディオ付き・3〜15秒・リーズナブルな秒単位料金で生成します。公式の MiniMax エンドポイントも同じファミリーで利用できます。

WaveSpeedAIのMiniMax H3ファミリーには、Motion-Control, Image-To-Video, Reference-To-Video, Image-To-Image, Text-To-Image, Video-To-Video, Video-Extend, Text-To-Videoのワークフローをカバーする20個のRESTエンドポイントがあります。各バリアントには、それぞれ独自の料金、パラメーター、作例があります。入力の種類と本番環境の制約に合うものを選ぶか、同じAPIキーで複数を呼び出して、多段階のパイプラインを組み立ててください。

WaveSpeedAIの他の1,000以上のAIモデルと同じAPIキー、請求アカウント、レート制限の枠組みでMiniMax H3を実行できます。ベンダーごとの設定も、プロバイダーごとのSDKも、ベンダーごとのレート制限も不要です。1つの連携で、テキストから画像、テキストから動画から、音声合成、3D生成、高画質化、編集まで、すべてカバーできます。

仕様

MiniMax H3 APIの機能とリリース状況

APIを選ぶ前に開発者が調べる、モデル固有の詳細情報:提供状況、想定される出力の長さ、参照入力への対応、そして現在利用できる代替エンドポイント。

テキストから動画

プロンプト駆動の生成

wavespeed-ai/minimax-h3/text-to-video で、文章のクリエイティブブリーフから、480p/540p/768p/1080pのまとまりある動画を生成します。ネイティブステレオオーディオ、3〜15秒の長さ、柔軟なアスペクト比に対応しています。

画像から動画

静止画をアニメーション化

wavespeed-ai/minimax-h3/image-to-video で、最初のフレームの画像(必要に応じて最後のフレームも)をアニメーション化し、被写体、フレーミング、アートディレクションをモーションに引き継いだ、ネイティブステレオオーディオ付きの動画を生成します。

リファレンスから動画

アイデンティティとスタイルをガイド

wavespeed-ai/minimax-h3/reference-to-video で、最大9枚の参照画像、3本の参照動画、3本の参照オーディオを使って生成をガイドします。被写体、キャラクターのアイデンティティ、スタイルを一貫させたいときに使います。

API連携

1つの WaveSpeedAI ワークフロー

WaveSpeedAI のAPIキーで H3 の予測を送信し、標準の予測ライフサイクルで各リクエストを追跡して、結果レスポンスから生成された動画を取得します。WaveSpeed ホストの版も公式の MiniMax エンドポイントも、同じワークフローで扱えます。

エンドポイント

MiniMax H3のAPIエンドポイント一覧

WaveSpeedAIで現在利用可能なMiniMax H3のエンドポイントは20個です。ワークフローに合ったバリアントを選んでください。

50%オフmotion-controlMinimax H3 Controlnet Union — MiniMaxのMiniMax H3 motion-controlプレビュー

Minimax H3 Controlnet Union

MiniMax H3 Open Weights ControlNet Union generates a new video that follows the motion and composition of a source video. Pose, depth, edges, lines, scribble or grayscale structure is extracted from the source automatically and guides the output, optionally with reference images for the subject or style, with native stereo audio generated in the same pass. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

$0.30$0.15から
50%オフimage-to-videoMinimax H3 Singularity Image To Video Lora — MiniMaxのMiniMax H3 image-to-videoプレビュー

Minimax H3 Singularity Image To Video Lora

MiniMax H3 Singularity Open Weights Image to Video animates a first-frame image, optionally with last-frame guidance, into coherent 480P / 540P / 768P videos with native stereo audio, 5-15 second duration, flexible aspect ratios, and per-second billing. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

$0.3125$0.1563から
50%オフimage-to-videoMinimax H3 Singularity Image To Video — MiniMaxのMiniMax H3 image-to-videoプレビュー

Minimax H3 Singularity Image To Video

MiniMax H3 Singularity Open Weights Image-to-Video animates a first-frame image, with optional last-frame guidance, into coherent 480P / 540P / 768P videos with native stereo audio, 5-15 second duration, flexible aspect ratios, and per-second billing. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

$0.25$0.125から
50%オフreference-to-videoMinimax H3 Singularity Reference To Video Lora — MiniMaxのMiniMax H3 reference-to-videoプレビュー

Minimax H3 Singularity Reference To Video Lora

MiniMax H3 Singularity Open Weights Reference to Video with custom LoRA support generates coherent 480P / 540P / 768P videos from prompts and multimodal references, guided by up to 9 reference images, 3 reference videos, and 3 reference audios, with native stereo audio and flexible reference-based video generation. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

$0.30$0.15から
50%オフreference-to-videoMinimax H3 Singularity Reference To Video — MiniMaxのMiniMax H3 reference-to-videoプレビュー

Minimax H3 Singularity Reference To Video

MiniMax H3 Singularity Open Weights Reference to Video generates coherent 480P / 540P / 768P videos from prompts and multimodal references, guided by up to 9 reference images, 3 reference videos, and 3 reference audios, with native stereo audio and flexible reference-based video generation on WaveSpeedAI infrastructure. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

$0.25$0.125から
image-to-imageMinimax H3 Image Edit Lora — MiniMaxのMiniMax H3 image-to-imageプレビュー

Minimax H3 Image Edit Lora

MiniMax H3 Open Weights Image Edit re-renders a subject from up to 9 reference images into a new scene, outfit or style from a text instruction, preserving identity at 1K or 2K resolution. Loads up to 3 custom LoRA weights per request. Ready-to-use REST inference API, best performance, no cold starts, affordable pricing.

$0.03から
text-to-imageMinimax H3 Text To Image Lora — MiniMaxのMiniMax H3 text-to-imageプレビュー

Minimax H3 Text To Image Lora

MiniMax H3 Open Weights Text-to-Image generates photorealistic, cinema-grade stills from a text prompt at 1K or 2K resolution with fifteen aspect ratios. Loads up to 3 custom LoRA weights per request. Ready-to-use REST inference API, best performance, no cold starts, affordable pricing.

$0.02から
image-to-imageMinimax H3 Image Edit — MiniMaxのMiniMax H3 image-to-imageプレビュー

Minimax H3 Image Edit

MiniMax H3 Open Weights Image Edit re-renders a subject from up to 9 reference images into a new scene, outfit or style from a text instruction, preserving identity at 1K or 2K resolution. Ready-to-use REST inference API, best performance, no cold starts, affordable pricing.

$0.03から
text-to-imageMinimax H3 Text To Image — MiniMaxのMiniMax H3 text-to-imageプレビュー

Minimax H3 Text To Image

MiniMax H3 Open Weights Text-to-Image generates photorealistic, cinema-grade stills from a text prompt at 1K or 2K resolution with fifteen aspect ratios. Ready-to-use REST inference API, best performance, no cold starts, affordable pricing.

$0.02から
50%オフvideo-to-videoMinimax H3 Video Edit — MiniMaxのMiniMax H3 video-to-videoプレビュー

Minimax H3 Video Edit

MiniMax H3 Open Weights Video-Edit edits an input video from a natural-language prompt. The input video drives subject identity, composition, and motion while the model rewrites lighting, style, weather, environment, or specific elements as instructed, with native stereo audio generated in the same pass. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

$0.25$0.125から
50%オフvideo-extendMinimax H3 Video Extend — MiniMaxのMiniMax H3 video-extendプレビュー

Minimax H3 Video Extend

MiniMax H3 Open Weights Video-Extend appends a new cinematic continuation to an existing video. A fresh segment with native stereo audio is generated from the input video's last frame and a natural-language prompt, then the original and new segment are concatenated into a single output. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

$0.20$0.10から
50%オフreference-to-videoMinimax H3 Reference To Video Lora — MiniMaxのMiniMax H3 reference-to-videoプレビュー

Minimax H3 Reference To Video Lora

MiniMax H3 Open Weights Reference to Video with custom LoRA support generates coherent 480P / 540P / 768P videos from prompts and multimodal references, guided by up to 9 reference images, 3 reference videos, and 3 reference audios, with native stereo audio and flexible reference-based video generation. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

$0.30$0.15から
50%オフimage-to-videoMinimax H3 Image To Video Lora — MiniMaxのMiniMax H3 image-to-videoプレビュー

Minimax H3 Image To Video Lora

MiniMax H3 Open Weights Image to Video animates a first-frame image, optionally with last-frame guidance, into coherent 480P / 540P / 768P videos with native stereo audio, 5-15 second duration, flexible aspect ratios, and per-second billing. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

$0.25$0.125から
50%オフtext-to-videoMinimax H3 Text To Video Lora — MiniMaxのMiniMax H3 text-to-videoプレビュー

Minimax H3 Text To Video Lora

MiniMax H3 Open Weights Text to Video with custom LoRA support generates coherent videos from text prompts, with 480P / 540P / 768P output, native stereo audio, 5-15 second duration, flexible aspect ratios, and per-second billing. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

$0.25$0.125から
50%オフreference-to-videoMinimax H3 Reference To Video — MiniMaxのMiniMax H3 reference-to-videoプレビュー

Minimax H3 Reference To Video

MiniMax H3 Open Weights Reference to Video generates coherent 480P / 540P / 768P videos from prompts and multimodal references, guided by up to 9 reference images, 3 reference videos, and 3 reference audios, with native stereo audio and flexible reference-based video generation on WaveSpeedAI infrastructure. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

$0.25$0.125から
50%オフimage-to-videoMinimax H3 Image To Video — MiniMaxのMiniMax H3 image-to-videoプレビュー

Minimax H3 Image To Video

MiniMax H3 Open Weights Image to Video animates a first-frame image, optionally with last-frame guidance, into coherent 480P / 540P / 768P videos with native stereo audio, 5-15 second duration, flexible aspect ratios, and per-second billing. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

$0.20$0.10から
50%オフtext-to-videoMinimax H3 Text To Video — MiniMaxのMiniMax H3 text-to-videoプレビュー

Minimax H3 Text To Video

MiniMax H3 Open Weights Text to Video generates coherent videos from text prompts, with 480P / 540P / 768P output, native stereo audio, 5-15 second duration, flexible aspect ratios, and per-second billing. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

$0.20$0.10から
reference-to-videoH3 Reference To Video — MiniMaxのMiniMax H3 reference-to-videoプレビュー

H3 Reference To Video

MiniMax H3 Reference to Video generates coherent 2K videos from natural-language prompts and multimodal references, including images, videos, and audio, guiding subject consistency, motion, timing, visual style, and scene continuity. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

$0.70から
image-to-videoH3 Image To Video — MiniMaxのMiniMax H3 image-to-videoプレビュー

H3 Image To Video

MiniMax H3 Image to Video animates a first-frame image into a coherent 2K video, with natural-language motion instructions and optional last-frame control for consistent motion, scene continuity, and cinematic video generation. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

$0.70から
text-to-videoH3 Text To Video — MiniMaxのMiniMax H3 text-to-videoプレビュー

H3 Text To Video

MiniMax H3 Text to Video generates coherent 2K videos from text prompts, with flexible 5-15 second duration and adaptive or custom aspect ratios for cinematic scenes, creative videos, and production workflows. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

$0.70から

作例

MiniMax H3の実力を見る

MiniMax H3 APIで実際に生成された出力です。動画にカーソルを合わせるとプレビュー、クリックすると原寸のビューアで開きます。

使い方

MiniMax H3 APIの使い方

登録から生成完了まで4ステップ。Python、Node.js、cURLの完全なサンプルは、下のAPIセクションにあります。

  1. 01

    APIキーを取得

    WaveSpeedAIのアカウントに登録し、ダッシュボードからAPIキーをコピーします。新規アカウントには無料のスタータークレジットが付くので、課金が始まる前にプレイグラウンドを数十回実行できます。

  2. 02

    予測を送信

    入力をJSONとしてhttps://api.wavespeed.ai/api/v3/wavespeed-ai/minimax-h3/text-to-videoにPOSTします。エンドポイントはすぐに予測IDを返します。生成は非同期なので、推論中に接続を開いたままにする必要はありません。

  3. 03

    完了までポーリング

    https://api.wavespeed.ai/api/v3/predictions/{request_id}/resultにGETします。completedなら出力を返し、failed、cancelled、timeout、deletedならエラーで停止し、それ以外のステータスの間はポーリングを続けます。

  4. 04

    出力URLを読み取る

    ステータスが"completed"になったら、data.outputs[0]からURLを読み取ります。URLは、WaveSpeedAIのCDN上にある生成されたメディアを指します。呼び出したMiniMax H3のバリアントに応じて、画像、動画、音声、3Dファイルのいずれかです。

活用例

MiniMax H3で作れるもの

開発者やクリエイターがMiniMax H3 APIでよく使うワークフロー。

01

オリジナルのコンセプト向けテキストから動画

wavespeed-ai/minimax-h3/text-to-video で、プロンプトからオリジナルのシネマティックなシーンを生成。480p/540p/768p/1080pのまとまりある出力、ネイティブステレオオーディオ、柔軟なアスペクト比に対応します。コンセプトの可視化、ストーリーテリング、キャンペーンのアイデア、既存素材から始める必要のないショットに最適です。

text-to-videopromptcinematic
02

商品紹介向けの画像から動画

商品写真、キャンペーンの静止画、キービジュアル、キャラクター画像を、元のビジュアルの方向性を保ったままアニメーション化します。商品のお披露目、SNS広告、モーション中心のランディングコンテンツに便利です。

image-to-videoproductadvertising
03

一貫した被写体のためのリファレンスから動画

最大9枚の参照画像、3本の参照動画、3本の参照オーディオで、キャラクターのアイデンティティ、被写体の見た目、シーンデザイン、スタイルをガイドできます。リファレンスから動画は、繰り返し登場するキャラクター、ブランドビジュアル、関連するクリエイティブセットのための H3 ワークフローです。

reference-to-videoconsistencyidentity
04

SNS・広告クリエイティブ

キャンペーン、ローンチ、SNSフィード、パフォーマンスマーケティング向けに、ショートフォームのクリエイティブを生成します。テキスト、完成した静止画、承認済みのビジュアルリファレンスのどこから始めても、APIプラットフォームを変える必要はありません。

social-videoad-creativemarketing
05

キャラクターシーンとビジュアルストーリーテリング

キャラクター主体のシーン、ストーリーのビート、ムード映像、プリビズを制作できます。素材に合わせてエンドポイントを選びます。新しいシーンにはテキスト、アニメーション化には画像、より強い連続性にはリファレンスです。

charactersstorytellingprevisualization
06

スケーラブルなAI動画生成パイプライン

WaveSpeedAI を通じて MiniMax H3 をアプリや自動化されたクリエイティブワークフローに組み込めます。H3 の3つの生成モードすべてで、1つのAPIキーと共通の予測ライフサイクルを使えます。

apiautomationproduction

コツ

MiniMax H3のプロンプトのコツ

MiniMax H3からより良い出力を得るための実践的なアドバイス。本番のパイプラインで動画モデル全般に通用するパターンに基づいています。

  1. 01

    素材に合わせてエンドポイントを選ぶ

    オリジナルのプロンプトにはテキストから動画、1枚の元画像のアニメーション化には画像から動画、ビジュアルの参照でアイデンティティ、被写体の見た目、シーンデザイン、スタイルをガイドしたい場合にはリファレンスから動画を使います。

  2. 02

    モーションを時系列で記述する

    開始時の状態、アクション、環境の反応、カメラの動き、終了時の状態を、時系列で書いてください。つながりのない視覚的な形容詞を並べるより、明確な流れのほうが、H3 により強いモーションの計画を与えます。

  3. 03

    構図が承認済みなら画像から動画を使う

    被写体、商品、フレーミング、アートディレクションが静止画で固まっているなら、テキストからシーンを作り直さず、画像から動画から始めてください。プロンプトは動きとカメラの挙動に集中させます。

  4. 04

    参照には、はっきりした役割を1つ与える

    リファレンスから動画では、参照が何を制御するのか(キャラクターのアイデンティティ、商品の見た目、ビジュアルスタイル、環境、構図)を明記します。同じ参照に、ショットの複数の相反する要素を定義させないでください。

  5. 05

    短いシーンでは、主なアクションを1つに絞る

    関連のない出来事を並べるより、1つのアクションに絞るほうが、まとまりよくレンダリングできます。ストーリーの区切りごとに別々のショットを生成し、コンセプトにより広い物語の範囲が必要なら、編集でつなげてください。

料金

MiniMax H3 APIの料金

料金は出力ごとです。最終的な請求額は、各バリアントのプレイグラウンドで設定するパラメーター(解像度、長さ、出力数、参照入力)に応じて変わります。

最低料金

$0.02/ 回

20個のエンドポイント · MiniMax H3

プレイグラウンドを開く
エンドポイントタイプ最低料金
wavespeed-ai/minimax-h3/controlnet-unionmotion-control50%オフ$0.30$0.15
wavespeed-ai/minimax-h3-singularity/image-to-video-loraimage-to-video50%オフ$0.3125$0.1563
wavespeed-ai/minimax-h3-singularity/image-to-videoimage-to-video50%オフ$0.25$0.125
wavespeed-ai/minimax-h3-singularity/reference-to-video-lorareference-to-video50%オフ$0.30$0.15
wavespeed-ai/minimax-h3-singularity/reference-to-videoreference-to-video50%オフ$0.25$0.125
wavespeed-ai/minimax-h3/image-edit-loraimage-to-image$0.03
wavespeed-ai/minimax-h3/text-to-image-loratext-to-image$0.02
wavespeed-ai/minimax-h3/image-editimage-to-image$0.03
wavespeed-ai/minimax-h3/text-to-imagetext-to-image$0.02
wavespeed-ai/minimax-h3/video-editvideo-to-video50%オフ$0.25$0.125
wavespeed-ai/minimax-h3/video-extendvideo-extend50%オフ$0.20$0.10
wavespeed-ai/minimax-h3/reference-to-video-lorareference-to-video50%オフ$0.30$0.15
wavespeed-ai/minimax-h3/image-to-video-loraimage-to-video50%オフ$0.25$0.125
wavespeed-ai/minimax-h3/text-to-video-loratext-to-video50%オフ$0.25$0.125
wavespeed-ai/minimax-h3/reference-to-videoreference-to-video50%オフ$0.25$0.125
wavespeed-ai/minimax-h3/image-to-videoimage-to-video50%オフ$0.20$0.10
wavespeed-ai/minimax-h3/text-to-videotext-to-video50%オフ$0.20$0.10
minimax/h3/reference-to-videoreference-to-video$0.70
minimax/h3/image-to-videoimage-to-video$0.70
minimax/h3/text-to-videotext-to-video$0.70

API

MiniMax H3 APIを呼び出す

wavespeed.ai/accesskeyでAPIキーに登録し、RESTで予測を送信します。プレイグラウンドでは、入力の組み合わせに応じて、そのまま貼り付けられるサンプルが生成されます。

POSThttps://api.wavespeed.ai/api/v3/wavespeed-ai/minimax-h3/text-to-video

# 1. Submit the prediction.
SUBMIT_RESPONSE=$(curl --silent --show-error --fail-with-body \
  -X POST "https://api.wavespeed.ai/api/v3/wavespeed-ai/minimax-h3/text-to-video" \
  -H "Content-Type: application/json" \
  -H "Authorization: Bearer $WAVESPEED_API_KEY" \
  -d '{
    "prompt": "A cinematic shot of a city at sunset, soft golden light",
    "aspect_ratio": "16:9",
    "resolution": "480p",
    "duration": 5
}')

TASK=$(printf '%s' "$SUBMIT_RESPONSE" | jq 'if has("data") then .data else . end')
PREDICTION_ID=$(printf '%s' "$TASK" | jq -r '.id')
if [ -z "$PREDICTION_ID" ] || [ "$PREDICTION_ID" = "null" ]; then
  printf 'Submission response did not contain a prediction id
' >&2
  exit 1
fi
RESULT_URL="https://api.wavespeed.ai/api/v3/predictions/$PREDICTION_ID/result"

# 2. Poll until the prediction finishes.
while true; do
  RESPONSE=$(curl --silent --show-error --fail-with-body "$RESULT_URL" \
    -H "Authorization: Bearer $WAVESPEED_API_KEY")
  RESULT=$(printf '%s' "$RESPONSE" | jq 'if has("data") then .data else . end')
  STATUS=$(printf '%s' "$RESULT" | jq -r '.status')
  case "$STATUS" in
    completed) printf '%s\n' "$RESULT" | jq '.outputs'; break ;;
    failed|cancelled|timeout|deleted) printf '%s\n' "$RESULT" | jq . >&2; exit 1 ;;
    *) sleep 2 ;;
  esac
done

比較

MiniMax H3と他モデルの比較

WaveSpeedAI上の類似モデルではなくMiniMax H3を選ぶべきケース。

MiniMax H3 vs Hailuo 2.3

Hailuo 2.3 は、Standard、Pro、Fast、Fast Pro のバリアントにまたがる階層型のテキストから動画・画像から動画エンドポイントを提供します。MiniMax H3 は3つのエンドポイントに絞ったスイートで、ビジュアルガイドと被写体の一貫性のためのリファレンスから動画を、主要なワークフローとして備えています。

MiniMax H3 vs Seedance 2.0

Seedance 2.0 は、テキストから動画、画像から動画、video-edit、video-extend、ネイティブオーディオ、複数のパフォーマンス階層を備えた幅広い制作向けファミリーです。テキスト、元画像、ビジュアルリファレンスからの生成が中心のワークフローなら、MiniMax H3 のほうがシンプルな選択肢です。

MiniMax H3 vs Kling 3.0

Kling 3.0 は、階層化された出力品質と専用のモーションコントロールエンドポイントに力を入れています。MiniMax H3 は素材の種類に沿ってAPIを構成しており、アイデンティティ、被写体、スタイルをガイドする生成のための専用のリファレンスから動画ルートがあります。

FAQ

MiniMax H3 API — よくある質問

料金、ライセンス、連携など、WaveSpeedAIでMiniMax H3を実行する際のよくある質問。

MiniMax H3 API とは何ですか?

MiniMax H3 API は、WaveSpeedAI で利用できる3モデル構成のAI動画生成スイートです。プロンプト駆動のシーン向けのテキストから動画、静止画をアニメーション化する画像から動画、ビジュアルリファレンスでガイドされる生成のためのリファレンスから動画に対応しています。

MiniMax H3 API ではどのエンドポイントを使えばよいですか?

まずは WaveSpeed のインフラ上で推奨されるオープンウェイト版から始めてください。プロンプトから始めるなら wavespeed-ai/minimax-h3/text-to-video、元画像をアニメーション化するなら wavespeed-ai/minimax-h3/image-to-video、ビジュアルリファレンスで被写体のアイデンティティ、見た目、スタイルをガイドしたいなら wavespeed-ai/minimax-h3/reference-to-video です。公式の minimax/h3 エンドポイントも同じファミリーで利用できます。

MiniMax H3 は画像から動画の生成に対応していますか?

はい。wavespeed-ai/minimax-h3/image-to-video エンドポイントは、最初のフレームの画像(必要に応じて最後のフレームも)を、ネイティブステレオオーディオ付きのまとまりある動画にアニメーション化します。商品写真、キャンペーンアート、キャラクター画像、コンセプトフレームなど、既存のビジュアルのアニメーション化に適しています。

MiniMax H3 のリファレンスから動画とは何ですか?

MiniMax H3 のリファレンスから動画は、最大9枚の参照画像、3本の参照動画、3本の参照オーディオを使い、被写体、キャラクターのアイデンティティ、スタイル、シーンの語り口をガイドして新しい動画を生成します。繰り返し登場するキャラクターや、ブランドに一貫したクリエイティブワークフローに最適な H3 エンドポイントです。

WaveSpeedAI で MiniMax H3 API を使うにはどうすればよいですか?

素材に合った H3 エンドポイントを選び、WaveSpeedAI のAPIキーで認証して、モデル入力をJSONで送信します。返された予測IDまたは結果URLで、生成された動画を取得できます。このページのエンドポイントカードとコードサンプルに、現在のリクエストスキーマが表示されています。

MiniMax H3 APIとは何ですか?

MiniMax H3は、MiniMaxの動画生成モデルで、WaveSpeedAI上でREST APIとして提供されています。MiniMax H3 API は、テキストから動画、画像から動画、リファレンスから動画のシネマティックな生成に対応し、ネイティブステレオオーディオ、優れたモーション品質、被写体の一貫性、シーンのまとまりを備えています。オープンウェイト版を WaveSpeed のインフラで、または公式の MiniMax エンドポイントを、1つの WaveSpeedAI API から実行できます。プログラムから呼び出すことも、上にリンクされたプレイグラウンドで試すこともできます。

MiniMax H3 APIはどう呼び出しますか?

WaveSpeedAIのアカウントに登録し、/accesskeyからAPIキーをコピーして、入力をJSONとしてhttps://api.wavespeed.ai/api/v3/wavespeed-ai/minimax-h3/text-to-videoにPOSTします。エンドポイントは予測IDを返します。結果のエンドポイントを約2秒ごとにポーリングし、長時間かかるタスクでは間隔を広げ、終端ステータスになったら停止してください。本番運用向けのPython / Node.js / cURLのサンプルは上にあります。

MiniMax H3 APIの料金はいくらですか?

MiniMax H3は1回あたり$0.02からです。実際の料金は、設定するパラメーター(解像度、長さ、出力数、参照入力)に応じて変わります。プレイグラウンドの「生成」ボタンの横に表示されるリアルタイムの料金プレビューで、現在の入力での正確な料金を確認できます。

MiniMax H3にはどのバリアントがありますか?

WaveSpeedAIでは、20個のMiniMax H3エンドポイントが利用可能です:wavespeed-ai/minimax-h3/controlnet-union, wavespeed-ai/minimax-h3-singularity/image-to-video-lora, wavespeed-ai/minimax-h3-singularity/image-to-video, wavespeed-ai/minimax-h3-singularity/reference-to-video-lora, wavespeed-ai/minimax-h3-singularity/reference-to-video, wavespeed-ai/minimax-h3/image-edit-lora, wavespeed-ai/minimax-h3/text-to-image-lora, wavespeed-ai/minimax-h3/image-editほか。各バリアントには、専用のプレイグラウンドページと料金があります。

MiniMax H3の出力を商用利用できますか?

商用利用の権利は、MiniMaxのモデルライセンスに従います。MiniMaxのほとんどのモデルは出力の商用利用を認めています。具体的なライセンスの概要は各モデルのプレイグラウンドページで、プラットフォーム全体の条件はWaveSpeedAIの利用規約でご確認ください。

なぜ直接ではなくWaveSpeedAIでMiniMax H3を使うのですか?

MiniMax H3と、他のプロバイダーの1,000以上のAIモデルを、1つのAPIキー、1つの請求アカウントで利用できます。ベンダーごとのSDK設定も、個別のレート制限も、ベンダーごとの連携コードの書き直しも不要です。料金は、通常MiniMaxの直接のAPIと同等かそれ以下です。

提供元

MiniMaxについて

WaveSpeedAIにおけるMiniMax H3と、MiniMaxのモデルラインナップ全体を手がけるチーム。

MiniMax は、Hailuo の動画生成と音声モデルで知られる、中国のAIラボです。Hailuo 2.3 は、Standard、Pro、Fast の階層で、物理を考慮したテキストから動画と画像から動画を提供し、クリエイターやマーケティングのワークフロー向けに、2.5倍の効率と、複雑な指示への高い正確さを打ち出しています。

WaveSpeedAIでMiniMax H3を使った開発を始めよう

登録時に無料のスタータークレジットを進呈。MiniMaxをはじめ、あらゆるプロバイダーの1,000以上のAIモデルを、1つのAPIキーで。