MiniMax H3 API
MiniMax H3 API để tạo video điện ảnh từ văn bản, ảnh và tham chiếu với âm thanh stereo gốc, chất lượng chuyển động mạnh, nhất quán chủ thể và mạch lạc cảnh. Chạy bản open-weights trên hạ tầng WaveSpeed hoặc các endpoint MiniMax chính thức qua một API WaveSpeedAI duy nhất.
Tạo video AI điện ảnh từ prompt văn bản, ảnh tĩnh hoặc tham chiếu hình ảnh. Bản triển khai open-weights được khuyên dùng trên hạ tầng WaveSpeed (wavespeed-ai/minimax-h3) cho video mạch lạc 480p/540p/768p/1080p với âm thanh stereo gốc, thời lượng 3–15 giây và giá theo giây phải chăng — các endpoint MiniMax chính thức cũng có trong cùng dòng.
Tổng quan
Giới thiệu về MiniMax H3 API
MiniMax H3 làm được gì, nằm ở đâu trong dòng mô hình của MiniMax, và vì sao các đội ngũ chọn nó.
MiniMax H3 là mô hình tạo video từ MiniMax, có sẵn qua REST API của WaveSpeedAI. MiniMax H3 API để tạo video điện ảnh từ văn bản, ảnh và tham chiếu với âm thanh stereo gốc, chất lượng chuyển động mạnh, nhất quán chủ thể và mạch lạc cảnh. Chạy bản open-weights trên hạ tầng WaveSpeed hoặc các endpoint MiniMax chính thức qua một API WaveSpeedAI duy nhất.
Tạo video AI điện ảnh từ prompt văn bản, ảnh tĩnh hoặc tham chiếu hình ảnh. Bản triển khai open-weights được khuyên dùng trên hạ tầng WaveSpeed (wavespeed-ai/minimax-h3) cho video mạch lạc 480p/540p/768p/1080p với âm thanh stereo gốc, thời lượng 3–15 giây và giá theo giây phải chăng — các endpoint MiniMax chính thức cũng có trong cùng dòng.
Họ MiniMax H3 trên WaveSpeedAI cung cấp 20 endpoint REST bao gồm 8 quy trình Motion-Control, Image-To-Video, Reference-To-Video, Image-To-Image, Text-To-Image, Video-To-Video, Video-Extend, Text-To-Video. Mỗi biến thể có giá, các tham số điều chỉnh và kết quả mẫu riêng — hãy chọn biến thể phù hợp với dạng đầu vào và ràng buộc sản xuất của bạn, hoặc gọi nhiều biến thể từ cùng một API key để ghép các pipeline nhiều bước.
Chạy MiniMax H3 bằng cùng API key, tài khoản thanh toán và hạn mức tốc độ bạn dùng cho hơn 1.000 mô hình AI khác trên WaveSpeedAI. Không cần thiết lập nhà cung cấp riêng, không cần SDK cho từng nhà cung cấp, không có hạn mức tốc độ riêng của từng hãng — một lần tích hợp bao quát mọi thứ, từ text-to-image và text-to-video đến tổng hợp âm thanh, tạo 3D, nâng cấp và chỉnh sửa.
Thông số
Khả năng và trạng thái phát hành của MiniMax H3 API
Các chi tiết riêng của mô hình mà nhà phát triển tìm kiếm trước khi chọn API: tình trạng khả dụng, độ dài đầu ra dự kiến, hỗ trợ tham chiếu và phương án thay thế đang hoạt động.
Text to video
Tạo theo prompt
Dùng wavespeed-ai/minimax-h3/text-to-video để biến brief sáng tạo bằng văn bản thành video mạch lạc 480p/540p/768p/1080p với âm thanh stereo gốc, thời lượng 3–15 giây và tỷ lệ khung hình linh hoạt.
Image to video
Làm động ảnh tĩnh
Dùng wavespeed-ai/minimax-h3/image-to-video để làm động ảnh khung đầu — có thể kèm khung cuối — thành video mạch lạc với âm thanh stereo gốc, đồng thời đưa chủ thể, bố cục và định hướng nghệ thuật của ảnh vào chuyển động.
Reference to video
Dẫn dắt danh tính và phong cách
Dùng wavespeed-ai/minimax-h3/reference-to-video để dẫn dắt việc tạo bằng tối đa 9 ảnh tham chiếu, 3 video tham chiếu và 3 âm thanh tham chiếu khi chủ thể, danh tính nhân vật hoặc phong cách cần nhất quán.
Tích hợp API
Một quy trình WaveSpeedAI
Gửi prediction H3 bằng API key WaveSpeedAI của bạn, theo dõi từng request qua vòng đời prediction tiêu chuẩn và lấy video đã tạo từ phản hồi kết quả — cùng một quy trình cho endpoint do WaveSpeed lưu trữ và endpoint MiniMax chính thức.
Endpoint
Tất cả endpoint API MiniMax H3
20 endpoint MiniMax H3 hiện có sẵn trên WaveSpeedAI — hãy chọn biến thể phù hợp với quy trình của bạn.
/filters:quality(82)/media/images/1790573118566544886_jzIR0oxG.webp)
Minimax H3 Controlnet Union
MiniMax H3 Open Weights ControlNet Union generates a new video that follows the motion and composition of a source video. Pose, depth, edges, lines, scribble or grayscale structure is extracted from the source automatically and guides the output, optionally with reference images for the subject or style, with native stereo audio generated in the same pass. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/1789888783447157992_hG8iqAJS.webp)
Minimax H3 Singularity Image To Video Lora
MiniMax H3 Singularity Open Weights Image to Video animates a first-frame image, optionally with last-frame guidance, into coherent 480P / 540P / 768P videos with native stereo audio, 5-15 second duration, flexible aspect ratios, and per-second billing. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/1789888766253281059_pb1bluDL.webp)
Minimax H3 Singularity Image To Video
MiniMax H3 Singularity Open Weights Image-to-Video animates a first-frame image, with optional last-frame guidance, into coherent 480P / 540P / 768P videos with native stereo audio, 5-15 second duration, flexible aspect ratios, and per-second billing. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/1789742735070040548_71aJS2bl.webp)
Minimax H3 Singularity Reference To Video Lora
MiniMax H3 Singularity Open Weights Reference to Video with custom LoRA support generates coherent 480P / 540P / 768P videos from prompts and multimodal references, guided by up to 9 reference images, 3 reference videos, and 3 reference audios, with native stereo audio and flexible reference-based video generation. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/1789742804979721138_b4T1aktE.webp)
Minimax H3 Singularity Reference To Video
MiniMax H3 Singularity Open Weights Reference to Video generates coherent 480P / 540P / 768P videos from prompts and multimodal references, guided by up to 9 reference images, 3 reference videos, and 3 reference audios, with native stereo audio and flexible reference-based video generation on WaveSpeedAI infrastructure. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/1788684887040987486_ILX6fU3e.webp)
Minimax H3 Image Edit Lora
MiniMax H3 Open Weights Image Edit re-renders a subject from up to 9 reference images into a new scene, outfit or style from a text instruction, preserving identity at 1K or 2K resolution. Loads up to 3 custom LoRA weights per request. Ready-to-use REST inference API, best performance, no cold starts, affordable pricing.
/filters:quality(82)/media/images/1788684121915845410_PlPY7gqA.webp)
Minimax H3 Text To Image Lora
MiniMax H3 Open Weights Text-to-Image generates photorealistic, cinema-grade stills from a text prompt at 1K or 2K resolution with fifteen aspect ratios. Loads up to 3 custom LoRA weights per request. Ready-to-use REST inference API, best performance, no cold starts, affordable pricing.
/filters:quality(82)/media/images/1788684872529921426_LGPY7gpy.webp)
Minimax H3 Image Edit
MiniMax H3 Open Weights Image Edit re-renders a subject from up to 9 reference images into a new scene, outfit or style from a text instruction, preserving identity at 1K or 2K resolution. Ready-to-use REST inference API, best performance, no cold starts, affordable pricing.
/filters:quality(82)/media/images/1788684072554630514_Lg4clvEO.webp)
Minimax H3 Text To Image
MiniMax H3 Open Weights Text-to-Image generates photorealistic, cinema-grade stills from a text prompt at 1K or 2K resolution with fifteen aspect ratios. Ready-to-use REST inference API, best performance, no cold starts, affordable pricing.
/filters:quality(82)/media/images/1788329903545692055_lEMW5fpy.webp)
Minimax H3 Video Edit
MiniMax H3 Open Weights Video-Edit edits an input video from a natural-language prompt. The input video drives subject identity, composition, and motion while the model rewrites lighting, style, weather, environment, or specific elements as instructed, with native stereo audio generated in the same pass. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/1788333421627652695_cNUHuh5T.webp)
Minimax H3 Video Extend
MiniMax H3 Open Weights Video-Extend appends a new cinematic continuation to an existing video. A fresh segment with native stereo audio is generated from the input video's last frame and a natural-language prompt, then the original and new segment are concatenated into a single output. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/1786952873106530692_cwjsCLT3.webp)
Minimax H3 Reference To Video Lora
MiniMax H3 Open Weights Reference to Video with custom LoRA support generates coherent 480P / 540P / 768P videos from prompts and multimodal references, guided by up to 9 reference images, 3 reference videos, and 3 reference audios, with native stereo audio and flexible reference-based video generation. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/1786952854061501695_t8heoxGQ.webp)
Minimax H3 Image To Video Lora
MiniMax H3 Open Weights Image to Video animates a first-frame image, optionally with last-frame guidance, into coherent 480P / 540P / 768P videos with native stereo audio, 5-15 second duration, flexible aspect ratios, and per-second billing. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/1788760867264586097_7BZ9isBK.webp)
Minimax H3 Text To Video Lora
MiniMax H3 Open Weights Text to Video with custom LoRA support generates coherent videos from text prompts, with 480P / 540P / 768P output, native stereo audio, 5-15 second duration, flexible aspect ratios, and per-second billing. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/1785814033019871155_vhrAKT2c.webp)
Minimax H3 Reference To Video
MiniMax H3 Open Weights Reference to Video generates coherent 480P / 540P / 768P videos from prompts and multimodal references, guided by up to 9 reference images, 3 reference videos, and 3 reference audios, with native stereo audio and flexible reference-based video generation on WaveSpeedAI infrastructure. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/1785815150161566057_2VF6wZoP.webp)
Minimax H3 Image To Video
MiniMax H3 Open Weights Image to Video animates a first-frame image, optionally with last-frame guidance, into coherent 480P / 540P / 768P videos with native stereo audio, 5-15 second duration, flexible aspect ratios, and per-second billing. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/1785814561183432100_tW5eoxGP.webp)
Minimax H3 Text To Video
MiniMax H3 Open Weights Text to Video generates coherent videos from text prompts, with 480P / 540P / 768P output, native stereo audio, 5-15 second duration, flexible aspect ratios, and per-second billing. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/1785415007635242852_UoxHQZ9H.webp)
H3 Reference To Video
MiniMax H3 Reference to Video generates coherent 2K videos from natural-language prompts and multimodal references, including images, videos, and audio, guiding subject consistency, motion, timing, visual style, and scene continuity. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/1785414895750297795_1JajtDMV.webp)
H3 Image To Video
MiniMax H3 Image to Video animates a first-frame image into a coherent 2K video, with natural-language motion instructions and optional last-frame control for consistent motion, scene continuity, and cinematic video generation. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/1785412796956877340_AktCLU4J.webp)
H3 Text To Video
MiniMax H3 Text to Video generates coherent 2K videos from text prompts, with flexible 5-15 second duration and adaptive or custom aspect ratios for cinematic scenes, creative videos, and production workflows. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
Ví dụ
Xem MiniMax H3 hoạt động
Kết quả thật do API MiniMax H3 tạo ra. Rê chuột lên video để xem trước, nhấp để mở trình xem kích thước đầy đủ.
Hướng dẫn
Cách dùng API MiniMax H3
Bốn bước từ đăng ký đến khi có kết quả. Ví dụ đầy đủ bằng Python, Node.js và cURL nằm ở phần API bên dưới.
- 01
Lấy API key
Đăng ký tài khoản WaveSpeedAI và sao chép API key từ bảng điều khiển. Tài khoản mới được tặng credit dùng thử miễn phí — đủ để chạy playground vài chục lần trước khi bắt đầu tính phí.
- 02
Gửi một prediction
POST đầu vào của bạn dưới dạng JSON tới https://api.wavespeed.ai/api/v3/wavespeed-ai/minimax-h3/text-to-video. Endpoint trả về prediction id ngay lập tức — việc tạo là bất đồng bộ nên bạn không phải giữ kết nối mở trong lúc suy luận.
- 03
Poll để chờ hoàn tất
GET https://api.wavespeed.ai/api/v3/predictions/{request_id}/result. Trả về kết quả khi completed; dừng và báo lỗi khi failed, cancelled, timeout hoặc deleted; tiếp tục poll với mọi trạng thái khác.
- 04
Đọc URL kết quả
Khi status là "completed", đọc URL từ data.outputs[0]. URL trỏ tới media bạn đã tạo trên CDN của WaveSpeedAI — ảnh, video, âm thanh hoặc tệp 3D tùy biến thể MiniMax H3 bạn đã gọi.
Trường hợp sử dụng
Bạn có thể xây dựng gì với MiniMax H3
Các quy trình phổ biến mà nhà phát triển và người sáng tạo dùng API MiniMax H3 cho.
Text-to-video cho ý tưởng gốc
Biến prompt thành một cảnh điện ảnh gốc với wavespeed-ai/minimax-h3/text-to-video — đầu ra 480p/540p/768p/1080p mạch lạc, âm thanh stereo gốc và tỷ lệ khung hình linh hoạt. Phù hợp để hình dung ý tưởng, kể chuyện, ý tưởng chiến dịch và các cảnh không cần bắt đầu từ hình có sẵn.
Image-to-video để giới thiệu sản phẩm
Làm động ảnh sản phẩm, ảnh chiến dịch, key art hoặc ảnh nhân vật trong khi giữ nguyên định hướng hình ảnh của nguồn. Hữu ích cho video ra mắt sản phẩm, quảng cáo mạng xã hội và nội dung landing page ưu tiên chuyển động.
Reference-to-video để giữ nhất quán chủ thể
Dùng tối đa 9 ảnh tham chiếu, 3 video tham chiếu và 3 âm thanh tham chiếu để dẫn dắt danh tính nhân vật, diện mạo chủ thể, thiết kế cảnh hoặc phong cách. Reference-to-video là quy trình H3 dành cho nhân vật lặp lại, hình ảnh thương hiệu và bộ sản phẩm sáng tạo liên kết.
Nội dung mạng xã hội và quảng cáo
Tạo ý tưởng sáng tạo dạng ngắn cho chiến dịch, ra mắt sản phẩm, bảng tin mạng xã hội và performance marketing. Đi từ văn bản, ảnh hoàn chỉnh hoặc tham chiếu hình ảnh đã duyệt mà không phải đổi nền tảng API.
Cảnh nhân vật và kể chuyện bằng hình ảnh
Tạo cảnh lấy nhân vật làm trung tâm, các nhịp truyện, đoạn tạo không khí và previsualization với endpoint phù hợp tư liệu nguồn: văn bản cho cảnh mới, ảnh để làm động, hoặc tham chiếu để liền mạch hơn.
Pipeline tạo video AI có thể mở rộng
Tích hợp MiniMax H3 vào ứng dụng và quy trình sáng tạo tự động qua WaveSpeedAI. Dùng một API key và vòng đời prediction nhất quán cho cả ba chế độ tạo của H3.
Mẹo
Mẹo viết prompt cho MiniMax H3
Lời khuyên thực tế để có kết quả tốt hơn từ MiniMax H3 — rút ra từ các mẫu hiệu quả trên các mô hình video trong các pipeline sản xuất.
- 01
Chọn endpoint theo tư liệu nguồn
Dùng text-to-video cho prompt gốc, image-to-video để tạo chuyển động cho một ảnh nguồn, và reference-to-video khi tham chiếu hình ảnh cần định hướng danh tính, diện mạo chủ thể, thiết kế cảnh hoặc phong cách.
- 02
Mô tả chuyển động như một chuỗi
Viết trạng thái ban đầu, hành động, phản ứng của môi trường, chuyển động máy quay và trạng thái kết thúc theo thứ tự thời gian. Một diễn tiến rõ ràng cho H3 kế hoạch chuyển động mạnh hơn một danh sách tính từ hình ảnh rời rạc.
- 03
Dùng image-to-video khi bố cục đã được duyệt
Nếu chủ thể, sản phẩm, khung hình hoặc định hướng nghệ thuật đã cố định trong một ảnh tĩnh, hãy bắt đầu bằng image-to-video thay vì dựng lại cảnh từ văn bản. Tập trung prompt vào chuyển động và hành vi máy quay.
- 04
Giao cho mỗi tham chiếu một nhiệm vụ rõ ràng
Với reference-to-video, hãy nêu tham chiếu cần kiểm soát gì: danh tính nhân vật, diện mạo sản phẩm, phong cách hình ảnh, môi trường hay bố cục. Tránh yêu cầu cùng một tham chiếu định nghĩa nhiều khía cạnh mâu thuẫn của cảnh quay.
- 05
Mỗi cảnh ngắn chỉ một hành động chính
Một hành động tập trung dễ render mạch lạc hơn cả chuỗi sự kiện không liên quan. Hãy tạo các cảnh quay riêng cho từng nhịp câu chuyện, rồi ghép lại khi dựng nếu ý tưởng cần phạm vi tường thuật rộng hơn.
Giá
Giá API MiniMax H3
Giá tính theo từng kết quả đầu ra. Khoản phí cuối cùng thay đổi theo các tham số bạn đặt trong playground của từng biến thể (độ phân giải, thời lượng, số kết quả, tham chiếu).
API
Gọi API MiniMax H3
Đăng ký API key tại wavespeed.ai/accesskey, rồi gửi prediction qua REST. Playground tạo sẵn mã mẫu có thể dán ngay cho mọi tổ hợp đầu vào.
POSThttps://api.wavespeed.ai/api/v3/wavespeed-ai/minimax-h3/text-to-video
# 1. Submit the prediction.
SUBMIT_RESPONSE=$(curl --silent --show-error --fail-with-body \
-X POST "https://api.wavespeed.ai/api/v3/wavespeed-ai/minimax-h3/text-to-video" \
-H "Content-Type: application/json" \
-H "Authorization: Bearer $WAVESPEED_API_KEY" \
-d '{
"prompt": "A cinematic shot of a city at sunset, soft golden light",
"aspect_ratio": "16:9",
"resolution": "480p",
"duration": 5
}')
TASK=$(printf '%s' "$SUBMIT_RESPONSE" | jq 'if has("data") then .data else . end')
PREDICTION_ID=$(printf '%s' "$TASK" | jq -r '.id')
if [ -z "$PREDICTION_ID" ] || [ "$PREDICTION_ID" = "null" ]; then
printf 'Submission response did not contain a prediction id
' >&2
exit 1
fi
RESULT_URL="https://api.wavespeed.ai/api/v3/predictions/$PREDICTION_ID/result"
# 2. Poll until the prediction finishes.
while true; do
RESPONSE=$(curl --silent --show-error --fail-with-body "$RESULT_URL" \
-H "Authorization: Bearer $WAVESPEED_API_KEY")
RESULT=$(printf '%s' "$RESPONSE" | jq 'if has("data") then .data else . end')
STATUS=$(printf '%s' "$RESULT" | jq -r '.status')
case "$STATUS" in
completed) printf '%s\n' "$RESULT" | jq '.outputs'; break ;;
failed|cancelled|timeout|deleted) printf '%s\n' "$RESULT" | jq . >&2; exit 1 ;;
*) sleep 2 ;;
esac
doneconst submitUrl = "https://api.wavespeed.ai/api/v3/wavespeed-ai/minimax-h3/text-to-video";
const apiKey = process.env.WAVESPEED_API_KEY;
if (!apiKey) throw new Error('Set WAVESPEED_API_KEY');
async function requestJson(url, options = {}) {
const response = await fetch(url, options);
if (!response.ok) throw new Error(await response.text());
return response.json();
}
// 1. Submit the prediction.
const body = await requestJson(submitUrl, {
method: "POST",
headers: {
"Authorization": `Bearer ${apiKey}`,
"Content-Type": "application/json",
},
body: JSON.stringify({
"prompt": "A cinematic shot of a city at sunset, soft golden light",
"aspect_ratio": "16:9",
"resolution": "480p",
"duration": 5
}),
});
const task = body.data ?? body;
const resultUrl = `https://api.wavespeed.ai/api/v3/predictions/${task.id}/result`;
// 2. Poll until the prediction finishes.
while (true) {
const resultBody = await requestJson(resultUrl, {
headers: { "Authorization": `Bearer ${apiKey}` },
});
const result = resultBody.data ?? resultBody;
if (result.status === "completed") {
console.log(result.outputs);
break;
}
if (["failed", "cancelled", "timeout", "deleted"].includes(result.status)) throw new Error(JSON.stringify(result));
await new Promise(resolve => setTimeout(resolve, 2000));
}import json
import os
import time
from urllib.request import Request, urlopen
api_key = os.environ["WAVESPEED_API_KEY"]
headers = {"Authorization": f"Bearer {api_key}", "Content-Type": "application/json"}
payload = {
"prompt": "A cinematic shot of a city at sunset, soft golden light",
"aspect_ratio": "16:9",
"resolution": "480p",
"duration": 5
}
def request_json(url, data=None):
request = Request(url, data=data, headers=headers, method="POST" if data else "GET")
with urlopen(request) as response:
return json.load(response)
# 1. Submit the prediction.
body = request_json("https://api.wavespeed.ai/api/v3/wavespeed-ai/minimax-h3/text-to-video", json.dumps(payload).encode())
task = body.get("data", body)
result_url = f"https://api.wavespeed.ai/api/v3/predictions/{task['id']}/result"
# 2. Poll until the prediction finishes.
while True:
result_body = request_json(result_url)
result = result_body.get("data", result_body)
status = result.get("status")
if status == "completed":
print(result.get("outputs", []))
break
if status in {"failed", "cancelled", "timeout", "deleted"}:
raise RuntimeError(result)
time.sleep(2)So sánh
MiniMax H3 so với các lựa chọn khác
Khi nào nên chọn MiniMax H3 thay vì các mô hình tương tự trên WaveSpeedAI.
MiniMax H3 so với Hailuo 2.3
Hailuo 2.3 cung cấp các endpoint text-to-video và image-to-video theo bậc gồm Standard, Pro, Fast và Fast Pro. MiniMax H3 dùng bộ ba endpoint gọn và bổ sung reference-to-video như một quy trình chính thức để dẫn dắt hình ảnh và giữ nhất quán chủ thể.
MiniMax H3 so với Seedance 2.0
Seedance 2.0 là dòng production rộng với text-to-video, image-to-video, video-edit, video-extend, âm thanh gốc và nhiều bậc hiệu năng. MiniMax H3 là lựa chọn đơn giản hơn khi quy trình xoay quanh việc tạo từ văn bản, ảnh nguồn hoặc tham chiếu hình ảnh.
MiniMax H3 so với Kling 3.0
Kling 3.0 nhấn mạnh các bậc chất lượng đầu ra và endpoint điều khiển chuyển động chuyên dụng. MiniMax H3 tổ chức API theo loại nguồn, gồm một route reference-to-video riêng cho việc tạo theo danh tính, chủ thể và phong cách.
Câu hỏi thường gặp
MiniMax H3 API — Câu hỏi thường gặp
Giá, giấy phép, tích hợp — những câu hỏi phổ biến về việc chạy MiniMax H3 trên WaveSpeedAI.
MiniMax H3 API là gì?
MiniMax H3 API là bộ ba mô hình tạo video AI trên WaveSpeedAI. Bộ này hỗ trợ text-to-video cho cảnh theo prompt, image-to-video để làm động ảnh tĩnh và reference-to-video để tạo theo tham chiếu hình ảnh.
Tôi nên dùng endpoint MiniMax H3 API nào?
Hãy bắt đầu với bản triển khai open-weights được khuyên dùng trên hạ tầng WaveSpeed: wavespeed-ai/minimax-h3/text-to-video khi bắt đầu từ prompt, wavespeed-ai/minimax-h3/image-to-video khi làm động ảnh nguồn, và wavespeed-ai/minimax-h3/reference-to-video khi tham chiếu hình ảnh cần dẫn dắt danh tính, diện mạo hoặc phong cách của chủ thể. Các endpoint minimax/h3 chính thức cũng có trong cùng dòng.
MiniMax H3 có hỗ trợ tạo image-to-video không?
Có. Endpoint wavespeed-ai/minimax-h3/image-to-video làm động ảnh khung đầu — có thể kèm khung cuối — thành video mạch lạc với âm thanh stereo gốc, phù hợp để làm động ảnh sản phẩm, ảnh chiến dịch, ảnh nhân vật, khung ý tưởng và các hình ảnh có sẵn khác.
MiniMax H3 reference-to-video là gì?
MiniMax H3 reference-to-video tạo video mới dựa trên tối đa 9 ảnh tham chiếu, 3 video tham chiếu và 3 âm thanh tham chiếu để dẫn dắt chủ thể, danh tính nhân vật, phong cách hoặc ngôn ngữ cảnh. Đây là endpoint H3 được ưu tiên cho nhân vật lặp lại và quy trình sáng tạo nhất quán với thương hiệu.
Làm thế nào để dùng MiniMax H3 API trên WaveSpeedAI?
Chọn endpoint H3 phù hợp với nguồn của bạn, xác thực bằng API key WaveSpeedAI, gửi đầu vào mô hình dưới dạng JSON, rồi dùng prediction ID hoặc URL kết quả được trả về để lấy video đã tạo. Thẻ endpoint và mẫu code trên trang này hiển thị schema request hiện tại.
API MiniMax H3 là gì?
MiniMax H3 là mô hình tạo video của MiniMax, được cung cấp dưới dạng REST API trên WaveSpeedAI. MiniMax H3 API để tạo video điện ảnh từ văn bản, ảnh và tham chiếu với âm thanh stereo gốc, chất lượng chuyển động mạnh, nhất quán chủ thể và mạch lạc cảnh. Chạy bản open-weights trên hạ tầng WaveSpeed hoặc các endpoint MiniMax chính thức qua một API WaveSpeedAI duy nhất. Bạn có thể gọi bằng lập trình hoặc thử từ playground ở liên kết phía trên.
Làm sao để gọi API MiniMax H3?
Đăng ký tài khoản WaveSpeedAI, sao chép API key từ /accesskey, rồi POST tới https://api.wavespeed.ai/api/v3/wavespeed-ai/minimax-h3/text-to-video với đầu vào dưới dạng JSON. Endpoint trả về prediction id. Hãy poll endpoint kết quả khoảng mỗi 2 giây, tăng khoảng cách với các tác vụ chạy lâu, và dừng ở bất kỳ trạng thái kết thúc nào. Ví dụ Python / Node.js / cURL hướng tới môi trường production ở phía trên.
API MiniMax H3 có giá bao nhiêu?
MiniMax H3 bắt đầu từ $0.02 mỗi lượt chạy. Chi phí chính xác thay đổi theo các tham số bạn đặt (độ phân giải, thời lượng, số kết quả, tham chiếu). Phần xem trước chi phí trực tiếp cạnh nút Tạo trong playground hiển thị giá chính xác cho đầu vào hiện tại của bạn.
Có những biến thể MiniMax H3 nào?
WaveSpeedAI cung cấp 20 endpoint MiniMax H3 đang hoạt động: wavespeed-ai/minimax-h3/controlnet-union, wavespeed-ai/minimax-h3-singularity/image-to-video-lora, wavespeed-ai/minimax-h3-singularity/image-to-video, wavespeed-ai/minimax-h3-singularity/reference-to-video-lora, wavespeed-ai/minimax-h3-singularity/reference-to-video, wavespeed-ai/minimax-h3/image-edit-lora, wavespeed-ai/minimax-h3/text-to-image-lora, wavespeed-ai/minimax-h3/image-edit, và nhiều hơn nữa. Mỗi biến thể có trang playground và giá riêng.
Tôi có thể dùng kết quả của MiniMax H3 cho mục đích thương mại không?
Quyền sử dụng thương mại tuân theo giấy phép mô hình của MiniMax. Hầu hết mô hình của MiniMax cho phép dùng kết quả đầu ra cho mục đích thương mại; hãy xem trang playground của từng mô hình để biết tóm tắt giấy phép cụ thể, và Điều khoản dịch vụ của WaveSpeedAI cho các điều kiện cấp nền tảng.
Tại sao dùng MiniMax H3 trên WaveSpeedAI thay vì gọi trực tiếp?
Một API key + một tài khoản thanh toán cho MiniMax H3 VÀ hơn 1.000 mô hình AI khác từ các nhà cung cấp khác. Không cần thiết lập SDK cho từng hãng, không có hạn mức tốc độ riêng, không phải viết lại mã tích hợp cho từng hãng. Giá thường ngang bằng hoặc thấp hơn API trực tiếp của MiniMax.
Nhà cung cấp
Giới thiệu về MiniMax
Đội ngũ đứng sau MiniMax H3 và dòng mô hình MiniMax rộng hơn trên WaveSpeedAI.
MiniMax là phòng thí nghiệm AI của Trung Quốc nổi tiếng với tạo video Hailuo và các mô hình giọng nói. Hailuo 2.3 có text-to-video và image-to-video hiểu vật lý với các bậc Standard, Pro và Fast — định vị quanh hiệu suất 2,5× và độ chính xác cao với chỉ dẫn phức tạp cho quy trình của nhà sáng tạo và marketing.
Bắt đầu xây dựng với MiniMax H3 trên WaveSpeedAI
Tặng credit dùng thử miễn phí khi đăng ký. Một API key cho hơn 1.000 mô hình AI từ MiniMax và mọi nhà cung cấp khác.