MiniMax H3 API
MiniMax H3 API สำหรับสร้างวิดีโอสไตล์ภาพยนตร์จากข้อความ ภาพ และภาพอ้างอิง พร้อมเสียงสเตอริโอในตัว คุณภาพการเคลื่อนไหวที่แข็งแกร่ง ความสม่ำเสมอของวัตถุ และความต่อเนื่องของฉาก รันรุ่น open-weights บนโครงสร้างพื้นฐานของ WaveSpeed หรือใช้ endpoint ทางการของ MiniMax ผ่าน WaveSpeedAI API เดียว
สร้างวิดีโอ AI สไตล์ภาพยนตร์จากพรอมต์ข้อความ ภาพนิ่ง หรือภาพอ้างอิง การติดตั้ง open-weights ที่แนะนำบนโครงสร้างพื้นฐานของ WaveSpeed (wavespeed-ai/minimax-h3) ให้วิดีโอ 480p/540p/768p/1080p ที่ต่อเนื่อง พร้อมเสียงสเตอริโอในตัว ความยาว 3-15 วินาที และราคาต่อวินาทีที่เข้าถึงได้ — และมี endpoint ทางการของ MiniMax ในตระกูลเดียวกัน
ภาพรวม
เกี่ยวกับ MiniMax H3 API
MiniMax H3 ทำอะไรได้ อยู่ตรงไหนในกลุ่มโมเดลของ MiniMax และทำไมทีมต่างๆ จึงเลือกใช้
MiniMax H3 คือโมเดลการสร้างวิดีโอจาก MiniMax ใช้งานได้ผ่าน REST API ของ WaveSpeedAI MiniMax H3 API สำหรับสร้างวิดีโอสไตล์ภาพยนตร์จากข้อความ ภาพ และภาพอ้างอิง พร้อมเสียงสเตอริโอในตัว คุณภาพการเคลื่อนไหวที่แข็งแกร่ง ความสม่ำเสมอของวัตถุ และความต่อเนื่องของฉาก รันรุ่น open-weights บนโครงสร้างพื้นฐานของ WaveSpeed หรือใช้ endpoint ทางการของ MiniMax ผ่าน WaveSpeedAI API เดียว
สร้างวิดีโอ AI สไตล์ภาพยนตร์จากพรอมต์ข้อความ ภาพนิ่ง หรือภาพอ้างอิง การติดตั้ง open-weights ที่แนะนำบนโครงสร้างพื้นฐานของ WaveSpeed (wavespeed-ai/minimax-h3) ให้วิดีโอ 480p/540p/768p/1080p ที่ต่อเนื่อง พร้อมเสียงสเตอริโอในตัว ความยาว 3-15 วินาที และราคาต่อวินาทีที่เข้าถึงได้ — และมี endpoint ทางการของ MiniMax ในตระกูลเดียวกัน
กลุ่ม MiniMax H3 บน WaveSpeedAI มี 20 REST endpoint ครอบคลุมเวิร์กโฟลว์ Motion-Control, Image-To-Video, Reference-To-Video, Image-To-Image, Text-To-Image, Video-To-Video, Video-Extend, Text-To-Video แต่ละแบบมีราคา พารามิเตอร์ที่ปรับได้ และตัวอย่างผลลัพธ์ของตัวเอง เลือกแบบที่ตรงกับประเภทอินพุตและข้อจำกัดของงานจริง หรือเรียกหลายแบบด้วย API key เดียวเพื่อประกอบเป็นไปป์ไลน์หลายขั้นตอน
รัน MiniMax H3 ผ่าน API key บัญชีเรียกเก็บเงิน และขีดจำกัดอัตราเดียวกับที่คุณใช้กับโมเดล AI กว่า 1,000 รายการอื่นบน WaveSpeedAI ไม่ต้องตั้งค่ากับผู้ให้บริการแยก ไม่ต้องใช้ SDK แยกตามผู้ให้บริการ และไม่ต้องจัดการขีดจำกัดอัตราแยกตามผู้ให้บริการ การเชื่อมต่อเดียวครอบคลุมทุกอย่างตั้งแต่ข้อความเป็นภาพและข้อความเป็นวิดีโอ ไปจนถึงการสังเคราะห์เสียง การสร้าง 3D การอัปสเกล และการแก้ไข
สเปก
ความสามารถและสถานะการเปิดตัวของ MiniMax H3 API
รายละเอียดเฉพาะของโมเดลที่นักพัฒนาค้นหาก่อนเลือก API: ความพร้อมใช้งาน ความยาวผลลัพธ์ที่คาดหวัง การรองรับภาพอ้างอิง และทางเลือกสำรองที่ใช้งานได้ในปัจจุบัน
ข้อความเป็นวิดีโอ
สร้างจากพรอมต์
ใช้ wavespeed-ai/minimax-h3/text-to-video เพื่อเปลี่ยนบรีฟครีเอทีฟที่เขียนไว้ให้เป็นวิดีโอ 480p/540p/768p/1080p ที่ต่อเนื่อง พร้อมเสียงสเตอริโอในตัว ความยาว 3-15 วินาที และอัตราส่วนภาพที่ยืดหยุ่น
ภาพเป็นวิดีโอ
ทำให้ภาพนิ่งเคลื่อนไหว
ใช้ wavespeed-ai/minimax-h3/image-to-video เพื่อทำให้ภาพเฟรมแรก — จะใส่เฟรมสุดท้ายด้วยก็ได้ — กลายเป็นวิดีโอที่ต่อเนื่องพร้อมเสียงสเตอริโอในตัว โดยนำวัตถุ การจัดเฟรม และอาร์ตไดเรกชันของภาพเข้าสู่การเคลื่อนไหว
Reference to video
กำกับตัวตนและสไตล์
ใช้ wavespeed-ai/minimax-h3/reference-to-video เพื่อกำกับการสร้างด้วยภาพอ้างอิงสูงสุด 9 ภาพ วิดีโออ้างอิง 3 รายการ และเสียงอ้างอิง 3 รายการ เมื่อวัตถุ ตัวตนของตัวละคร หรือสไตล์ต้องคงที่
การเชื่อมต่อ API
เวิร์กโฟลว์ WaveSpeedAI เดียว
ส่ง prediction ของ H3 ด้วย WaveSpeedAI API key ติดตามแต่ละคำขอผ่านวงจร prediction มาตรฐาน และรับวิดีโอที่สร้างจากผลลัพธ์ที่ตอบกลับ — เวิร์กโฟลว์เดียวกันทั้ง endpoint ที่ WaveSpeed โฮสต์และ endpoint ทางการของ MiniMax
Endpoint
API endpoint ของ MiniMax H3 ทั้งหมด
20 endpoint ของ MiniMax H3 พร้อมใช้งานบน WaveSpeedAI แล้ว — เลือกแบบที่ตรงกับเวิร์กโฟลว์ของคุณ
/filters:quality(82)/media/images/1790573118566544886_jzIR0oxG.webp)
Minimax H3 Controlnet Union
MiniMax H3 Open Weights ControlNet Union generates a new video that follows the motion and composition of a source video. Pose, depth, edges, lines, scribble or grayscale structure is extracted from the source automatically and guides the output, optionally with reference images for the subject or style, with native stereo audio generated in the same pass. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/1789888783447157992_hG8iqAJS.webp)
Minimax H3 Singularity Image To Video Lora
MiniMax H3 Singularity Open Weights Image to Video animates a first-frame image, optionally with last-frame guidance, into coherent 480P / 540P / 768P videos with native stereo audio, 5-15 second duration, flexible aspect ratios, and per-second billing. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/1789888766253281059_pb1bluDL.webp)
Minimax H3 Singularity Image To Video
MiniMax H3 Singularity Open Weights Image-to-Video animates a first-frame image, with optional last-frame guidance, into coherent 480P / 540P / 768P videos with native stereo audio, 5-15 second duration, flexible aspect ratios, and per-second billing. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/1789742735070040548_71aJS2bl.webp)
Minimax H3 Singularity Reference To Video Lora
MiniMax H3 Singularity Open Weights Reference to Video with custom LoRA support generates coherent 480P / 540P / 768P videos from prompts and multimodal references, guided by up to 9 reference images, 3 reference videos, and 3 reference audios, with native stereo audio and flexible reference-based video generation. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/1789742804979721138_b4T1aktE.webp)
Minimax H3 Singularity Reference To Video
MiniMax H3 Singularity Open Weights Reference to Video generates coherent 480P / 540P / 768P videos from prompts and multimodal references, guided by up to 9 reference images, 3 reference videos, and 3 reference audios, with native stereo audio and flexible reference-based video generation on WaveSpeedAI infrastructure. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/1788684887040987486_ILX6fU3e.webp)
Minimax H3 Image Edit Lora
MiniMax H3 Open Weights Image Edit re-renders a subject from up to 9 reference images into a new scene, outfit or style from a text instruction, preserving identity at 1K or 2K resolution. Loads up to 3 custom LoRA weights per request. Ready-to-use REST inference API, best performance, no cold starts, affordable pricing.
/filters:quality(82)/media/images/1788684121915845410_PlPY7gqA.webp)
Minimax H3 Text To Image Lora
MiniMax H3 Open Weights Text-to-Image generates photorealistic, cinema-grade stills from a text prompt at 1K or 2K resolution with fifteen aspect ratios. Loads up to 3 custom LoRA weights per request. Ready-to-use REST inference API, best performance, no cold starts, affordable pricing.
/filters:quality(82)/media/images/1788684872529921426_LGPY7gpy.webp)
Minimax H3 Image Edit
MiniMax H3 Open Weights Image Edit re-renders a subject from up to 9 reference images into a new scene, outfit or style from a text instruction, preserving identity at 1K or 2K resolution. Ready-to-use REST inference API, best performance, no cold starts, affordable pricing.
/filters:quality(82)/media/images/1788684072554630514_Lg4clvEO.webp)
Minimax H3 Text To Image
MiniMax H3 Open Weights Text-to-Image generates photorealistic, cinema-grade stills from a text prompt at 1K or 2K resolution with fifteen aspect ratios. Ready-to-use REST inference API, best performance, no cold starts, affordable pricing.
/filters:quality(82)/media/images/1788329903545692055_lEMW5fpy.webp)
Minimax H3 Video Edit
MiniMax H3 Open Weights Video-Edit edits an input video from a natural-language prompt. The input video drives subject identity, composition, and motion while the model rewrites lighting, style, weather, environment, or specific elements as instructed, with native stereo audio generated in the same pass. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/1788333421627652695_cNUHuh5T.webp)
Minimax H3 Video Extend
MiniMax H3 Open Weights Video-Extend appends a new cinematic continuation to an existing video. A fresh segment with native stereo audio is generated from the input video's last frame and a natural-language prompt, then the original and new segment are concatenated into a single output. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/1786952873106530692_cwjsCLT3.webp)
Minimax H3 Reference To Video Lora
MiniMax H3 Open Weights Reference to Video with custom LoRA support generates coherent 480P / 540P / 768P videos from prompts and multimodal references, guided by up to 9 reference images, 3 reference videos, and 3 reference audios, with native stereo audio and flexible reference-based video generation. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/1786952854061501695_t8heoxGQ.webp)
Minimax H3 Image To Video Lora
MiniMax H3 Open Weights Image to Video animates a first-frame image, optionally with last-frame guidance, into coherent 480P / 540P / 768P videos with native stereo audio, 5-15 second duration, flexible aspect ratios, and per-second billing. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/1788760867264586097_7BZ9isBK.webp)
Minimax H3 Text To Video Lora
MiniMax H3 Open Weights Text to Video with custom LoRA support generates coherent videos from text prompts, with 480P / 540P / 768P output, native stereo audio, 5-15 second duration, flexible aspect ratios, and per-second billing. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/1785814033019871155_vhrAKT2c.webp)
Minimax H3 Reference To Video
MiniMax H3 Open Weights Reference to Video generates coherent 480P / 540P / 768P videos from prompts and multimodal references, guided by up to 9 reference images, 3 reference videos, and 3 reference audios, with native stereo audio and flexible reference-based video generation on WaveSpeedAI infrastructure. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/1785815150161566057_2VF6wZoP.webp)
Minimax H3 Image To Video
MiniMax H3 Open Weights Image to Video animates a first-frame image, optionally with last-frame guidance, into coherent 480P / 540P / 768P videos with native stereo audio, 5-15 second duration, flexible aspect ratios, and per-second billing. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/1785814561183432100_tW5eoxGP.webp)
Minimax H3 Text To Video
MiniMax H3 Open Weights Text to Video generates coherent videos from text prompts, with 480P / 540P / 768P output, native stereo audio, 5-15 second duration, flexible aspect ratios, and per-second billing. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/1785415007635242852_UoxHQZ9H.webp)
H3 Reference To Video
MiniMax H3 Reference to Video generates coherent 2K videos from natural-language prompts and multimodal references, including images, videos, and audio, guiding subject consistency, motion, timing, visual style, and scene continuity. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/1785414895750297795_1JajtDMV.webp)
H3 Image To Video
MiniMax H3 Image to Video animates a first-frame image into a coherent 2K video, with natural-language motion instructions and optional last-frame control for consistent motion, scene continuity, and cinematic video generation. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/1785412796956877340_AktCLU4J.webp)
H3 Text To Video
MiniMax H3 Text to Video generates coherent 2K videos from text prompts, with flexible 5-15 second duration and adaptive or custom aspect ratios for cinematic scenes, creative videos, and production workflows. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
ตัวอย่าง
ดู MiniMax H3 ทำงานจริง
ผลลัพธ์จริงที่สร้างโดย MiniMax H3 API วางเมาส์บนวิดีโอเพื่อดูตัวอย่าง คลิกเพื่อเปิดตัวดูขนาดเต็ม
วิธีใช้
วิธีใช้ MiniMax H3 API
สี่ขั้นตอนจากการสมัครไปจนถึงผลลัพธ์ที่เสร็จสมบูรณ์ ตัวอย่าง Python, Node.js และ cURL แบบเต็มอยู่ในส่วน API ด้านล่าง
- 01
รับ API key
สมัครบัญชี WaveSpeedAI แล้วคัดลอก API key จากแดชบอร์ด บัญชีใหม่มาพร้อมเครดิตเริ่มต้นฟรี เพียงพอให้ลองใช้ playground ได้หลายสิบครั้งก่อนเริ่มเรียกเก็บเงิน
- 02
ส่ง prediction
POST อินพุตของคุณเป็น JSON ไปที่ https://api.wavespeed.ai/api/v3/wavespeed-ai/minimax-h3/text-to-video endpoint จะส่ง prediction id กลับมาทันที การสร้างทำงานแบบอะซิงโครนัส คุณจึงไม่ต้องเปิดการเชื่อมต่อค้างไว้ระหว่างประมวลผล
- 03
poll เพื่อรอให้เสร็จ
GET https://api.wavespeed.ai/api/v3/predictions/{request_id}/result เมื่อสถานะเป็น completed ให้นำผลลัพธ์ไปใช้ เมื่อเป็น failed, cancelled, timeout หรือ deleted ให้หยุดพร้อมแจ้งข้อผิดพลาด และ poll ต่อไปสำหรับสถานะอื่นทั้งหมด
- 04
อ่าน URL ของผลลัพธ์
เมื่อสถานะเป็น "completed" ให้อ่าน URL จาก data.outputs[0] URL ชี้ไปยังสื่อที่สร้างขึ้นของคุณบน CDN ของ WaveSpeedAI ซึ่งเป็นภาพ วิดีโอ เสียง หรือไฟล์ 3D ตามแบบของ MiniMax H3 ที่คุณเรียกใช้
กรณีการใช้งาน
สิ่งที่คุณสร้างได้ด้วย MiniMax H3
เวิร์กโฟลว์ทั่วไปที่นักพัฒนาและครีเอเตอร์ใช้ MiniMax H3 API
ข้อความเป็นวิดีโอสำหรับไอเดียต้นฉบับ
เปลี่ยนพรอมต์ให้เป็นฉากสไตล์ภาพยนตร์ต้นฉบับด้วย wavespeed-ai/minimax-h3/text-to-video — เอาต์พุต 480p/540p/768p/1080p ที่ต่อเนื่อง พร้อมเสียงสเตอริโอในตัวและอัตราส่วนภาพที่ยืดหยุ่น ใช้สำหรับการสร้างภาพคอนเซปต์ การเล่าเรื่อง ไอเดียแคมเปญ และช็อตที่ไม่จำเป็นต้องเริ่มจากงานภาพที่มีอยู่
ภาพเป็นวิดีโอสำหรับโชว์สินค้า
ทำให้ภาพถ่ายสินค้า ภาพนิ่งแคมเปญ คีย์อาร์ต หรือภาพตัวละครเคลื่อนไหว โดยรักษาทิศทางภาพของต้นฉบับ เหมาะกับการเปิดตัวสินค้า โฆษณาโซเชียล และคอนเทนต์แลนดิ้งที่เน้นการเคลื่อนไหว
Reference-to-video เพื่อความสม่ำเสมอของวัตถุ
ใช้ภาพอ้างอิงได้สูงสุด 9 ภาพ วิดีโออ้างอิง 3 รายการ และเสียงอ้างอิง 3 รายการ เพื่อกำหนดตัวตนของตัวละคร รูปลักษณ์ของวัตถุ การออกแบบฉาก หรือสไตล์ Reference-to-video คือเวิร์กโฟลว์ของ H3 สำหรับตัวละครที่กลับมาซ้ำ ภาพของแบรนด์ และชุดงานสร้างสรรค์ที่เชื่อมโยงกัน
โซเชียลมีเดียและครีเอทีฟโฆษณา
สร้างคอนเซปต์ครีเอทีฟแบบสั้นสำหรับแคมเปญ การเปิดตัว ฟีดโซเชียล และการตลาดแบบเน้นผลลัพธ์ เริ่มจากข้อความ ภาพนิ่งที่เสร็จแล้ว หรือภาพอ้างอิงที่อนุมัติแล้ว โดยไม่ต้องเปลี่ยนแพลตฟอร์ม API
ฉากตัวละครและการเล่าเรื่องด้วยภาพ
สร้างฉากที่มีตัวละครเป็นศูนย์กลาง จังหวะเรื่องราว ชิ้นงานบรรยากาศ และพรีวิชวลไลซ์เซชัน ด้วย endpoint ที่เลือกให้ตรงกับวัสดุต้นฉบับ: ข้อความสำหรับฉากใหม่ ภาพสำหรับการทำให้เคลื่อนไหว หรือภาพอ้างอิงเพื่อความต่อเนื่องที่แข็งแกร่งขึ้น
ไปป์ไลน์สร้างวิดีโอ AI ที่ขยายขนาดได้
ผสานรวม MiniMax H3 เข้ากับแอปพลิเคชันและเวิร์กโฟลว์ครีเอทีฟอัตโนมัติผ่าน WaveSpeedAI ใช้ API key เดียวและวงจร prediction ที่สอดคล้องกันในโหมดสร้างทั้งสามของ H3
เคล็ดลับ
เคล็ดลับการเขียนพรอมต์สำหรับ MiniMax H3
คำแนะนำเชิงปฏิบัติเพื่อให้ได้ผลลัพธ์ที่ดีขึ้นจาก MiniMax H3 — มาจากรูปแบบที่ได้ผลกับโมเดลวิดีโอในไปป์ไลน์การผลิตจริง
- 01
เลือก endpoint ตามวัสดุต้นฉบับ
ใช้ text-to-video สำหรับพรอมต์ต้นฉบับ image-to-video เพื่อทำให้ภาพต้นฉบับหนึ่งภาพเคลื่อนไหว และ reference-to-video เมื่อต้องการให้ภาพอ้างอิงกำกับตัวตน รูปลักษณ์ของวัตถุ การออกแบบฉาก หรือสไตล์
- 02
อธิบายการเคลื่อนไหวเป็นลำดับ
เขียนสถานะเริ่มต้น แอ็กชัน ปฏิกิริยาของสภาพแวดล้อม การเคลื่อนกล้อง และสถานะสุดท้ายตามลำดับเวลา ความก้าวหน้าที่ชัดเจนให้แผนการเคลื่อนไหวที่แข็งแกร่งกว่าการลิสต์คำคุณศัพท์ภาพที่ไม่เกี่ยวข้องกัน
- 03
ใช้ image-to-video เมื่อองค์ประกอบได้รับอนุมัติแล้ว
หากวัตถุ สินค้า การจัดเฟรม หรืออาร์ตไดเรกชันล็อกไว้ในภาพนิ่งแล้ว ให้เริ่มด้วย image-to-video แทนการสร้างฉากใหม่จากข้อความ โฟกัสพรอมต์ที่การเคลื่อนไหวและพฤติกรรมของกล้อง
- 04
ให้ภาพอ้างอิงแต่ละภาพมีหน้าที่ชัดเจนหนึ่งอย่าง
สำหรับ reference-to-video ระบุว่าภาพอ้างอิงควรควบคุมอะไร: ตัวตนของตัวละคร รูปลักษณ์สินค้า สไตล์ภาพ สภาพแวดล้อม หรือองค์ประกอบ หลีกเลี่ยงการให้ภาพอ้างอิงเดียวกำหนดหลายแง่มุมของช็อตที่ขัดแย้งกัน
- 05
ใช้หนึ่งแอ็กชันหลักต่อหนึ่งฉากสั้น
แอ็กชันที่โฟกัสเรนเดอร์ให้ต่อเนื่องได้ง่ายกว่าลำดับเหตุการณ์ที่ไม่เกี่ยวข้องกัน สร้างช็อตแยกสำหรับแต่ละจังหวะเรื่องราว แล้วประกอบกันในการตัดต่อเมื่อคอนเซปต์ต้องการขอบเขตการเล่าเรื่องที่กว้างขึ้น
ราคา
ราคา MiniMax H3 API
คิดราคาตามผลลัพธ์ ค่าใช้จ่ายสุดท้ายขึ้นอยู่กับพารามิเตอร์ที่คุณตั้งใน playground ของแต่ละแบบ (ความละเอียด ความยาว จำนวนผลลัพธ์ ภาพอ้างอิง)
API
เรียกใช้ MiniMax H3 API
สมัครรับ API key ที่ wavespeed.ai/accesskey แล้วส่ง prediction ผ่าน REST playground สร้างตัวอย่างพร้อมวางได้สำหรับอินพุตทุกแบบ
POSThttps://api.wavespeed.ai/api/v3/wavespeed-ai/minimax-h3/text-to-video
# 1. Submit the prediction.
SUBMIT_RESPONSE=$(curl --silent --show-error --fail-with-body \
-X POST "https://api.wavespeed.ai/api/v3/wavespeed-ai/minimax-h3/text-to-video" \
-H "Content-Type: application/json" \
-H "Authorization: Bearer $WAVESPEED_API_KEY" \
-d '{
"prompt": "A cinematic shot of a city at sunset, soft golden light",
"aspect_ratio": "16:9",
"resolution": "480p",
"duration": 5
}')
TASK=$(printf '%s' "$SUBMIT_RESPONSE" | jq 'if has("data") then .data else . end')
PREDICTION_ID=$(printf '%s' "$TASK" | jq -r '.id')
if [ -z "$PREDICTION_ID" ] || [ "$PREDICTION_ID" = "null" ]; then
printf 'Submission response did not contain a prediction id
' >&2
exit 1
fi
RESULT_URL="https://api.wavespeed.ai/api/v3/predictions/$PREDICTION_ID/result"
# 2. Poll until the prediction finishes.
while true; do
RESPONSE=$(curl --silent --show-error --fail-with-body "$RESULT_URL" \
-H "Authorization: Bearer $WAVESPEED_API_KEY")
RESULT=$(printf '%s' "$RESPONSE" | jq 'if has("data") then .data else . end')
STATUS=$(printf '%s' "$RESULT" | jq -r '.status')
case "$STATUS" in
completed) printf '%s\n' "$RESULT" | jq '.outputs'; break ;;
failed|cancelled|timeout|deleted) printf '%s\n' "$RESULT" | jq . >&2; exit 1 ;;
*) sleep 2 ;;
esac
doneconst submitUrl = "https://api.wavespeed.ai/api/v3/wavespeed-ai/minimax-h3/text-to-video";
const apiKey = process.env.WAVESPEED_API_KEY;
if (!apiKey) throw new Error('Set WAVESPEED_API_KEY');
async function requestJson(url, options = {}) {
const response = await fetch(url, options);
if (!response.ok) throw new Error(await response.text());
return response.json();
}
// 1. Submit the prediction.
const body = await requestJson(submitUrl, {
method: "POST",
headers: {
"Authorization": `Bearer ${apiKey}`,
"Content-Type": "application/json",
},
body: JSON.stringify({
"prompt": "A cinematic shot of a city at sunset, soft golden light",
"aspect_ratio": "16:9",
"resolution": "480p",
"duration": 5
}),
});
const task = body.data ?? body;
const resultUrl = `https://api.wavespeed.ai/api/v3/predictions/${task.id}/result`;
// 2. Poll until the prediction finishes.
while (true) {
const resultBody = await requestJson(resultUrl, {
headers: { "Authorization": `Bearer ${apiKey}` },
});
const result = resultBody.data ?? resultBody;
if (result.status === "completed") {
console.log(result.outputs);
break;
}
if (["failed", "cancelled", "timeout", "deleted"].includes(result.status)) throw new Error(JSON.stringify(result));
await new Promise(resolve => setTimeout(resolve, 2000));
}import json
import os
import time
from urllib.request import Request, urlopen
api_key = os.environ["WAVESPEED_API_KEY"]
headers = {"Authorization": f"Bearer {api_key}", "Content-Type": "application/json"}
payload = {
"prompt": "A cinematic shot of a city at sunset, soft golden light",
"aspect_ratio": "16:9",
"resolution": "480p",
"duration": 5
}
def request_json(url, data=None):
request = Request(url, data=data, headers=headers, method="POST" if data else "GET")
with urlopen(request) as response:
return json.load(response)
# 1. Submit the prediction.
body = request_json("https://api.wavespeed.ai/api/v3/wavespeed-ai/minimax-h3/text-to-video", json.dumps(payload).encode())
task = body.get("data", body)
result_url = f"https://api.wavespeed.ai/api/v3/predictions/{task['id']}/result"
# 2. Poll until the prediction finishes.
while True:
result_body = request_json(result_url)
result = result_body.get("data", result_body)
status = result.get("status")
if status == "completed":
print(result.get("outputs", []))
break
if status in {"failed", "cancelled", "timeout", "deleted"}:
raise RuntimeError(result)
time.sleep(2)เปรียบเทียบ
MiniMax H3 เทียบกับทางเลือกอื่น
เมื่อไรควรเลือก MiniMax H3 แทนโมเดลที่คล้ายกันบน WaveSpeedAI
MiniMax H3 เทียบกับ Hailuo 2.3
Hailuo 2.3 มี endpoint ข้อความเป็นวิดีโอและภาพเป็นวิดีโอแบบแบ่งระดับ ได้แก่ Standard, Pro, Fast และ Fast Pro ส่วน MiniMax H3 ใช้ชุด endpoint สามรายการที่โฟกัส และเพิ่ม reference-to-video เป็นเวิร์กโฟลว์หลักสำหรับการกำกับด้วยภาพและความสม่ำเสมอของวัตถุ
MiniMax H3 เทียบกับ Seedance 2.0
Seedance 2.0 เป็นตระกูลโปรดักชันที่ครอบคลุม มีข้อความเป็นวิดีโอ ภาพเป็นวิดีโอ video-edit video-extend เสียงในตัว และหลายระดับประสิทธิภาพ MiniMax H3 เป็นทางเลือกที่เรียบง่ายกว่าเมื่อเวิร์กโฟลว์เน้นการสร้างจากข้อความ ภาพต้นฉบับ หรือภาพอ้างอิง
MiniMax H3 เทียบกับ Kling 3.0
Kling 3.0 เน้นคุณภาพเอาต์พุตแบบแบ่งระดับและ endpoint ควบคุมการเคลื่อนไหวโดยเฉพาะ MiniMax H3 จัดระเบียบ API ตามประเภทต้นฉบับ รวมถึงเส้นทาง reference-to-video โดยเฉพาะสำหรับการสร้างที่อิงตัวตน วัตถุ และสไตล์
คำถามที่พบบ่อย
MiniMax H3 API — คำถามที่พบบ่อย
ราคา สัญญาอนุญาต การเชื่อมต่อ — คำถามที่พบบ่อยเกี่ยวกับการรัน MiniMax H3 บน WaveSpeedAI
MiniMax H3 API คืออะไร
MiniMax H3 API คือชุดโมเดลสร้างวิดีโอ AI สามโมเดลบน WaveSpeedAI รองรับข้อความเป็นวิดีโอสำหรับฉากที่ขับเคลื่อนด้วยพรอมต์ ภาพเป็นวิดีโอสำหรับทำให้ภาพนิ่งเคลื่อนไหว และ reference-to-video สำหรับการสร้างที่กำกับด้วยภาพอ้างอิง
ควรเลือกใช้ endpoint ไหนของ MiniMax H3 API
เริ่มจากการติดตั้ง open-weights ที่แนะนำบนโครงสร้างพื้นฐานของ WaveSpeed: wavespeed-ai/minimax-h3/text-to-video เมื่อเริ่มจากพรอมต์ wavespeed-ai/minimax-h3/image-to-video เมื่อทำให้ภาพต้นฉบับเคลื่อนไหว และ wavespeed-ai/minimax-h3/reference-to-video เมื่อต้องการให้ภาพอ้างอิงกำกับตัวตน รูปลักษณ์ หรือสไตล์ของวัตถุ endpoint ทางการ minimax/h3 มีให้ในตระกูลเดียวกัน
MiniMax H3 รองรับการสร้างแบบภาพเป็นวิดีโอหรือไม่
รองรับ endpoint wavespeed-ai/minimax-h3/image-to-video ทำให้ภาพเฟรมแรก — จะใส่เฟรมสุดท้ายด้วยก็ได้ — กลายเป็นวิดีโอที่ต่อเนื่องพร้อมเสียงสเตอริโอในตัว เหมาะกับการทำให้ภาพถ่ายสินค้า งานแคมเปญ ภาพตัวละคร เฟรมคอนเซปต์ และภาพอื่นที่มีอยู่แล้วเคลื่อนไหว
MiniMax H3 reference-to-video คืออะไร
MiniMax H3 reference-to-video สร้างวิดีโอใหม่โดยมีภาพอ้างอิงสูงสุด 9 ภาพ วิดีโออ้างอิง 3 รายการ และเสียงอ้างอิง 3 รายการเป็นตัวกำกับวัตถุ ตัวตนของตัวละคร สไตล์ หรือภาษาของฉาก เป็น endpoint ของ H3 ที่แนะนำสำหรับตัวละครที่กลับมาซ้ำและเวิร์กโฟลว์ครีเอทีฟที่ต้องสอดคล้องกับแบรนด์
ใช้ MiniMax H3 API บน WaveSpeedAI อย่างไร
เลือก endpoint ของ H3 ที่ตรงกับต้นฉบับของคุณ ยืนยันตัวตนด้วย WaveSpeedAI API key ส่งอินพุตของโมเดลเป็น JSON แล้วใช้ prediction ID หรือ URL ผลลัพธ์ที่ได้กลับมาเพื่อดึงวิดีโอที่สร้าง การ์ด endpoint และโค้ดตัวอย่างในหน้านี้แสดงสคีมาคำขอล่าสุด
เรียกใช้ MiniMax H3 API อย่างไร
สมัครบัญชี WaveSpeedAI คัดลอก API key จาก /accesskey แล้ว POST ไปที่ https://api.wavespeed.ai/api/v3/wavespeed-ai/minimax-h3/text-to-video พร้อมอินพุตของคุณเป็น JSON endpoint จะส่ง prediction id กลับมา poll endpoint ผลลัพธ์โดยเริ่มประมาณทุก 2 วินาที เพิ่มช่วงเวลาสำหรับงานที่ใช้เวลานาน และหยุดเมื่อได้สถานะสิ้นสุดใดๆ ตัวอย่าง Python / Node.js / cURL ที่เหมาะกับงานจริงอยู่ด้านบน
MiniMax H3 API มีค่าใช้จ่ายเท่าไร
MiniMax H3 เริ่มต้นที่ $0.02 ต่อครั้ง ค่าใช้จ่ายจริงขึ้นอยู่กับพารามิเตอร์ที่คุณตั้ง (ความละเอียด ความยาว จำนวนผลลัพธ์ ภาพอ้างอิง) ตัวอย่างค่าใช้จ่ายแบบเรียลไทม์ข้างปุ่ม Generate ใน playground แสดงราคาที่แน่นอนสำหรับอินพุตปัจจุบันของคุณ
MiniMax H3 มีแบบไหนให้ใช้บ้าง
WaveSpeedAI มี endpoint MiniMax H3 ที่ใช้งานได้ 20 รายการ: wavespeed-ai/minimax-h3/controlnet-union, wavespeed-ai/minimax-h3-singularity/image-to-video-lora, wavespeed-ai/minimax-h3-singularity/image-to-video, wavespeed-ai/minimax-h3-singularity/reference-to-video-lora, wavespeed-ai/minimax-h3-singularity/reference-to-video, wavespeed-ai/minimax-h3/image-edit-lora, wavespeed-ai/minimax-h3/text-to-image-lora, wavespeed-ai/minimax-h3/image-edit, และอื่นๆ แต่ละแบบมีหน้า playground และราคาของตัวเอง
ใช้ผลลัพธ์ของ MiniMax H3 เชิงพาณิชย์ได้ไหม
สิทธิ์ในการใช้เชิงพาณิชย์เป็นไปตามสัญญาอนุญาตของโมเดล MiniMax โมเดลของ MiniMax ส่วนใหญ่อนุญาตให้ใช้ผลลัพธ์เชิงพาณิชย์ ดูสรุปสัญญาอนุญาตเฉพาะในหน้า playground ของแต่ละโมเดล และข้อกำหนดการใช้บริการของ WaveSpeedAI สำหรับเงื่อนไขระดับแพลตฟอร์ม
ทำไมต้องใช้ MiniMax H3 บน WaveSpeedAI แทนการเชื่อมต่อตรง
API key เดียวและบัญชีเรียกเก็บเงินเดียวสำหรับ MiniMax H3 และโมเดล AI อื่นกว่า 1,000 รายการจากผู้ให้บริการอื่น ไม่ต้องตั้งค่า SDK แยกตามผู้ให้บริการ ไม่ต้องจัดการขีดจำกัดอัตราแยกกัน และไม่ต้องเขียนโค้ดเชื่อมต่อใหม่ทุกครั้งที่เปลี่ยนผู้ให้บริการ โดยทั่วไปราคาเท่ากับหรือต่ำกว่า API ตรงของ MiniMax
ผู้พัฒนา
เกี่ยวกับ MiniMax
ทีมเบื้องหลัง MiniMax H3 และกลุ่มโมเดลอื่นๆ ของ MiniMax บน WaveSpeedAI
MiniMax เป็นห้องปฏิบัติการ AI ของจีนที่ขึ้นชื่อด้านการสร้างวิดีโอ Hailuo และโมเดลเสียงพูด Hailuo 2.3 มีข้อความเป็นวิดีโอและภาพเป็นวิดีโอที่ตระหนักถึงฟิสิกส์ ในระดับ Standard, Pro และ Fast — วางตำแหน่งรอบประสิทธิภาพ 2.5× และความแม่นยำคำสั่งซับซ้อนที่แข็งแกร่งสำหรับเวิร์กโฟลว์ของครีเอเตอร์และการตลาด
เริ่มสร้างด้วย MiniMax H3 บน WaveSpeedAI
รับเครดิตเริ่มต้นฟรีเมื่อสมัคร API key เดียวใช้ได้กับโมเดล AI กว่า 1,000 รายการจาก MiniMax และผู้ให้บริการอื่นทุกราย