MiniMax H3 API
सिनेमैटिक text-to-video, image-to-video और reference-to-video जनरेशन के लिए MiniMax H3 API — नेटिव स्टीरियो ऑडियो, दमदार मोशन क्वालिटी, विषय की निरंतरता और सीन कोहेरेंस के साथ। एक ही WaveSpeedAI API से WaveSpeed इन्फ्रास्ट्रक्चर पर ओपन-वेट्स रिलीज़ या आधिकारिक MiniMax एंडपॉइंट चलाएं।
टेक्स्ट प्रॉम्प्ट, स्टिल इमेज या विज़ुअल रेफ़रेंस से सिनेमैटिक AI वीडियो बनाएं। WaveSpeed इन्फ्रास्ट्रक्चर पर अनुशंसित ओपन-वेट्स डिप्लॉयमेंट (wavespeed-ai/minimax-h3) नेटिव स्टीरियो ऑडियो के साथ सुसंगत 480p/540p/768p/1080p वीडियो, 3-15 सेकंड की अवधि और किफ़ायती प्रति-सेकंड प्राइसिंग देता है — आधिकारिक MiniMax एंडपॉइंट उसी फ़ैमिली में उपलब्ध हैं।
अवलोकन
MiniMax H3 API के बारे में
MiniMax H3 क्या करता है, MiniMax की मॉडल श्रृंखला में इसकी जगह क्या है, और टीमें इसे क्यों चुनती हैं।
MiniMax H3, MiniMax का वीडियो जनरेशन मॉडल है, जो WaveSpeedAI REST API के ज़रिए उपलब्ध है। सिनेमैटिक text-to-video, image-to-video और reference-to-video जनरेशन के लिए MiniMax H3 API — नेटिव स्टीरियो ऑडियो, दमदार मोशन क्वालिटी, विषय की निरंतरता और सीन कोहेरेंस के साथ। एक ही WaveSpeedAI API से WaveSpeed इन्फ्रास्ट्रक्चर पर ओपन-वेट्स रिलीज़ या आधिकारिक MiniMax एंडपॉइंट चलाएं।
टेक्स्ट प्रॉम्प्ट, स्टिल इमेज या विज़ुअल रेफ़रेंस से सिनेमैटिक AI वीडियो बनाएं। WaveSpeed इन्फ्रास्ट्रक्चर पर अनुशंसित ओपन-वेट्स डिप्लॉयमेंट (wavespeed-ai/minimax-h3) नेटिव स्टीरियो ऑडियो के साथ सुसंगत 480p/540p/768p/1080p वीडियो, 3-15 सेकंड की अवधि और किफ़ायती प्रति-सेकंड प्राइसिंग देता है — आधिकारिक MiniMax एंडपॉइंट उसी फ़ैमिली में उपलब्ध हैं।
WaveSpeedAI पर MiniMax H3 परिवार में Motion-Control, Image-To-Video, Reference-To-Video, Image-To-Image, Text-To-Image, Video-To-Video, Video-Extend, Text-To-Video वर्कफ़्लो को कवर करने वाले 20 REST एंडपॉइंट हैं। हर वेरिएंट की अपनी प्राइसिंग, पैरामीटर और उदाहरण आउटपुट हैं — अपने इनपुट मोडैलिटी और प्रोडक्शन ज़रूरतों के अनुसार चुनें, या एक ही API कुंजी से कई वेरिएंट कॉल करके मल्टी-स्टेप पाइपलाइन बनाएं।
MiniMax H3 को उसी API कुंजी, बिलिंग अकाउंट और रेट-लिमिट के साथ चलाएं जिसका उपयोग आप WaveSpeedAI के अन्य 1,000+ AI मॉडल के लिए करते हैं। अलग वेंडर सेटअप नहीं, प्रति-प्रदाता SDK नहीं, प्रति-वेंडर रेट-लिमिट नहीं — एक ही इंटीग्रेशन टेक्स्ट-टू-इमेज और टेक्स्ट-टू-वीडियो से लेकर ऑडियो सिंथेसिस, 3D जनरेशन, अपस्केलिंग और एडिटिंग तक सब कुछ कवर करता है।
स्पेसिफ़िकेशन
MiniMax H3 API की क्षमताएँ और रिलीज़ स्थिति
मॉडल-विशिष्ट जानकारी जो डेवलपर API चुनने से पहले खोजते हैं: उपलब्धता, अपेक्षित आउटपुट अवधि, रेफरेंस सपोर्ट और वर्तमान लाइव विकल्प।
Text to video
प्रॉम्प्ट-आधारित जनरेशन
लिखित क्रिएटिव ब्रीफ़ को नेटिव स्टीरियो ऑडियो, 3-15 सेकंड की अवधि और लचीले आस्पेक्ट रेशियो के साथ सुसंगत 480p/540p/768p/1080p वीडियो में बदलने के लिए wavespeed-ai/minimax-h3/text-to-video इस्तेमाल करें।
Image to video
स्टिल इमेज को एनिमेट करें
पहले फ़्रेम की इमेज — वैकल्पिक रूप से अंतिम फ़्रेम के साथ — को नेटिव स्टीरियो ऑडियो वाले सुसंगत वीडियो में एनिमेट करने के लिए wavespeed-ai/minimax-h3/image-to-video इस्तेमाल करें; इसमें विषय, फ़्रेमिंग और आर्ट डायरेक्शन मोशन में आगे बढ़ते हैं।
Reference to video
पहचान और स्टाइल निर्देशित करें
जब विषय, कैरेक्टर की पहचान या स्टाइल एक जैसी रखनी हो, तब 9 रेफ़रेंस इमेज, 3 रेफ़रेंस वीडियो और 3 रेफ़रेंस ऑडियो तक से जनरेशन निर्देशित करने के लिए wavespeed-ai/minimax-h3/reference-to-video इस्तेमाल करें।
API इंटीग्रेशन
एक WaveSpeedAI वर्कफ़्लो
अपनी WaveSpeedAI API key से H3 प्रेडिक्शन सबमिट करें, स्टैंडर्ड प्रेडिक्शन लाइफ़साइकिल से हर रिक्वेस्ट को ट्रैक करें और रिज़ल्ट रिस्पॉन्स से जनरेट किए गए वीडियो प्राप्त करें — WaveSpeed-होस्टेड और आधिकारिक MiniMax एंडपॉइंट के लिए एक ही वर्कफ़्लो।
एंडपॉइंट
सभी MiniMax H3 API एंडपॉइंट
WaveSpeedAI पर अभी 20 MiniMax H3 एंडपॉइंट उपलब्ध हैं — अपने वर्कफ़्लो के अनुकूल वेरिएंट चुनें।
/filters:quality(82)/media/images/1790573118566544886_jzIR0oxG.webp)
Minimax H3 Controlnet Union
MiniMax H3 Open Weights ControlNet Union generates a new video that follows the motion and composition of a source video. Pose, depth, edges, lines, scribble or grayscale structure is extracted from the source automatically and guides the output, optionally with reference images for the subject or style, with native stereo audio generated in the same pass. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/1789888783447157992_hG8iqAJS.webp)
Minimax H3 Singularity Image To Video Lora
MiniMax H3 Singularity Open Weights Image to Video animates a first-frame image, optionally with last-frame guidance, into coherent 480P / 540P / 768P videos with native stereo audio, 5-15 second duration, flexible aspect ratios, and per-second billing. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/1789888766253281059_pb1bluDL.webp)
Minimax H3 Singularity Image To Video
MiniMax H3 Singularity Open Weights Image-to-Video animates a first-frame image, with optional last-frame guidance, into coherent 480P / 540P / 768P videos with native stereo audio, 5-15 second duration, flexible aspect ratios, and per-second billing. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/1789742735070040548_71aJS2bl.webp)
Minimax H3 Singularity Reference To Video Lora
MiniMax H3 Singularity Open Weights Reference to Video with custom LoRA support generates coherent 480P / 540P / 768P videos from prompts and multimodal references, guided by up to 9 reference images, 3 reference videos, and 3 reference audios, with native stereo audio and flexible reference-based video generation. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/1789742804979721138_b4T1aktE.webp)
Minimax H3 Singularity Reference To Video
MiniMax H3 Singularity Open Weights Reference to Video generates coherent 480P / 540P / 768P videos from prompts and multimodal references, guided by up to 9 reference images, 3 reference videos, and 3 reference audios, with native stereo audio and flexible reference-based video generation on WaveSpeedAI infrastructure. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/1788684887040987486_ILX6fU3e.webp)
Minimax H3 Image Edit Lora
MiniMax H3 Open Weights Image Edit re-renders a subject from up to 9 reference images into a new scene, outfit or style from a text instruction, preserving identity at 1K or 2K resolution. Loads up to 3 custom LoRA weights per request. Ready-to-use REST inference API, best performance, no cold starts, affordable pricing.
/filters:quality(82)/media/images/1788684121915845410_PlPY7gqA.webp)
Minimax H3 Text To Image Lora
MiniMax H3 Open Weights Text-to-Image generates photorealistic, cinema-grade stills from a text prompt at 1K or 2K resolution with fifteen aspect ratios. Loads up to 3 custom LoRA weights per request. Ready-to-use REST inference API, best performance, no cold starts, affordable pricing.
/filters:quality(82)/media/images/1788684872529921426_LGPY7gpy.webp)
Minimax H3 Image Edit
MiniMax H3 Open Weights Image Edit re-renders a subject from up to 9 reference images into a new scene, outfit or style from a text instruction, preserving identity at 1K or 2K resolution. Ready-to-use REST inference API, best performance, no cold starts, affordable pricing.
/filters:quality(82)/media/images/1788684072554630514_Lg4clvEO.webp)
Minimax H3 Text To Image
MiniMax H3 Open Weights Text-to-Image generates photorealistic, cinema-grade stills from a text prompt at 1K or 2K resolution with fifteen aspect ratios. Ready-to-use REST inference API, best performance, no cold starts, affordable pricing.
/filters:quality(82)/media/images/1788329903545692055_lEMW5fpy.webp)
Minimax H3 Video Edit
MiniMax H3 Open Weights Video-Edit edits an input video from a natural-language prompt. The input video drives subject identity, composition, and motion while the model rewrites lighting, style, weather, environment, or specific elements as instructed, with native stereo audio generated in the same pass. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/1788333421627652695_cNUHuh5T.webp)
Minimax H3 Video Extend
MiniMax H3 Open Weights Video-Extend appends a new cinematic continuation to an existing video. A fresh segment with native stereo audio is generated from the input video's last frame and a natural-language prompt, then the original and new segment are concatenated into a single output. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/1786952873106530692_cwjsCLT3.webp)
Minimax H3 Reference To Video Lora
MiniMax H3 Open Weights Reference to Video with custom LoRA support generates coherent 480P / 540P / 768P videos from prompts and multimodal references, guided by up to 9 reference images, 3 reference videos, and 3 reference audios, with native stereo audio and flexible reference-based video generation. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/1786952854061501695_t8heoxGQ.webp)
Minimax H3 Image To Video Lora
MiniMax H3 Open Weights Image to Video animates a first-frame image, optionally with last-frame guidance, into coherent 480P / 540P / 768P videos with native stereo audio, 5-15 second duration, flexible aspect ratios, and per-second billing. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/1788760867264586097_7BZ9isBK.webp)
Minimax H3 Text To Video Lora
MiniMax H3 Open Weights Text to Video with custom LoRA support generates coherent videos from text prompts, with 480P / 540P / 768P output, native stereo audio, 5-15 second duration, flexible aspect ratios, and per-second billing. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/1785814033019871155_vhrAKT2c.webp)
Minimax H3 Reference To Video
MiniMax H3 Open Weights Reference to Video generates coherent 480P / 540P / 768P videos from prompts and multimodal references, guided by up to 9 reference images, 3 reference videos, and 3 reference audios, with native stereo audio and flexible reference-based video generation on WaveSpeedAI infrastructure. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/1785815150161566057_2VF6wZoP.webp)
Minimax H3 Image To Video
MiniMax H3 Open Weights Image to Video animates a first-frame image, optionally with last-frame guidance, into coherent 480P / 540P / 768P videos with native stereo audio, 5-15 second duration, flexible aspect ratios, and per-second billing. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/1785814561183432100_tW5eoxGP.webp)
Minimax H3 Text To Video
MiniMax H3 Open Weights Text to Video generates coherent videos from text prompts, with 480P / 540P / 768P output, native stereo audio, 5-15 second duration, flexible aspect ratios, and per-second billing. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/1785415007635242852_UoxHQZ9H.webp)
H3 Reference To Video
MiniMax H3 Reference to Video generates coherent 2K videos from natural-language prompts and multimodal references, including images, videos, and audio, guiding subject consistency, motion, timing, visual style, and scene continuity. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/1785414895750297795_1JajtDMV.webp)
H3 Image To Video
MiniMax H3 Image to Video animates a first-frame image into a coherent 2K video, with natural-language motion instructions and optional last-frame control for consistent motion, scene continuity, and cinematic video generation. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/1785412796956877340_AktCLU4J.webp)
H3 Text To Video
MiniMax H3 Text to Video generates coherent 2K videos from text prompts, with flexible 5-15 second duration and adaptive or custom aspect ratios for cinematic scenes, creative videos, and production workflows. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
उदाहरण
MiniMax H3 को काम करते देखें
MiniMax H3 API से बनाए गए असली आउटपुट। किसी भी वीडियो पर होवर करके प्रीव्यू देखें, पूर्ण आकार का व्यूअर खोलने के लिए क्लिक करें।
कैसे करें
MiniMax H3 API का उपयोग कैसे करें
साइनअप से तैयार जनरेशन तक चार चरण। पूरे Python, Node.js और cURL उदाहरण नीचे API सेक्शन में हैं।
- 01
API कुंजी प्राप्त करें
WaveSpeedAI अकाउंट बनाएं और डैशबोर्ड से अपनी API कुंजी कॉपी करें। नए अकाउंट के साथ मुफ़्त शुरुआती क्रेडिट मिलते हैं — बिलिंग शुरू होने से पहले प्लेग्राउंड को कई दर्जन बार चलाने के लिए पर्याप्त।
- 02
prediction सबमिट करें
अपना इनपुट JSON के रूप में https://api.wavespeed.ai/api/v3/wavespeed-ai/minimax-h3/text-to-video पर POST करें। एंडपॉइंट तुरंत एक prediction id लौटाता है — जनरेशन असिंक्रोनस होते हैं, इसलिए इनफ़रेंस के दौरान आपको कनेक्शन खुला नहीं रखना पड़ता।
- 03
पूर्ण होने तक पोल करें
https://api.wavespeed.ai/api/v3/predictions/{request_id}/result पर GET करें। completed पर आउटपुट लौटाएं; failed, cancelled, timeout या deleted पर त्रुटि के साथ रुकें; बाकी हर status पर पोलिंग जारी रखें।
- 04
आउटपुट URL पढ़ें
status "completed" होने पर data.outputs[0] से URL पढ़ें। यह URL WaveSpeedAI CDN पर आपके बनाए मीडिया की ओर इशारा करता है — आपके कॉल किए गए MiniMax H3 वेरिएंट के अनुसार इमेज, वीडियो, ऑडियो या 3D फ़ाइल।
उपयोग के मामले
MiniMax H3 से आप क्या बना सकते हैं
आम वर्कफ़्लो जिनके लिए डेवलपर और क्रिएटर MiniMax H3 API का उपयोग करते हैं।
मौलिक कॉन्सेप्ट के लिए Text-to-video
wavespeed-ai/minimax-h3/text-to-video से प्रॉम्प्ट को मौलिक सिनेमैटिक सीन में बदलें — नेटिव स्टीरियो ऑडियो और लचीले आस्पेक्ट रेशियो के साथ सुसंगत 480p/540p/768p/1080p आउटपुट। इसे कॉन्सेप्ट विज़ुअलाइज़ेशन, स्टोरीटेलिंग, कैंपेन आइडिया और ऐसे शॉट्स के लिए इस्तेमाल करें जिन्हें मौजूदा आर्टवर्क से शुरू करने की ज़रूरत नहीं।
प्रोडक्ट शोकेस के लिए Image-to-video
प्रोडक्ट फ़ोटोग्राफ़ी, कैंपेन स्टिल, की आर्ट या कैरेक्टर इमेज को एनिमेट करें, और सोर्स की विज़ुअल दिशा बनाए रखें। प्रोडक्ट रीवील, सोशल विज्ञापन और मोशन-फ़र्स्ट लैंडिंग कंटेंट के लिए उपयोगी।
सुसंगत विषयों के लिए Reference-to-video
कैरेक्टर की पहचान, विषय के रूप-रंग, सीन डिज़ाइन या स्टाइल को निर्देशित करने के लिए 9 रेफ़रेंस इमेज, 3 रेफ़रेंस वीडियो और 3 रेफ़रेंस ऑडियो तक का उपयोग करें। बार-बार आने वाले कैरेक्टर, ब्रांडेड विज़ुअल और जुड़े हुए क्रिएटिव सेट के लिए H3 का वर्कफ़्लो Reference-to-video है।
सोशल मीडिया और विज्ञापन क्रिएटिव
कैंपेन, लॉन्च, सोशल फ़ीड और परफ़ॉर्मेंस मार्केटिंग के लिए शॉर्ट-फ़ॉर्म क्रिएटिव कॉन्सेप्ट बनाएं। API प्लेटफ़ॉर्म बदले बिना टेक्स्ट, तैयार स्टिल या स्वीकृत विज़ुअल रेफ़रेंस से आगे बढ़ें।
कैरेक्टर सीन और विज़ुअल स्टोरीटेलिंग
सोर्स मटेरियल के अनुसार चुने गए एंडपॉइंट से कैरेक्टर-केंद्रित सीन, स्टोरी बीट, मूड पीस और प्रीविज़ुअलाइज़ेशन बनाएं: नए सीन के लिए टेक्स्ट, एनिमेशन के लिए इमेज, या मज़बूत निरंतरता के लिए रेफ़रेंस।
स्केलेबल AI वीडियो जनरेशन पाइपलाइन
WaveSpeedAI के ज़रिए MiniMax H3 को एप्लिकेशन और ऑटोमेटेड क्रिएटिव वर्कफ़्लो में इंटीग्रेट करें। H3 के तीनों जनरेशन मोड में एक ही API key और एकसमान प्रेडिक्शन लाइफ़साइकिल इस्तेमाल करें।
सुझाव
MiniMax H3 के लिए प्रॉम्प्ट लिखने के सुझाव
MiniMax H3 से बेहतर आउटपुट पाने के व्यावहारिक सुझाव — प्रोडक्शन पाइपलाइन में वीडियो मॉडल पर काम करने वाले पैटर्न से लिए गए।
- 01
सोर्स मटेरियल के अनुसार एंडपॉइंट चुनें
मौलिक प्रॉम्प्ट के लिए text-to-video, एक सोर्स इमेज को एनिमेट करने के लिए image-to-video, और जब विज़ुअल रेफ़रेंस से पहचान, विषय के रूप-रंग, सीन डिज़ाइन या स्टाइल को निर्देशित करना हो तब reference-to-video इस्तेमाल करें।
- 02
मोशन को क्रम के रूप में बताएं
शुरुआती स्थिति, एक्शन, परिवेश की प्रतिक्रिया, कैमरा मूवमेंट और अंतिम स्थिति को कालक्रम से लिखें। असंबद्ध विज़ुअल विशेषणों की सूची के बजाय स्पष्ट क्रम H3 को मज़बूत मोशन योजना देता है।
- 03
कंपोज़िशन स्वीकृत हो तो image-to-video इस्तेमाल करें
अगर विषय, प्रोडक्ट, फ़्रेमिंग या आर्ट डायरेक्शन किसी स्टिल इमेज में तय है, तो टेक्स्ट से सीन दोबारा बनाने के बजाय image-to-video से शुरू करें। प्रॉम्प्ट को गति और कैमरा व्यवहार पर केंद्रित रखें।
- 04
हर रेफ़रेंस को एक साफ़ काम दें
reference-to-video के लिए बताएं कि रेफ़रेंस किसे नियंत्रित करें: कैरेक्टर की पहचान, प्रोडक्ट का रूप, विज़ुअल स्टाइल, परिवेश या कंपोज़िशन। एक ही रेफ़रेंस से शॉट के कई परस्पर विरोधी पहलू तय कराने से बचें।
- 05
छोटे सीन में एक मुख्य एक्शन रखें
असंबद्ध घटनाओं की पूरी श्रृंखला के बजाय एक केंद्रित एक्शन सुसंगत रूप से रेंडर करना आसान है। अलग स्टोरी बीट के लिए अलग शॉट बनाएं, फिर कॉन्सेप्ट को ज़्यादा नैरेटिव दायरा चाहिए तो उन्हें एडिटिंग में जोड़ें।
प्राइसिंग
MiniMax H3 API प्राइसिंग
प्राइसिंग प्रति आउटपुट है। अंतिम शुल्क उन पैरामीटर के अनुसार बदलता है जो आप हर वेरिएंट के प्लेग्राउंड में सेट करते हैं (रिज़ॉल्यूशन, अवधि, आउटपुट संख्या, रेफरेंस)।
API
MiniMax H3 API कॉल करें
wavespeed.ai/accesskey पर API कुंजी बनाएं, फिर REST के ज़रिए prediction सबमिट करें। प्लेग्राउंड इनपुट के किसी भी संयोजन के लिए सीधे पेस्ट करने योग्य नमूने तैयार करता है।
POSThttps://api.wavespeed.ai/api/v3/wavespeed-ai/minimax-h3/text-to-video
# 1. Submit the prediction.
SUBMIT_RESPONSE=$(curl --silent --show-error --fail-with-body \
-X POST "https://api.wavespeed.ai/api/v3/wavespeed-ai/minimax-h3/text-to-video" \
-H "Content-Type: application/json" \
-H "Authorization: Bearer $WAVESPEED_API_KEY" \
-d '{
"prompt": "A cinematic shot of a city at sunset, soft golden light",
"aspect_ratio": "16:9",
"resolution": "480p",
"duration": 5
}')
TASK=$(printf '%s' "$SUBMIT_RESPONSE" | jq 'if has("data") then .data else . end')
PREDICTION_ID=$(printf '%s' "$TASK" | jq -r '.id')
if [ -z "$PREDICTION_ID" ] || [ "$PREDICTION_ID" = "null" ]; then
printf 'Submission response did not contain a prediction id
' >&2
exit 1
fi
RESULT_URL="https://api.wavespeed.ai/api/v3/predictions/$PREDICTION_ID/result"
# 2. Poll until the prediction finishes.
while true; do
RESPONSE=$(curl --silent --show-error --fail-with-body "$RESULT_URL" \
-H "Authorization: Bearer $WAVESPEED_API_KEY")
RESULT=$(printf '%s' "$RESPONSE" | jq 'if has("data") then .data else . end')
STATUS=$(printf '%s' "$RESULT" | jq -r '.status')
case "$STATUS" in
completed) printf '%s\n' "$RESULT" | jq '.outputs'; break ;;
failed|cancelled|timeout|deleted) printf '%s\n' "$RESULT" | jq . >&2; exit 1 ;;
*) sleep 2 ;;
esac
doneconst submitUrl = "https://api.wavespeed.ai/api/v3/wavespeed-ai/minimax-h3/text-to-video";
const apiKey = process.env.WAVESPEED_API_KEY;
if (!apiKey) throw new Error('Set WAVESPEED_API_KEY');
async function requestJson(url, options = {}) {
const response = await fetch(url, options);
if (!response.ok) throw new Error(await response.text());
return response.json();
}
// 1. Submit the prediction.
const body = await requestJson(submitUrl, {
method: "POST",
headers: {
"Authorization": `Bearer ${apiKey}`,
"Content-Type": "application/json",
},
body: JSON.stringify({
"prompt": "A cinematic shot of a city at sunset, soft golden light",
"aspect_ratio": "16:9",
"resolution": "480p",
"duration": 5
}),
});
const task = body.data ?? body;
const resultUrl = `https://api.wavespeed.ai/api/v3/predictions/${task.id}/result`;
// 2. Poll until the prediction finishes.
while (true) {
const resultBody = await requestJson(resultUrl, {
headers: { "Authorization": `Bearer ${apiKey}` },
});
const result = resultBody.data ?? resultBody;
if (result.status === "completed") {
console.log(result.outputs);
break;
}
if (["failed", "cancelled", "timeout", "deleted"].includes(result.status)) throw new Error(JSON.stringify(result));
await new Promise(resolve => setTimeout(resolve, 2000));
}import json
import os
import time
from urllib.request import Request, urlopen
api_key = os.environ["WAVESPEED_API_KEY"]
headers = {"Authorization": f"Bearer {api_key}", "Content-Type": "application/json"}
payload = {
"prompt": "A cinematic shot of a city at sunset, soft golden light",
"aspect_ratio": "16:9",
"resolution": "480p",
"duration": 5
}
def request_json(url, data=None):
request = Request(url, data=data, headers=headers, method="POST" if data else "GET")
with urlopen(request) as response:
return json.load(response)
# 1. Submit the prediction.
body = request_json("https://api.wavespeed.ai/api/v3/wavespeed-ai/minimax-h3/text-to-video", json.dumps(payload).encode())
task = body.get("data", body)
result_url = f"https://api.wavespeed.ai/api/v3/predictions/{task['id']}/result"
# 2. Poll until the prediction finishes.
while True:
result_body = request_json(result_url)
result = result_body.get("data", result_body)
status = result.get("status")
if status == "completed":
print(result.get("outputs", []))
break
if status in {"failed", "cancelled", "timeout", "deleted"}:
raise RuntimeError(result)
time.sleep(2)तुलना
MiniMax H3 बनाम विकल्प
WaveSpeedAI पर समान मॉडल की तुलना में MiniMax H3 कब चुनें।
MiniMax H3 बनाम Hailuo 2.3
Hailuo 2.3 में Standard, Pro, Fast और Fast Pro वेरिएंट में टियर वाले text-to-video और image-to-video एंडपॉइंट मिलते हैं। MiniMax H3 तीन एंडपॉइंट का केंद्रित सुइट इस्तेमाल करता है और विज़ुअल गाइडेंस व विषय की निरंतरता के लिए reference-to-video को प्रमुख वर्कफ़्लो के रूप में जोड़ता है।
MiniMax H3 बनाम Seedance 2.0
Seedance 2.0 text-to-video, image-to-video, video-edit, video-extend, नेटिव ऑडियो और कई परफ़ॉर्मेंस टियर के साथ एक व्यापक प्रोडक्शन फ़ैमिली देता है। जब वर्कफ़्लो टेक्स्ट, सोर्स इमेज या विज़ुअल रेफ़रेंस से जनरेट करने पर केंद्रित हो, तब MiniMax H3 सरल विकल्प है।
MiniMax H3 बनाम Kling 3.0
Kling 3.0 टियर वाली आउटपुट क्वालिटी और समर्पित मोशन-कंट्रोल एंडपॉइंट पर ज़ोर देता है। MiniMax H3 अपना API सोर्स टाइप के आसपास व्यवस्थित करता है, जिसमें पहचान, विषय और स्टाइल-निर्देशित जनरेशन के लिए समर्पित reference-to-video रूट शामिल है।
FAQ
MiniMax H3 API — अक्सर पूछे जाने वाले प्रश्न
प्राइसिंग, लाइसेंस, इंटीग्रेशन — WaveSpeedAI पर MiniMax H3 चलाने के बारे में आम सवाल।
MiniMax H3 API क्या है?
MiniMax H3 API WaveSpeedAI पर तीन मॉडल वाला AI वीडियो जनरेशन सुइट है। यह प्रॉम्प्ट-आधारित सीन के लिए text-to-video, स्टिल इमेज को एनिमेट करने के लिए image-to-video, और विज़ुअल रेफ़रेंस से निर्देशित जनरेशन के लिए reference-to-video सपोर्ट करता है।
मुझे कौन-सा MiniMax H3 API एंडपॉइंट इस्तेमाल करना चाहिए?
WaveSpeed इन्फ्रास्ट्रक्चर पर अनुशंसित ओपन-वेट्स डिप्लॉयमेंट से शुरू करें: प्रॉम्प्ट से शुरू करते समय wavespeed-ai/minimax-h3/text-to-video, सोर्स इमेज को एनिमेट करते समय wavespeed-ai/minimax-h3/image-to-video, और जब विज़ुअल रेफ़रेंस से विषय की पहचान, रूप-रंग या स्टाइल निर्देशित करनी हो तब wavespeed-ai/minimax-h3/reference-to-video। आधिकारिक minimax/h3 एंडपॉइंट उसी फ़ैमिली में उपलब्ध हैं।
क्या MiniMax H3 image-to-video जनरेशन सपोर्ट करता है?
हाँ। wavespeed-ai/minimax-h3/image-to-video एंडपॉइंट पहले फ़्रेम की इमेज — वैकल्पिक रूप से अंतिम फ़्रेम के साथ — को नेटिव स्टीरियो ऑडियो वाले सुसंगत वीडियो में एनिमेट करता है, जिससे यह प्रोडक्ट फ़ोटो, कैंपेन आर्ट, कैरेक्टर इमेज, कॉन्सेप्ट फ़्रेम और अन्य मौजूदा विज़ुअल को एनिमेट करने के लिए उपयुक्त है।
MiniMax H3 reference-to-video क्या है?
MiniMax H3 reference-to-video 9 रेफ़रेंस इमेज, 3 रेफ़रेंस वीडियो और 3 रेफ़रेंस ऑडियो तक से निर्देशित नया वीडियो बनाता है, जो विषय, कैरेक्टर की पहचान, स्टाइल या सीन की भाषा को दिशा देते हैं। बार-बार आने वाले कैरेक्टर और ब्रांड-संगत क्रिएटिव वर्कफ़्लो के लिए यह पसंदीदा H3 एंडपॉइंट है।
मैं WaveSpeedAI पर MiniMax H3 API कैसे इस्तेमाल करूं?
अपने सोर्स के अनुरूप H3 एंडपॉइंट चुनें, अपनी WaveSpeedAI API key से ऑथेंटिकेट करें, मॉडल इनपुट JSON के रूप में सबमिट करें, और जनरेट किया गया वीडियो पाने के लिए लौटाई गई प्रेडिक्शन ID या रिज़ल्ट URL का उपयोग करें। इस पेज के एंडपॉइंट कार्ड और कोड सैंपल मौजूदा रिक्वेस्ट स्कीमा देते हैं।
MiniMax H3 API को कैसे कॉल करें?
WaveSpeedAI अकाउंट बनाएं, /accesskey से अपनी API कुंजी कॉपी करें, फिर अपना इनपुट JSON के रूप में https://api.wavespeed.ai/api/v3/wavespeed-ai/minimax-h3/text-to-video पर POST करें। एंडपॉइंट एक prediction id लौटाता है। रिज़ल्ट एंडपॉइंट को लगभग हर 2 सेकंड पर पोल करना शुरू करें, लंबे टास्क के लिए अंतराल बढ़ाएं, और किसी भी अंतिम status पर रुक जाएं। प्रोडक्शन-उन्मुख Python / Node.js / cURL उदाहरण ऊपर हैं।
MiniMax H3 API की कीमत कितनी है?
MiniMax H3 $0.02 प्रति रन से शुरू होता है। सटीक लागत आपके सेट किए पैरामीटर (रिज़ॉल्यूशन, अवधि, आउटपुट संख्या, रेफरेंस) के अनुसार बदलती है। प्लेग्राउंड में Generate बटन के पास लाइव कॉस्ट प्रीव्यू आपके वर्तमान इनपुट की सटीक कीमत दिखाता है।
MiniMax H3 के कौन-से वेरिएंट उपलब्ध हैं?
WaveSpeedAI पर 20 लाइव MiniMax H3 एंडपॉइंट हैं: wavespeed-ai/minimax-h3/controlnet-union, wavespeed-ai/minimax-h3-singularity/image-to-video-lora, wavespeed-ai/minimax-h3-singularity/image-to-video, wavespeed-ai/minimax-h3-singularity/reference-to-video-lora, wavespeed-ai/minimax-h3-singularity/reference-to-video, wavespeed-ai/minimax-h3/image-edit-lora, wavespeed-ai/minimax-h3/text-to-image-lora, wavespeed-ai/minimax-h3/image-edit, और अन्य। हर वेरिएंट का अपना प्लेग्राउंड पेज और प्राइसिंग है।
क्या मैं MiniMax H3 के आउटपुट का व्यावसायिक उपयोग कर सकता हूँ?
व्यावसायिक उपयोग के अधिकार MiniMax मॉडल लाइसेंस के अनुसार होते हैं। अधिकांश MiniMax मॉडल आउटपुट के व्यावसायिक उपयोग की अनुमति देते हैं; विशिष्ट लाइसेंस सारांश के लिए हर मॉडल का प्लेग्राउंड पेज और प्लेटफ़ॉर्म-स्तर की शर्तों के लिए WaveSpeedAI की सेवा की शर्तें देखें।
सीधे जाने के बजाय WaveSpeedAI पर MiniMax H3 क्यों इस्तेमाल करें?
एक API कुंजी + एक बिलिंग अकाउंट, MiniMax H3 और अन्य प्रदाताओं के 1,000+ दूसरे AI मॉडल के लिए। प्रति-वेंडर SDK सेटअप नहीं, अलग रेट-लिमिट नहीं, हर वेंडर के लिए इंटीग्रेशन कोड दोबारा लिखना नहीं। प्राइसिंग आमतौर पर MiniMax के सीधे API के बराबर या उससे कम होती है।
प्रदाता
MiniMax के बारे में
WaveSpeedAI पर MiniMax H3 और MiniMax की व्यापक मॉडल श्रृंखला के पीछे की टीम।
MiniMax एक चीनी AI लैब है जो Hailuo वीडियो जनरेशन और स्पीच मॉडल के लिए जानी जाती है। Hailuo 2.3 Standard, Pro और Fast टियर के साथ फ़िज़िक्स-सजग text-to-video और image-to-video देता है — क्रिएटर और मार्केटिंग वर्कफ़्लो के लिए 2.5× दक्षता और मज़बूत जटिल-निर्देश सटीकता के आसपास पेश किया गया।
WaveSpeedAI पर MiniMax H3 के साथ बनाना शुरू करें
साइनअप पर मुफ़्त शुरुआती क्रेडिट। MiniMax और हर अन्य प्रदाता के 1,000+ AI मॉडल के लिए एक API कुंजी।