Nano Banana 2.1 출시 — Google 최신 모델 | 지금 체험 →
MiniMax H3 Models

MiniMax H3 Models

MiniMax H3 delivers cinematic AI video generation for text, image, and reference-based workflows

MiniMax H3 delivers cinematic AI video generation for text, image, and reference-based workflows

전체 모델

20개 모델
wavespeed-ai/minimax-h3/video-edit50% OFF
video-to-video$0.2500$0.1250

wavespeed-ai/minimax-h3/video-edit

MiniMax H3 Open Weights Video-Edit edits an input video from a natural-language prompt. The input video drives subject identity, composition, and motion while the model rewrites lighting, style, weather, environment, or specific elements as instructed, with native stereo audio generated in the same pass. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

wavespeed-ai/minimax-h3/text-to-image-lora
text-to-image$0.0200

wavespeed-ai/minimax-h3/text-to-image-lora

MiniMax H3 Open Weights Text-to-Image generates photorealistic, cinema-grade stills from a text prompt at 1K or 2K resolution with fifteen aspect ratios. Loads up to 3 custom LoRA weights per request. Ready-to-use REST inference API, best performance, no cold starts, affordable pricing.

wavespeed-ai/minimax-h3/controlnet-union50% OFF
motion-control$0.3000$0.1500

wavespeed-ai/minimax-h3/controlnet-union

MiniMax H3 Open Weights ControlNet Union generates a new video that follows the motion and composition of a source video. Pose, depth, edges, lines, scribble or grayscale structure is extracted from the source automatically and guides the output, optionally with reference images for the subject or style, with native stereo audio generated in the same pass. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

wavespeed-ai/minimax-h3/text-to-video50% OFF
text-to-video$0.2000$0.1000

wavespeed-ai/minimax-h3/text-to-video

MiniMax H3 Open Weights Text to Video generates coherent videos from text prompts, with 480P / 540P / 768P output, native stereo audio, 5-15 second duration, flexible aspect ratios, and per-second billing. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

wavespeed-ai/minimax-h3/image-to-video50% OFF
image-to-video$0.2000$0.1000

wavespeed-ai/minimax-h3/image-to-video

MiniMax H3 Open Weights Image to Video animates a first-frame image, optionally with last-frame guidance, into coherent 480P / 540P / 768P videos with native stereo audio, 5-15 second duration, flexible aspect ratios, and per-second billing. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

wavespeed-ai/minimax-h3/reference-to-video50% OFF
reference-to-video$0.2500$0.1250

wavespeed-ai/minimax-h3/reference-to-video

MiniMax H3 Open Weights Reference to Video generates coherent 480P / 540P / 768P videos from prompts and multimodal references, guided by up to 9 reference images, 3 reference videos, and 3 reference audios, with native stereo audio and flexible reference-based video generation on WaveSpeedAI infrastructure. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

wavespeed-ai/minimax-h3/text-to-video-lora50% OFF
text-to-video$0.2500$0.1250

wavespeed-ai/minimax-h3/text-to-video-lora

MiniMax H3 Open Weights Text to Video with custom LoRA support generates coherent videos from text prompts, with 480P / 540P / 768P output, native stereo audio, 5-15 second duration, flexible aspect ratios, and per-second billing. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

wavespeed-ai/minimax-h3/image-to-video-lora50% OFF
image-to-video$0.2500$0.1250

wavespeed-ai/minimax-h3/image-to-video-lora

MiniMax H3 Open Weights Image to Video animates a first-frame image, optionally with last-frame guidance, into coherent 480P / 540P / 768P videos with native stereo audio, 5-15 second duration, flexible aspect ratios, and per-second billing. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

wavespeed-ai/minimax-h3/video-extend50% OFF
video-extend$0.2000$0.1000

wavespeed-ai/minimax-h3/video-extend

MiniMax H3 Open Weights Video-Extend appends a new cinematic continuation to an existing video. A fresh segment with native stereo audio is generated from the input video's last frame and a natural-language prompt, then the original and new segment are concatenated into a single output. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

wavespeed-ai/minimax-h3-singularity/image-to-video-lora50% OFF
image-to-video$0.3125$0.1563

wavespeed-ai/minimax-h3-singularity/image-to-video-lora

MiniMax H3 Singularity Open Weights Image to Video animates a first-frame image, optionally with last-frame guidance, into coherent 480P / 540P / 768P videos with native stereo audio, 5-15 second duration, flexible aspect ratios, and per-second billing. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

wavespeed-ai/minimax-h3/image-edit
image-to-image$0.0300

wavespeed-ai/minimax-h3/image-edit

MiniMax H3 Open Weights Image Edit re-renders a subject from up to 9 reference images into a new scene, outfit or style from a text instruction, preserving identity at 1K or 2K resolution. Ready-to-use REST inference API, best performance, no cold starts, affordable pricing.

wavespeed-ai/minimax-h3/text-to-image
text-to-image$0.0200

wavespeed-ai/minimax-h3/text-to-image

MiniMax H3 Open Weights Text-to-Image generates photorealistic, cinema-grade stills from a text prompt at 1K or 2K resolution with fifteen aspect ratios. Ready-to-use REST inference API, best performance, no cold starts, affordable pricing.

wavespeed-ai/minimax-h3-singularity/image-to-video50% OFF
image-to-video$0.2500$0.1250

wavespeed-ai/minimax-h3-singularity/image-to-video

MiniMax H3 Singularity Open Weights Image-to-Video animates a first-frame image, with optional last-frame guidance, into coherent 480P / 540P / 768P videos with native stereo audio, 5-15 second duration, flexible aspect ratios, and per-second billing. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

wavespeed-ai/minimax-h3/reference-to-video-lora50% OFF
reference-to-video$0.3000$0.1500

wavespeed-ai/minimax-h3/reference-to-video-lora

MiniMax H3 Open Weights Reference to Video with custom LoRA support generates coherent 480P / 540P / 768P videos from prompts and multimodal references, guided by up to 9 reference images, 3 reference videos, and 3 reference audios, with native stereo audio and flexible reference-based video generation. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

wavespeed-ai/minimax-h3/image-edit-lora
image-to-image$0.0300

wavespeed-ai/minimax-h3/image-edit-lora

MiniMax H3 Open Weights Image Edit re-renders a subject from up to 9 reference images into a new scene, outfit or style from a text instruction, preserving identity at 1K or 2K resolution. Loads up to 3 custom LoRA weights per request. Ready-to-use REST inference API, best performance, no cold starts, affordable pricing.

wavespeed-ai/minimax-h3-singularity/reference-to-video50% OFF
reference-to-video$0.2500$0.1250

wavespeed-ai/minimax-h3-singularity/reference-to-video

MiniMax H3 Singularity Open Weights Reference to Video generates coherent 480P / 540P / 768P videos from prompts and multimodal references, guided by up to 9 reference images, 3 reference videos, and 3 reference audios, with native stereo audio and flexible reference-based video generation on WaveSpeedAI infrastructure. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

wavespeed-ai/minimax-h3-singularity/reference-to-video-lora50% OFF
reference-to-video$0.3000$0.1500

wavespeed-ai/minimax-h3-singularity/reference-to-video-lora

MiniMax H3 Singularity Open Weights Reference to Video with custom LoRA support generates coherent 480P / 540P / 768P videos from prompts and multimodal references, guided by up to 9 reference images, 3 reference videos, and 3 reference audios, with native stereo audio and flexible reference-based video generation. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

minimax/h3/text-to-video
text-to-video$0.7000

minimax/h3/text-to-video

MiniMax H3 Text to Video generates coherent 2K videos from text prompts, with flexible 5-15 second duration and adaptive or custom aspect ratios for cinematic scenes, creative videos, and production workflows. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

minimax/h3/image-to-video
image-to-video$0.7000

minimax/h3/image-to-video

MiniMax H3 Image to Video animates a first-frame image into a coherent 2K video, with natural-language motion instructions and optional last-frame control for consistent motion, scene continuity, and cinematic video generation. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

minimax/h3/reference-to-video
reference-to-video$0.7000

minimax/h3/reference-to-video

MiniMax H3 Reference to Video generates coherent 2K videos from natural-language prompts and multimodal references, including images, videos, and audio, guiding subject consistency, motion, timing, visual style, and scene continuity. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

MiniMax H3 Models

MiniMax H3 provides a focused AI image and video generation model suite for text-to-video, image-to-video, reference-to-video, text-to-image, image editing, LoRA-powered generation, video editing, video extension, and MiniMax H3 Singularity workflows. The collection is designed to help creators, developers, marketers, and AI applications generate high-quality visual content from prompts, still images, visual references, and custom LoRA styles with strong motion quality, subject consistency, and scene coherence.

Built for fast and flexible creative production, MiniMax H3 supports a wide range of use cases including social media videos, ad creatives, product showcases, character scenes, storytelling, concept visualization, and scalable video generation pipelines.

Core Model Capabilities

Text-to-Video Generation:

Generate cinematic videos directly from natural-language prompts with strong scene understanding, camera movement, motion direction, lighting control, and visual style alignment.

Image-to-Video Generation:

Animate still images into dynamic video clips while preserving the original subject, composition, identity, and visual style.

Reference-to-Video Generation:

Create videos guided by reference images or visual inputs, helping maintain subject identity, style consistency, and scene continuity across generated motion.

Text-to-Image Generation:

Generate high-quality images directly from text prompts for concept art, product visuals, key art, social media assets, marketing content, and branded creative workflows.

Image Editing:

Edit and transform existing images with natural-language instructions while preserving composition, subject identity, and visual consistency. This is useful for creative retouching, scene changes, style adjustments, object replacement, and production-ready visual refinement.

LoRA-Powered Generation:

Use MiniMax H3 LoRA models for customized image generation workflows, including text-to-image, and image edit. LoRA support helps generate consistent characters, personalized styles, branded visual identities, and repeatable creative direction across both image and video outputs.

LoRA-Powered Video Generation:

Use MiniMax H3 LoRA models for customized text-to-video, image-to-video, and reference-to-video workflows. LoRA support helps generate videos with consistent characters, personalized styles, branded visual identities, and repeatable creative direction.

Video Editing:

Edit and transform existing videos with natural-language prompts while preserving core motion, subject structure, and scene continuity. This is useful for visual restyling, subject refinement, scene adjustment, and creative video adaptation.

Video Extension:

Extend existing video clips into longer continuous sequences while maintaining motion flow, character consistency, visual style, and cinematic scene coherence.

Cinematic Motion Quality:

Produce videos with smooth movement, stable subjects, natural transitions, and coherent visual structure for creative and commercial workflows.

Creative Video Production:

Support use cases such as short-form content, product videos, character animation, brand visuals, storytelling scenes, and AI-generated promotional videos.

Developer-Friendly Video API:

Access MiniMax H3 models through scalable APIs for automated video generation, high-volume creative workflows, and production-ready AI video applications.

MiniMax H3 Singularity:

Use MiniMax H3 Singularity models for image-to-video and reference-to-video generation, with dedicated LoRA-enabled variants for customized creative workflows. The Singularity lineup expands MiniMax H3 with additional options for animating source images, generating reference-guided video, and applying custom LoRA styles for more personalized and repeatable video creation.

MiniMax H3 Models on WaveSpeedAI give creators and developers fast access to text-to-video, image-to-video, and reference-to-video generation with flexible pricing, scalable API access, and production-ready video quality.

MiniMax H3 Models API — 가격 및 성능

MiniMax H3 Models 컬렉션의 모든 모델을 단일 REST API로 실행하세요. 생성당 과금 — 구독 없음, 최소 요금 없음 — 99.9% 가동률 인프라에서 업계 최고의 지연 시간을 제공합니다.

WaveSpeedAI에서 MiniMax H3 Models을 사용하는 이유

투명한 가격

모든 MiniMax H3 Models 모델에 대한 호출당 가격. 가격은 각 모델 페이지에 표시되며 플랫폼 수수료는 추가되지 않습니다.

낮은 지연 시간에 최적화

대부분의 MiniMax H3 Models 이미지 모델은 2초 이내에 완료됩니다. 비디오 및 3D 모델은 셀프 호스팅 대안보다 몇 배 더 빠릅니다.

99.9% 가동률

다중 리전 페일오버와 자동 재시도로 프로바이더 장애 중에도 운영 트래픽을 온라인 상태로 유지합니다.

자주 묻는 질문

MiniMax H3 Models API는 얼마인가요?+

각 모델에는 모델 페이지에 호출당 자체 가격이 표시되어 있습니다. 성공한 생성 단위로 청구되며 구독 요금이나 최소 요금은 없습니다.

WaveSpeedAI에서 MiniMax H3 Models 모델은 얼마나 빠릅니까?+

이 컬렉션의 이미지 모델은 일반적으로 2초 이내에 완료됩니다. 비디오 및 3D 모델은 길이와 해상도에 따라 다르지만 보통 셀프 호스팅 실행보다 몇 배 더 빠릅니다.

신용카드 없이 API를 시험해 볼 수 있나요?+

조건을 충족하는 신규 계정은 신용카드 없이 MiniMax H3 Models 모델을 체험할 수 있는 $1 프로모션 크레딧을 받을 수 있습니다. 모든 가입에 체험 크레딧이 보장되지는 않으므로 생성 전에 계정 잔액을 확인하세요.

속도 제한이 있나요?+

표준 계정에는 넉넉한 동시 작업 제한이 있습니다. Enterprise 플랜은 맞춤형 RPM, 더 높은 동시성, 전용 용량을 제공합니다 — 자세한 내용은 영업팀에 문의하세요.