Nano Banana 2.1 출시 — Google 최신 모델 | 지금 체험 →
OpenAI Models

OpenAI Models

OpenAI's state-of-the-art AI models for text, image, and multimodal applications, Sora 2 is included

OpenAI's state-of-the-art AI models for text, image, and multimodal applications, Sora 2 is included

전체 모델

19개 모델
openai/gpt-image-2.5-sunburst/edit
image-to-image$0.0390

openai/gpt-image-2.5-sunburst/edit

OpenAI's GPT Image 2.5 Sunburst Edit edits one or more reference images from natural-language instructions, with five quality tiers up to 4K. Sunburst is the precision-focused GPT Image 2.5 tier that spends more time per image for extra fidelity on intricate detail. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

openai/gpt-image-2.5-sunburst/text-to-image
text-to-image$0.0240

openai/gpt-image-2.5-sunburst/text-to-image

OpenAI's GPT Image 2.5 Sunburst Text-to-Image generates high-quality images from natural-language prompts, with five quality tiers up to 4K. Sunburst is the precision-focused GPT Image 2.5 tier that spends more time per image for extra fidelity on intricate detail. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

openai/gpt-image-2.5-flare/edit
image-to-image$0.0390

openai/gpt-image-2.5-flare/edit

OpenAI's GPT Image 2.5 Flare Edit edits one or more reference images from natural-language instructions, with five quality tiers up to 4K. Flare is the fast, balanced GPT Image 2.5 tier for everyday generation at low latency. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

openai/gpt-image-2.5-flare/text-to-image
text-to-image$0.0240

openai/gpt-image-2.5-flare/text-to-image

OpenAI's GPT Image 2.5 Flare Text-to-Image generates high-quality images from natural-language prompts, with five quality tiers up to 4K. Flare is the fast, balanced GPT Image 2.5 tier for everyday generation at low latency. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

openai/gpt-image-2/edit5% OFF
image-to-image$0.0700$0.0665

openai/gpt-image-2/edit

OpenAI's GPT Image 2 Edit enables image editing from natural-language instructions with one or more reference images. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

openai/gpt-image-2/text-to-image5% OFF
text-to-image$0.0600$0.0570

openai/gpt-image-2/text-to-image

OpenAI's GPT Image 2 Text-to-Image generates high-quality images from natural-language prompts. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

openai/gpt-6-astra/image-to-3d
image-to-3d$8.0000

openai/gpt-6-astra/image-to-3d

GPT-6 Astra Image-to-3D generates textured 3D models or scenes from reference images, with optional text guidance for image-guided 3D asset creation, product visualization, concept design, virtual scenes, and production workflows. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

openai/gpt-6-astra/text-to-3d
text-to-3d$8.0000

openai/gpt-6-astra/text-to-3d

GPT-6 Astra Text-to-3D generates textured 3D models or scenes from text descriptions, supporting fast 3D asset creation for game assets, product visualization, concept design, virtual scenes, and production workflows. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

wavespeed-ai/openai-whisper-with-video
speech-to-text$0.0010

wavespeed-ai/openai-whisper-with-video

OpenAI Whisper Large v3 (Video-to-Text) delivers high-accuracy multilingual transcription directly from video files, with automatic language detection and optional timestamped, subtitle-ready segments. Built for stable production use with a ready-to-use REST API, fast response, no cold starts, and predictable pricing.

openai/sora-2-pro/text-to-video
text-to-video$1.2000

openai/sora-2-pro/text-to-video

OpenAI Sora 2 Pro is a state-of-the-art text-to-video model with realistic physics, synchronized audio, and strong steerability. Supports multiple resolutions up to 1080p and durations up to 20 seconds. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

openai/sora-2-pro/image-to-video
image-to-video$1.2000

openai/sora-2-pro/image-to-video

OpenAI Sora 2 Pro Image-to-Video creates physics-aware, realistic videos from reference images with synchronized audio and strong steerability. Supports 720p and 1080p resolutions with durations up to 20 seconds. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

openai/sora-2/image-to-video
image-to-video$0.4000

openai/sora-2/image-to-video

OpenAI Sora 2 generates realistic image-to-video content with synchronized audio, improved physics, sharper realism and steerability. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

openai/gpt-image-1.5/edit
image-to-image$0.1000

openai/gpt-image-1.5/edit

GPT Image 1.5 Edit is OpenAI’s image model for precise, natural-language edits. Add/remove objects, swap backgrounds, retouch faces, adjust colors/lighting, edit text/graphics, crop/resize, and apply hex color control. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

openai/gpt-image-1.5/text-to-image
text-to-image$0.0400

openai/gpt-image-1.5/text-to-image

GPT Image 1.5 text to image is OpenAI’s fast, cost-efficient text-to-image generator powered by GPT-5 guidance. Create photorealistic shots, product renders, concept art, and stylized graphics from natural-language prompts (optionally conditioned with an image). Supports custom aspect ratios, seeds, negative prompts, hex color hints, and style presets. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

wavespeed-ai/openai-whisper-turbo
speech-to-text$0.0007

wavespeed-ai/openai-whisper-turbo

Accurate speech-to-text with OpenAI Whisper Large v3 Turbo: multilingual transcripts with auto language detection and punctuation. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

openai/sora-2/text-to-video-pro
text-to-video$1.2000

openai/sora-2/text-to-video-pro

OpenAI Sora 2 Text-to-Video Pro creates high-fidelity videos with synchronized audio, realistic physics, and enhanced steerability. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

openai/sora-2/text-to-video
text-to-video$0.4000

openai/sora-2/text-to-video

OpenAI Sora 2 is a state-of-the-art text-to-video model with realistic visuals, accurate physics, synchronized audio, and strong steerability. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

openai/sora-2/image-to-video-pro
image-to-video$1.2000

openai/sora-2/image-to-video-pro

OpenAI Sora 2 Image-to-Video Pro creates physics-aware, realistic videos with synchronized audio and greater steerability. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

wavespeed-ai/openai-whisper
speech-to-text$0.0010

wavespeed-ai/openai-whisper

Whisper Large v3 speech-to-text: instant, accurate multilingual transcripts with automatic language detection and punctuation. Upload audio to get transcripts. Ready-to-use REST API, no coldstarts, affordable pricing.

OpenAI Models

OpenAI Models on WaveSpeedAI bring together advanced image, video, and speech AI models for creative and production workflows. The collection is centered on GPT Image 2, OpenAI’s latest image generation and editing model family, while also including Sora video generation, GPT Image 1.5, GPT Image 1 Mini, DALL·E models, and Whisper speech recognition.

Built for creators, developers, designers, marketers, and AI applications, OpenAI Models support high-quality text-to-image generation, natural-language image editing, image-to-video generation, text-to-video creation, transcription, and multimodal creative workflows. The suite is suitable for marketing visuals, product concepts, UI mockups, social media assets, campaign creatives, cinematic videos, and scalable content production.

Core Model Capabilities

GPT Image 2 — Flagship Image Generation and Editing:

GPT Image 2 is the main highlight of this collection, delivering high-quality image generation and editing with strong prompt understanding, clean composition, polished aesthetics, and improved visual coherence. It is designed for professional creative workflows that need reliable results, detailed visual control, and production-ready output.

Text-to-Image Generation:

Generate high-quality images from natural-language prompts for campaign assets, UI concepts, product visuals, concept art, social media content, brand creatives, and rapid visual ideation.

Natural-Language Image Editing:

Use GPT Image 2 Edit to modify images with text instructions and reference inputs while preserving visual consistency, style coherence, composition, and fine details. It is useful for marketing asset refinement, product image editing, design iteration, and creative retouching.

GPT Image 1.5 and GPT Image 1 Mini:

Use GPT Image 1.5 and GPT Image 1 Mini for cost-efficient image generation, fast creative iteration, lightweight image editing, and scalable visual production workflows.

Sora Video Generation:

Use Sora and Sora 2 models for image-to-video and text-to-video workflows, turning prompts or still images into cinematic video clips with coherent motion, stable identities, and smooth camera movement.

DALL·E Image Models:

DALL·E models provide additional text-to-image options for illustration, concept exploration, quick drafts, and stylized image creation.

Whisper Speech Recognition:

Whisper and Whisper Turbo provide multilingual speech recognition for transcription, automatic language detection, punctuation, and large-scale audio processing workflows.

Text-to-3D Generation:

Use GPT-6 Astra Text-to-3D to generate 3D assets directly from natural-language prompts, making it easier to move from concept to usable 3D content for design, visualization, and creative production.

Image-to-3D Generation:

Use GPT-6 Astra Image-to-3D to convert reference images into 3D assets, helping creators build 3D objects from visual inputs while preserving shape cues and overall appearance.

OpenAI Models on WaveSpeedAI give creators and developers fast access to OpenAI’s image, video, and speech models through scalable APIs, flexible pricing, and production-ready creative capabilities.

OpenAI Models API — 가격 및 성능

OpenAI Models 컬렉션의 모든 모델을 단일 REST API로 실행하세요. 생성당 과금 — 구독 없음, 최소 요금 없음 — 99.9% 가동률 인프라에서 업계 최고의 지연 시간을 제공합니다.

WaveSpeedAI에서 OpenAI Models을 사용하는 이유

투명한 가격

모든 OpenAI Models 모델에 대한 호출당 가격. 가격은 각 모델 페이지에 표시되며 플랫폼 수수료는 추가되지 않습니다.

낮은 지연 시간에 최적화

대부분의 OpenAI Models 이미지 모델은 2초 이내에 완료됩니다. 비디오 및 3D 모델은 셀프 호스팅 대안보다 몇 배 더 빠릅니다.

99.9% 가동률

다중 리전 페일오버와 자동 재시도로 프로바이더 장애 중에도 운영 트래픽을 온라인 상태로 유지합니다.

자주 묻는 질문

OpenAI Models API는 얼마인가요?+

각 모델에는 모델 페이지에 호출당 자체 가격이 표시되어 있습니다. 성공한 생성 단위로 청구되며 구독 요금이나 최소 요금은 없습니다.

WaveSpeedAI에서 OpenAI Models 모델은 얼마나 빠릅니까?+

이 컬렉션의 이미지 모델은 일반적으로 2초 이내에 완료됩니다. 비디오 및 3D 모델은 길이와 해상도에 따라 다르지만 보통 셀프 호스팅 실행보다 몇 배 더 빠릅니다.

신용카드 없이 API를 시험해 볼 수 있나요?+

조건을 충족하는 신규 계정은 신용카드 없이 OpenAI Models 모델을 체험할 수 있는 $1 프로모션 크레딧을 받을 수 있습니다. 모든 가입에 체험 크레딧이 보장되지는 않으므로 생성 전에 계정 잔액을 확인하세요.

속도 제한이 있나요?+

표준 계정에는 넉넉한 동시 작업 제한이 있습니다. Enterprise 플랜은 맞춤형 RPM, 더 높은 동시성, 전용 용량을 제공합니다 — 자세한 내용은 영업팀에 문의하세요.