Nano Banana 2.1 HADIR — Terbaru dari Google | Coba sekarang →
OpenAI Models

OpenAI Models

OpenAI's state-of-the-art AI models for text, image, and multimodal applications, Sora 2 is included

OpenAI's state-of-the-art AI models for text, image, and multimodal applications, Sora 2 is included

Semua model

19 model
openai/gpt-image-2.5-sunburst/edit
image-to-image$0.0390

openai/gpt-image-2.5-sunburst/edit

OpenAI's GPT Image 2.5 Sunburst Edit edits one or more reference images from natural-language instructions, with five quality tiers up to 4K. Sunburst is the precision-focused GPT Image 2.5 tier that spends more time per image for extra fidelity on intricate detail. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

openai/gpt-image-2.5-sunburst/text-to-image
text-to-image$0.0240

openai/gpt-image-2.5-sunburst/text-to-image

OpenAI's GPT Image 2.5 Sunburst Text-to-Image generates high-quality images from natural-language prompts, with five quality tiers up to 4K. Sunburst is the precision-focused GPT Image 2.5 tier that spends more time per image for extra fidelity on intricate detail. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

openai/gpt-image-2.5-flare/edit
image-to-image$0.0390

openai/gpt-image-2.5-flare/edit

OpenAI's GPT Image 2.5 Flare Edit edits one or more reference images from natural-language instructions, with five quality tiers up to 4K. Flare is the fast, balanced GPT Image 2.5 tier for everyday generation at low latency. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

openai/gpt-image-2.5-flare/text-to-image
text-to-image$0.0240

openai/gpt-image-2.5-flare/text-to-image

OpenAI's GPT Image 2.5 Flare Text-to-Image generates high-quality images from natural-language prompts, with five quality tiers up to 4K. Flare is the fast, balanced GPT Image 2.5 tier for everyday generation at low latency. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

openai/gpt-image-2/edit5% OFF
image-to-image$0.0700$0.0665

openai/gpt-image-2/edit

OpenAI's GPT Image 2 Edit enables image editing from natural-language instructions with one or more reference images. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

openai/gpt-image-2/text-to-image5% OFF
text-to-image$0.0600$0.0570

openai/gpt-image-2/text-to-image

OpenAI's GPT Image 2 Text-to-Image generates high-quality images from natural-language prompts. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

openai/gpt-6-astra/image-to-3d
image-to-3d$8.0000

openai/gpt-6-astra/image-to-3d

GPT-6 Astra Image-to-3D generates textured 3D models or scenes from reference images, with optional text guidance for image-guided 3D asset creation, product visualization, concept design, virtual scenes, and production workflows. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

openai/gpt-6-astra/text-to-3d
text-to-3d$8.0000

openai/gpt-6-astra/text-to-3d

GPT-6 Astra Text-to-3D generates textured 3D models or scenes from text descriptions, supporting fast 3D asset creation for game assets, product visualization, concept design, virtual scenes, and production workflows. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

wavespeed-ai/openai-whisper-with-video
speech-to-text$0.0010

wavespeed-ai/openai-whisper-with-video

OpenAI Whisper Large v3 (Video-to-Text) delivers high-accuracy multilingual transcription directly from video files, with automatic language detection and optional timestamped, subtitle-ready segments. Built for stable production use with a ready-to-use REST API, fast response, no cold starts, and predictable pricing.

openai/sora-2-pro/text-to-video
text-to-video$1.2000

openai/sora-2-pro/text-to-video

OpenAI Sora 2 Pro is a state-of-the-art text-to-video model with realistic physics, synchronized audio, and strong steerability. Supports multiple resolutions up to 1080p and durations up to 20 seconds. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

openai/sora-2-pro/image-to-video
image-to-video$1.2000

openai/sora-2-pro/image-to-video

OpenAI Sora 2 Pro Image-to-Video creates physics-aware, realistic videos from reference images with synchronized audio and strong steerability. Supports 720p and 1080p resolutions with durations up to 20 seconds. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

openai/sora-2/image-to-video
image-to-video$0.4000

openai/sora-2/image-to-video

OpenAI Sora 2 generates realistic image-to-video content with synchronized audio, improved physics, sharper realism and steerability. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

openai/gpt-image-1.5/edit
image-to-image$0.1000

openai/gpt-image-1.5/edit

GPT Image 1.5 Edit is OpenAI’s image model for precise, natural-language edits. Add/remove objects, swap backgrounds, retouch faces, adjust colors/lighting, edit text/graphics, crop/resize, and apply hex color control. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

openai/gpt-image-1.5/text-to-image
text-to-image$0.0400

openai/gpt-image-1.5/text-to-image

GPT Image 1.5 text to image is OpenAI’s fast, cost-efficient text-to-image generator powered by GPT-5 guidance. Create photorealistic shots, product renders, concept art, and stylized graphics from natural-language prompts (optionally conditioned with an image). Supports custom aspect ratios, seeds, negative prompts, hex color hints, and style presets. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

wavespeed-ai/openai-whisper-turbo
speech-to-text$0.0007

wavespeed-ai/openai-whisper-turbo

Accurate speech-to-text with OpenAI Whisper Large v3 Turbo: multilingual transcripts with auto language detection and punctuation. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

openai/sora-2/text-to-video-pro
text-to-video$1.2000

openai/sora-2/text-to-video-pro

OpenAI Sora 2 Text-to-Video Pro creates high-fidelity videos with synchronized audio, realistic physics, and enhanced steerability. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

openai/sora-2/text-to-video
text-to-video$0.4000

openai/sora-2/text-to-video

OpenAI Sora 2 is a state-of-the-art text-to-video model with realistic visuals, accurate physics, synchronized audio, and strong steerability. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

openai/sora-2/image-to-video-pro
image-to-video$1.2000

openai/sora-2/image-to-video-pro

OpenAI Sora 2 Image-to-Video Pro creates physics-aware, realistic videos with synchronized audio and greater steerability. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

wavespeed-ai/openai-whisper
speech-to-text$0.0010

wavespeed-ai/openai-whisper

Whisper Large v3 speech-to-text: instant, accurate multilingual transcripts with automatic language detection and punctuation. Upload audio to get transcripts. Ready-to-use REST API, no coldstarts, affordable pricing.

OpenAI Models

OpenAI Models on WaveSpeedAI bring together advanced image, video, and speech AI models for creative and production workflows. The collection is centered on GPT Image 2, OpenAI’s latest image generation and editing model family, while also including Sora video generation, GPT Image 1.5, GPT Image 1 Mini, DALL·E models, and Whisper speech recognition.

Built for creators, developers, designers, marketers, and AI applications, OpenAI Models support high-quality text-to-image generation, natural-language image editing, image-to-video generation, text-to-video creation, transcription, and multimodal creative workflows. The suite is suitable for marketing visuals, product concepts, UI mockups, social media assets, campaign creatives, cinematic videos, and scalable content production.

Core Model Capabilities

GPT Image 2 — Flagship Image Generation and Editing:

GPT Image 2 is the main highlight of this collection, delivering high-quality image generation and editing with strong prompt understanding, clean composition, polished aesthetics, and improved visual coherence. It is designed for professional creative workflows that need reliable results, detailed visual control, and production-ready output.

Text-to-Image Generation:

Generate high-quality images from natural-language prompts for campaign assets, UI concepts, product visuals, concept art, social media content, brand creatives, and rapid visual ideation.

Natural-Language Image Editing:

Use GPT Image 2 Edit to modify images with text instructions and reference inputs while preserving visual consistency, style coherence, composition, and fine details. It is useful for marketing asset refinement, product image editing, design iteration, and creative retouching.

GPT Image 1.5 and GPT Image 1 Mini:

Use GPT Image 1.5 and GPT Image 1 Mini for cost-efficient image generation, fast creative iteration, lightweight image editing, and scalable visual production workflows.

Sora Video Generation:

Use Sora and Sora 2 models for image-to-video and text-to-video workflows, turning prompts or still images into cinematic video clips with coherent motion, stable identities, and smooth camera movement.

DALL·E Image Models:

DALL·E models provide additional text-to-image options for illustration, concept exploration, quick drafts, and stylized image creation.

Whisper Speech Recognition:

Whisper and Whisper Turbo provide multilingual speech recognition for transcription, automatic language detection, punctuation, and large-scale audio processing workflows.

Text-to-3D Generation:

Use GPT-6 Astra Text-to-3D to generate 3D assets directly from natural-language prompts, making it easier to move from concept to usable 3D content for design, visualization, and creative production.

Image-to-3D Generation:

Use GPT-6 Astra Image-to-3D to convert reference images into 3D assets, helping creators build 3D objects from visual inputs while preserving shape cues and overall appearance.

OpenAI Models on WaveSpeedAI give creators and developers fast access to OpenAI’s image, video, and speech models through scalable APIs, flexible pricing, and production-ready creative capabilities.

API OpenAI Models — harga & performa

Jalankan model apa pun di koleksi OpenAI Models melalui satu REST API. Bayar per generasi — tanpa langganan, tanpa minimum — dengan latensi terdepan di infrastruktur dengan uptime 99,9%.

Mengapa menjalankan OpenAI Models di WaveSpeedAI

Harga transparan

Harga per panggilan untuk setiap model OpenAI Models. Harga tercantum di halaman setiap model — tanpa biaya platform tambahan.

Dioptimalkan untuk latensi rendah

Sebagian besar model gambar OpenAI Models selesai di bawah 2 detik. Model video dan 3D beberapa kali lebih cepat daripada alternatif yang di-hosting sendiri.

Uptime 99,9%

Failover multi-region dan retry otomatis menjaga lalu lintas produksi tetap online — bahkan saat provider mengalami gangguan.

Pertanyaan yang sering diajukan

Berapa biaya API OpenAI Models?+

Setiap model memiliki harga per panggilan tersendiri yang tercantum di halaman model. Kami menagih per generasi berhasil, tanpa biaya langganan atau minimum.

Seberapa cepat model OpenAI Models di WaveSpeedAI?+

Model gambar di koleksi ini biasanya selesai di bawah 2 detik. Model video dan 3D bergantung pada durasi dan resolusi, tetapi biasanya beberapa kali lebih cepat dari run yang di-hosting sendiri.

Bisakah saya mencoba API tanpa kartu kredit?+

Akun baru yang memenuhi syarat dapat menerima kredit promosi $1 untuk mencoba model OpenAI Models tanpa kartu kredit. Kredit uji coba tidak dijamin untuk setiap pendaftaran; periksa saldo akun sebelum membuat konten.

Apakah ada rate limit?+

Akun standar memiliki batas concurrent job yang murah hati. Paket Enterprise menawarkan RPM khusus, concurrency lebih tinggi, dan kapasitas khusus — hubungi sales untuk detailnya.