Nano Banana 2.1 公開中 — Google 最新モデル | 今すぐ試す →
OpenAI Models

OpenAI Models

OpenAI's state-of-the-art AI models for text, image, and multimodal applications, Sora 2 is included

OpenAI's state-of-the-art AI models for text, image, and multimodal applications, Sora 2 is included

すべてのモデル

19 モデル
openai/gpt-image-2.5-sunburst/edit
image-to-image$0.0390

openai/gpt-image-2.5-sunburst/edit

OpenAI's GPT Image 2.5 Sunburst Edit edits one or more reference images from natural-language instructions, with five quality tiers up to 4K. Sunburst is the precision-focused GPT Image 2.5 tier that spends more time per image for extra fidelity on intricate detail. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

openai/gpt-image-2.5-sunburst/text-to-image
text-to-image$0.0240

openai/gpt-image-2.5-sunburst/text-to-image

OpenAI's GPT Image 2.5 Sunburst Text-to-Image generates high-quality images from natural-language prompts, with five quality tiers up to 4K. Sunburst is the precision-focused GPT Image 2.5 tier that spends more time per image for extra fidelity on intricate detail. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

openai/gpt-image-2.5-flare/edit
image-to-image$0.0390

openai/gpt-image-2.5-flare/edit

OpenAI's GPT Image 2.5 Flare Edit edits one or more reference images from natural-language instructions, with five quality tiers up to 4K. Flare is the fast, balanced GPT Image 2.5 tier for everyday generation at low latency. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

openai/gpt-image-2.5-flare/text-to-image
text-to-image$0.0240

openai/gpt-image-2.5-flare/text-to-image

OpenAI's GPT Image 2.5 Flare Text-to-Image generates high-quality images from natural-language prompts, with five quality tiers up to 4K. Flare is the fast, balanced GPT Image 2.5 tier for everyday generation at low latency. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

openai/gpt-image-2/edit5% OFF
image-to-image$0.0700$0.0665

openai/gpt-image-2/edit

OpenAI's GPT Image 2 Edit enables image editing from natural-language instructions with one or more reference images. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

openai/gpt-image-2/text-to-image5% OFF
text-to-image$0.0600$0.0570

openai/gpt-image-2/text-to-image

OpenAI's GPT Image 2 Text-to-Image generates high-quality images from natural-language prompts. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

openai/gpt-6-astra/image-to-3d
image-to-3d$8.0000

openai/gpt-6-astra/image-to-3d

GPT-6 Astra Image-to-3D generates textured 3D models or scenes from reference images, with optional text guidance for image-guided 3D asset creation, product visualization, concept design, virtual scenes, and production workflows. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

openai/gpt-6-astra/text-to-3d
text-to-3d$8.0000

openai/gpt-6-astra/text-to-3d

GPT-6 Astra Text-to-3D generates textured 3D models or scenes from text descriptions, supporting fast 3D asset creation for game assets, product visualization, concept design, virtual scenes, and production workflows. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

wavespeed-ai/openai-whisper-with-video
speech-to-text$0.0010

wavespeed-ai/openai-whisper-with-video

OpenAI Whisper Large v3 (Video-to-Text) delivers high-accuracy multilingual transcription directly from video files, with automatic language detection and optional timestamped, subtitle-ready segments. Built for stable production use with a ready-to-use REST API, fast response, no cold starts, and predictable pricing.

openai/sora-2-pro/text-to-video
text-to-video$1.2000

openai/sora-2-pro/text-to-video

OpenAI Sora 2 Pro is a state-of-the-art text-to-video model with realistic physics, synchronized audio, and strong steerability. Supports multiple resolutions up to 1080p and durations up to 20 seconds. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

openai/sora-2-pro/image-to-video
image-to-video$1.2000

openai/sora-2-pro/image-to-video

OpenAI Sora 2 Pro Image-to-Video creates physics-aware, realistic videos from reference images with synchronized audio and strong steerability. Supports 720p and 1080p resolutions with durations up to 20 seconds. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

openai/sora-2/image-to-video
image-to-video$0.4000

openai/sora-2/image-to-video

OpenAI Sora 2 generates realistic image-to-video content with synchronized audio, improved physics, sharper realism and steerability. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

openai/gpt-image-1.5/edit
image-to-image$0.1000

openai/gpt-image-1.5/edit

GPT Image 1.5 Edit is OpenAI’s image model for precise, natural-language edits. Add/remove objects, swap backgrounds, retouch faces, adjust colors/lighting, edit text/graphics, crop/resize, and apply hex color control. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

openai/gpt-image-1.5/text-to-image
text-to-image$0.0400

openai/gpt-image-1.5/text-to-image

GPT Image 1.5 text to image is OpenAI’s fast, cost-efficient text-to-image generator powered by GPT-5 guidance. Create photorealistic shots, product renders, concept art, and stylized graphics from natural-language prompts (optionally conditioned with an image). Supports custom aspect ratios, seeds, negative prompts, hex color hints, and style presets. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

wavespeed-ai/openai-whisper-turbo
speech-to-text$0.0007

wavespeed-ai/openai-whisper-turbo

Accurate speech-to-text with OpenAI Whisper Large v3 Turbo: multilingual transcripts with auto language detection and punctuation. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

openai/sora-2/text-to-video-pro
text-to-video$1.2000

openai/sora-2/text-to-video-pro

OpenAI Sora 2 Text-to-Video Pro creates high-fidelity videos with synchronized audio, realistic physics, and enhanced steerability. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

openai/sora-2/text-to-video
text-to-video$0.4000

openai/sora-2/text-to-video

OpenAI Sora 2 is a state-of-the-art text-to-video model with realistic visuals, accurate physics, synchronized audio, and strong steerability. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

openai/sora-2/image-to-video-pro
image-to-video$1.2000

openai/sora-2/image-to-video-pro

OpenAI Sora 2 Image-to-Video Pro creates physics-aware, realistic videos with synchronized audio and greater steerability. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

wavespeed-ai/openai-whisper
speech-to-text$0.0010

wavespeed-ai/openai-whisper

Whisper Large v3 speech-to-text: instant, accurate multilingual transcripts with automatic language detection and punctuation. Upload audio to get transcripts. Ready-to-use REST API, no coldstarts, affordable pricing.

OpenAI Models

OpenAI Models on WaveSpeedAI bring together advanced image, video, and speech AI models for creative and production workflows. The collection is centered on GPT Image 2, OpenAI’s latest image generation and editing model family, while also including Sora video generation, GPT Image 1.5, GPT Image 1 Mini, DALL·E models, and Whisper speech recognition.

Built for creators, developers, designers, marketers, and AI applications, OpenAI Models support high-quality text-to-image generation, natural-language image editing, image-to-video generation, text-to-video creation, transcription, and multimodal creative workflows. The suite is suitable for marketing visuals, product concepts, UI mockups, social media assets, campaign creatives, cinematic videos, and scalable content production.

Core Model Capabilities

GPT Image 2 — Flagship Image Generation and Editing:

GPT Image 2 is the main highlight of this collection, delivering high-quality image generation and editing with strong prompt understanding, clean composition, polished aesthetics, and improved visual coherence. It is designed for professional creative workflows that need reliable results, detailed visual control, and production-ready output.

Text-to-Image Generation:

Generate high-quality images from natural-language prompts for campaign assets, UI concepts, product visuals, concept art, social media content, brand creatives, and rapid visual ideation.

Natural-Language Image Editing:

Use GPT Image 2 Edit to modify images with text instructions and reference inputs while preserving visual consistency, style coherence, composition, and fine details. It is useful for marketing asset refinement, product image editing, design iteration, and creative retouching.

GPT Image 1.5 and GPT Image 1 Mini:

Use GPT Image 1.5 and GPT Image 1 Mini for cost-efficient image generation, fast creative iteration, lightweight image editing, and scalable visual production workflows.

Sora Video Generation:

Use Sora and Sora 2 models for image-to-video and text-to-video workflows, turning prompts or still images into cinematic video clips with coherent motion, stable identities, and smooth camera movement.

DALL·E Image Models:

DALL·E models provide additional text-to-image options for illustration, concept exploration, quick drafts, and stylized image creation.

Whisper Speech Recognition:

Whisper and Whisper Turbo provide multilingual speech recognition for transcription, automatic language detection, punctuation, and large-scale audio processing workflows.

Text-to-3D Generation:

Use GPT-6 Astra Text-to-3D to generate 3D assets directly from natural-language prompts, making it easier to move from concept to usable 3D content for design, visualization, and creative production.

Image-to-3D Generation:

Use GPT-6 Astra Image-to-3D to convert reference images into 3D assets, helping creators build 3D objects from visual inputs while preserving shape cues and overall appearance.

OpenAI Models on WaveSpeedAI give creators and developers fast access to OpenAI’s image, video, and speech models through scalable APIs, flexible pricing, and production-ready creative capabilities.

OpenAI Models API — 料金とパフォーマンス

OpenAI Models コレクションのすべてのモデルを単一の REST API で実行できます。生成ごとに課金 — サブスクなし、最低料金なし — で、稼働率 99.9% のインフラ上の業界トップクラスのレイテンシを提供します。

WaveSpeedAI で OpenAI Models を使う理由

透明な料金体系

各 OpenAI Models モデルにコールごとの料金が設定されています。料金は各モデルのページに表示され、プラットフォーム手数料はかかりません。

低レイテンシに最適化

ほとんどの OpenAI Models 画像モデルは 2 秒以内に完了します。動画や 3D モデルはセルフホスト構成より数倍高速です。

稼働率 99.9%

マルチリージョンのフェイルオーバーと自動リトライで、プロバイダー障害時にも本番トラフィックを維持します。

よくある質問

OpenAI Models API の料金はいくらですか?+

各モデルにはモデルページ上にコール単価が記載されています。成功した生成ごとに課金され、サブスクリプション料金や最低料金はありません。

WaveSpeedAI 上の OpenAI Models モデルはどのくらい高速ですか?+

このコレクションの画像モデルは通常 2 秒以内に完了します。動画や 3D モデルは長さや解像度に依存しますが、セルフホスト実行より数倍高速なことが多いです。

クレジットカードなしで API を試せますか?+

条件を満たす新規アカウントは、クレジットカードなしで OpenAI Models モデルを試すためのプロモーションクレジット 1 ドル分を受け取れる場合があります。すべての登録で付与されるわけではありません。生成前にアカウント残高をご確認ください。

レート制限はありますか?+

標準アカウントには十分な同時実行ジョブ枠があります。Enterprise プランではカスタム RPM、より高い同時実行性、専用キャパシティを提供します — 詳細は営業へお問い合わせください。