Nano Banana 2.1 现已上线 — Google 最新模型 | 立即体验 →
OpenAI Models

OpenAI Models

OpenAI's state-of-the-art AI models for text, image, and multimodal applications, Sora 2 is included

OpenAI's state-of-the-art AI models for text, image, and multimodal applications, Sora 2 is included

所有模型

19 个模型
openai/gpt-image-2.5-sunburst/edit
image-to-image$0.0390

openai/gpt-image-2.5-sunburst/edit

OpenAI's GPT Image 2.5 Sunburst Edit edits one or more reference images from natural-language instructions, with five quality tiers up to 4K. Sunburst is the precision-focused GPT Image 2.5 tier that spends more time per image for extra fidelity on intricate detail. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

openai/gpt-image-2.5-sunburst/text-to-image
text-to-image$0.0240

openai/gpt-image-2.5-sunburst/text-to-image

OpenAI's GPT Image 2.5 Sunburst Text-to-Image generates high-quality images from natural-language prompts, with five quality tiers up to 4K. Sunburst is the precision-focused GPT Image 2.5 tier that spends more time per image for extra fidelity on intricate detail. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

openai/gpt-image-2.5-flare/edit
image-to-image$0.0390

openai/gpt-image-2.5-flare/edit

OpenAI's GPT Image 2.5 Flare Edit edits one or more reference images from natural-language instructions, with five quality tiers up to 4K. Flare is the fast, balanced GPT Image 2.5 tier for everyday generation at low latency. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

openai/gpt-image-2.5-flare/text-to-image
text-to-image$0.0240

openai/gpt-image-2.5-flare/text-to-image

OpenAI's GPT Image 2.5 Flare Text-to-Image generates high-quality images from natural-language prompts, with five quality tiers up to 4K. Flare is the fast, balanced GPT Image 2.5 tier for everyday generation at low latency. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

openai/gpt-image-2/edit5% OFF
image-to-image$0.0700$0.0665

openai/gpt-image-2/edit

OpenAI's GPT Image 2 Edit enables image editing from natural-language instructions with one or more reference images. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

openai/gpt-image-2/text-to-image5% OFF
text-to-image$0.0600$0.0570

openai/gpt-image-2/text-to-image

OpenAI's GPT Image 2 Text-to-Image generates high-quality images from natural-language prompts. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

openai/gpt-6-astra/image-to-3d
image-to-3d$8.0000

openai/gpt-6-astra/image-to-3d

GPT-6 Astra Image-to-3D generates textured 3D models or scenes from reference images, with optional text guidance for image-guided 3D asset creation, product visualization, concept design, virtual scenes, and production workflows. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

openai/gpt-6-astra/text-to-3d
text-to-3d$8.0000

openai/gpt-6-astra/text-to-3d

GPT-6 Astra Text-to-3D generates textured 3D models or scenes from text descriptions, supporting fast 3D asset creation for game assets, product visualization, concept design, virtual scenes, and production workflows. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

wavespeed-ai/openai-whisper-with-video
speech-to-text$0.0010

wavespeed-ai/openai-whisper-with-video

OpenAI Whisper Large v3 (Video-to-Text) delivers high-accuracy multilingual transcription directly from video files, with automatic language detection and optional timestamped, subtitle-ready segments. Built for stable production use with a ready-to-use REST API, fast response, no cold starts, and predictable pricing.

openai/sora-2-pro/text-to-video
text-to-video$1.2000

openai/sora-2-pro/text-to-video

OpenAI Sora 2 Pro is a state-of-the-art text-to-video model with realistic physics, synchronized audio, and strong steerability. Supports multiple resolutions up to 1080p and durations up to 20 seconds. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

openai/sora-2-pro/image-to-video
image-to-video$1.2000

openai/sora-2-pro/image-to-video

OpenAI Sora 2 Pro Image-to-Video creates physics-aware, realistic videos from reference images with synchronized audio and strong steerability. Supports 720p and 1080p resolutions with durations up to 20 seconds. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

openai/sora-2/image-to-video
image-to-video$0.4000

openai/sora-2/image-to-video

OpenAI Sora 2 generates realistic image-to-video content with synchronized audio, improved physics, sharper realism and steerability. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

openai/gpt-image-1.5/edit
image-to-image$0.1000

openai/gpt-image-1.5/edit

GPT Image 1.5 Edit is OpenAI’s image model for precise, natural-language edits. Add/remove objects, swap backgrounds, retouch faces, adjust colors/lighting, edit text/graphics, crop/resize, and apply hex color control. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

openai/gpt-image-1.5/text-to-image
text-to-image$0.0400

openai/gpt-image-1.5/text-to-image

GPT Image 1.5 text to image is OpenAI’s fast, cost-efficient text-to-image generator powered by GPT-5 guidance. Create photorealistic shots, product renders, concept art, and stylized graphics from natural-language prompts (optionally conditioned with an image). Supports custom aspect ratios, seeds, negative prompts, hex color hints, and style presets. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

wavespeed-ai/openai-whisper-turbo
speech-to-text$0.0007

wavespeed-ai/openai-whisper-turbo

Accurate speech-to-text with OpenAI Whisper Large v3 Turbo: multilingual transcripts with auto language detection and punctuation. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

openai/sora-2/text-to-video-pro
text-to-video$1.2000

openai/sora-2/text-to-video-pro

OpenAI Sora 2 Text-to-Video Pro creates high-fidelity videos with synchronized audio, realistic physics, and enhanced steerability. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

openai/sora-2/text-to-video
text-to-video$0.4000

openai/sora-2/text-to-video

OpenAI Sora 2 is a state-of-the-art text-to-video model with realistic visuals, accurate physics, synchronized audio, and strong steerability. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

openai/sora-2/image-to-video-pro
image-to-video$1.2000

openai/sora-2/image-to-video-pro

OpenAI Sora 2 Image-to-Video Pro creates physics-aware, realistic videos with synchronized audio and greater steerability. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

wavespeed-ai/openai-whisper
speech-to-text$0.0010

wavespeed-ai/openai-whisper

Whisper Large v3 speech-to-text: instant, accurate multilingual transcripts with automatic language detection and punctuation. Upload audio to get transcripts. Ready-to-use REST API, no coldstarts, affordable pricing.

OpenAI Models

OpenAI Models on WaveSpeedAI bring together advanced image, video, and speech AI models for creative and production workflows. The collection is centered on GPT Image 2, OpenAI’s latest image generation and editing model family, while also including Sora video generation, GPT Image 1.5, GPT Image 1 Mini, DALL·E models, and Whisper speech recognition.

Built for creators, developers, designers, marketers, and AI applications, OpenAI Models support high-quality text-to-image generation, natural-language image editing, image-to-video generation, text-to-video creation, transcription, and multimodal creative workflows. The suite is suitable for marketing visuals, product concepts, UI mockups, social media assets, campaign creatives, cinematic videos, and scalable content production.

Core Model Capabilities

GPT Image 2 — Flagship Image Generation and Editing:

GPT Image 2 is the main highlight of this collection, delivering high-quality image generation and editing with strong prompt understanding, clean composition, polished aesthetics, and improved visual coherence. It is designed for professional creative workflows that need reliable results, detailed visual control, and production-ready output.

Text-to-Image Generation:

Generate high-quality images from natural-language prompts for campaign assets, UI concepts, product visuals, concept art, social media content, brand creatives, and rapid visual ideation.

Natural-Language Image Editing:

Use GPT Image 2 Edit to modify images with text instructions and reference inputs while preserving visual consistency, style coherence, composition, and fine details. It is useful for marketing asset refinement, product image editing, design iteration, and creative retouching.

GPT Image 1.5 and GPT Image 1 Mini:

Use GPT Image 1.5 and GPT Image 1 Mini for cost-efficient image generation, fast creative iteration, lightweight image editing, and scalable visual production workflows.

Sora Video Generation:

Use Sora and Sora 2 models for image-to-video and text-to-video workflows, turning prompts or still images into cinematic video clips with coherent motion, stable identities, and smooth camera movement.

DALL·E Image Models:

DALL·E models provide additional text-to-image options for illustration, concept exploration, quick drafts, and stylized image creation.

Whisper Speech Recognition:

Whisper and Whisper Turbo provide multilingual speech recognition for transcription, automatic language detection, punctuation, and large-scale audio processing workflows.

Text-to-3D Generation:

Use GPT-6 Astra Text-to-3D to generate 3D assets directly from natural-language prompts, making it easier to move from concept to usable 3D content for design, visualization, and creative production.

Image-to-3D Generation:

Use GPT-6 Astra Image-to-3D to convert reference images into 3D assets, helping creators build 3D objects from visual inputs while preserving shape cues and overall appearance.

OpenAI Models on WaveSpeedAI give creators and developers fast access to OpenAI’s image, video, and speech models through scalable APIs, flexible pricing, and production-ready creative capabilities.

OpenAI Models API — 价格与性能

通过单一 REST API 运行 OpenAI Models 系列中的任意模型。按生成计费 — 无订阅、无最低消费 — 在 99.9% 可用性的基础设施上提供行业领先的延迟。

为什么在 WaveSpeedAI 上运行 OpenAI Models

透明定价

每个 OpenAI Models 模型都有按调用计价。价格在每个模型的页面上列出 — 不收取额外的平台费。

为低延迟优化

大多数 OpenAI Models 图像模型在 2 秒内完成。视频和 3D 模型比自托管方案快数倍。

99.9% 可用性

多区域故障转移和自动重试可确保您的生产流量保持在线 — 即使在供应商故障期间。

常见问题

OpenAI Models API 多少钱?+

每个模型在其模型页面上都列有自己的按调用价格。我们按每次成功生成计费,没有订阅费或最低消费。

OpenAI Models 模型在 WaveSpeedAI 上有多快?+

本系列中的图像模型通常在 2 秒内完成。视频和 3D 模型取决于时长和分辨率,但通常比自托管运行快数倍。

不用信用卡可以试用 API 吗?+

符合条件的新账户可能获得 $1 的推广额度,用于无需信用卡试用 OpenAI Models 模型。并非每次注册都会获得试用额度,请在生成前查看账户余额。

有速率限制吗?+

标准账户有充足的并发任务限制。企业版计划提供自定义 RPM、更高并发和专用容量 — 详情请联系销售。