GPT Image 2.5 ya está DISPONIBLE — Flare y Sunburst | Pruébalo en el generador de imágenes →
OpenAI Models

OpenAI Models

OpenAI's state-of-the-art AI models for text, image, and multimodal applications, Sora 2 is included

OpenAI's state-of-the-art AI models for text, image, and multimodal applications, Sora 2 is included

Todos los modelos

15 modelos
openai/gpt-image-2/edit5% OFF
image-to-image$0.0700$0.0665

openai/gpt-image-2/edit

OpenAI's GPT Image 2 Edit enables image editing from natural-language instructions with one or more reference images. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

openai/gpt-image-2/text-to-image5% OFF
text-to-image$0.0600$0.0570

openai/gpt-image-2/text-to-image

OpenAI's GPT Image 2 Text-to-Image generates high-quality images from natural-language prompts. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

openai/gpt-6-astra/image-to-3d
image-to-3d$8.0000

openai/gpt-6-astra/image-to-3d

GPT-6 Astra Image-to-3D generates textured 3D models or scenes from reference images, with optional text guidance for image-guided 3D asset creation, product visualization, concept design, virtual scenes, and production workflows. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

openai/gpt-6-astra/text-to-3d
text-to-3d$8.0000

openai/gpt-6-astra/text-to-3d

GPT-6 Astra Text-to-3D generates textured 3D models or scenes from text descriptions, supporting fast 3D asset creation for game assets, product visualization, concept design, virtual scenes, and production workflows. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

openai/gpt-image-1.5/text-to-image
text-to-image$0.0400

openai/gpt-image-1.5/text-to-image

GPT Image 1.5 text to image is OpenAI’s fast, cost-efficient text-to-image generator powered by GPT-5 guidance. Create photorealistic shots, product renders, concept art, and stylized graphics from natural-language prompts (optionally conditioned with an image). Supports custom aspect ratios, seeds, negative prompts, hex color hints, and style presets. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

wavespeed-ai/openai-whisper-turbo
speech-to-text$0.0007

wavespeed-ai/openai-whisper-turbo

Accurate speech-to-text with OpenAI Whisper Large v3 Turbo: multilingual transcripts with auto language detection and punctuation. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

openai/sora-2/image-to-video
image-to-video$0.4000

openai/sora-2/image-to-video

OpenAI Sora 2 generates realistic image-to-video content with synchronized audio, improved physics, sharper realism and steerability. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

openai/gpt-image-1.5/edit
image-to-image$0.1000

openai/gpt-image-1.5/edit

GPT Image 1.5 Edit is OpenAI’s image model for precise, natural-language edits. Add/remove objects, swap backgrounds, retouch faces, adjust colors/lighting, edit text/graphics, crop/resize, and apply hex color control. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

wavespeed-ai/openai-whisper-with-video
speech-to-text$0.0010

wavespeed-ai/openai-whisper-with-video

OpenAI Whisper Large v3 (Video-to-Text) delivers high-accuracy multilingual transcription directly from video files, with automatic language detection and optional timestamped, subtitle-ready segments. Built for stable production use with a ready-to-use REST API, fast response, no cold starts, and predictable pricing.

openai/sora-2-pro/text-to-video
text-to-video$1.2000

openai/sora-2-pro/text-to-video

OpenAI Sora 2 Pro is a state-of-the-art text-to-video model with realistic physics, synchronized audio, and strong steerability. Supports multiple resolutions up to 1080p and durations up to 20 seconds. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

openai/sora-2-pro/image-to-video
image-to-video$1.2000

openai/sora-2-pro/image-to-video

OpenAI Sora 2 Pro Image-to-Video creates physics-aware, realistic videos from reference images with synchronized audio and strong steerability. Supports 720p and 1080p resolutions with durations up to 20 seconds. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

wavespeed-ai/openai-whisper
speech-to-text$0.0010

wavespeed-ai/openai-whisper

Whisper Large v3 speech-to-text: instant, accurate multilingual transcripts with automatic language detection and punctuation. Upload audio to get transcripts. Ready-to-use REST API, no coldstarts, affordable pricing.

openai/sora-2/text-to-video-pro
text-to-video$1.2000

openai/sora-2/text-to-video-pro

OpenAI Sora 2 Text-to-Video Pro creates high-fidelity videos with synchronized audio, realistic physics, and enhanced steerability. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

openai/sora-2/text-to-video
text-to-video$0.4000

openai/sora-2/text-to-video

OpenAI Sora 2 is a state-of-the-art text-to-video model with realistic visuals, accurate physics, synchronized audio, and strong steerability. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

openai/sora-2/image-to-video-pro
image-to-video$1.2000

openai/sora-2/image-to-video-pro

OpenAI Sora 2 Image-to-Video Pro creates physics-aware, realistic videos with synchronized audio and greater steerability. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

OpenAI Models

OpenAI Models on WaveSpeedAI bring together advanced image, video, and speech AI models for creative and production workflows. The collection is centered on GPT Image 2, OpenAI’s latest image generation and editing model family, while also including Sora video generation, GPT Image 1.5, GPT Image 1 Mini, DALL·E models, and Whisper speech recognition.

Built for creators, developers, designers, marketers, and AI applications, OpenAI Models support high-quality text-to-image generation, natural-language image editing, image-to-video generation, text-to-video creation, transcription, and multimodal creative workflows. The suite is suitable for marketing visuals, product concepts, UI mockups, social media assets, campaign creatives, cinematic videos, and scalable content production.

Core Model Capabilities

GPT Image 2 — Flagship Image Generation and Editing:

GPT Image 2 is the main highlight of this collection, delivering high-quality image generation and editing with strong prompt understanding, clean composition, polished aesthetics, and improved visual coherence. It is designed for professional creative workflows that need reliable results, detailed visual control, and production-ready output.

Text-to-Image Generation:

Generate high-quality images from natural-language prompts for campaign assets, UI concepts, product visuals, concept art, social media content, brand creatives, and rapid visual ideation.

Natural-Language Image Editing:

Use GPT Image 2 Edit to modify images with text instructions and reference inputs while preserving visual consistency, style coherence, composition, and fine details. It is useful for marketing asset refinement, product image editing, design iteration, and creative retouching.

GPT Image 1.5 and GPT Image 1 Mini:

Use GPT Image 1.5 and GPT Image 1 Mini for cost-efficient image generation, fast creative iteration, lightweight image editing, and scalable visual production workflows.

Sora Video Generation:

Use Sora and Sora 2 models for image-to-video and text-to-video workflows, turning prompts or still images into cinematic video clips with coherent motion, stable identities, and smooth camera movement.

DALL·E Image Models:

DALL·E models provide additional text-to-image options for illustration, concept exploration, quick drafts, and stylized image creation.

Whisper Speech Recognition:

Whisper and Whisper Turbo provide multilingual speech recognition for transcription, automatic language detection, punctuation, and large-scale audio processing workflows.

Text-to-3D Generation:

Use GPT-6 Astra Text-to-3D to generate 3D assets directly from natural-language prompts, making it easier to move from concept to usable 3D content for design, visualization, and creative production.

Image-to-3D Generation:

Use GPT-6 Astra Image-to-3D to convert reference images into 3D assets, helping creators build 3D objects from visual inputs while preserving shape cues and overall appearance.

OpenAI Models on WaveSpeedAI give creators and developers fast access to OpenAI’s image, video, and speech models through scalable APIs, flexible pricing, and production-ready creative capabilities.

API de OpenAI Models — precios y rendimiento

Ejecuta cualquier modelo de la colección OpenAI Models a través de una sola API REST. Paga por generación — sin suscripciones ni mínimos — con latencia líder del sector sobre una infraestructura con 99,9 % de disponibilidad.

Por qué ejecutar OpenAI Models en WaveSpeedAI

Precios transparentes

Precio por llamada para cada modelo OpenAI Models. El precio aparece en la página de cada modelo — sin recargos de plataforma.

Optimizado para baja latencia

La mayoría de los modelos de imagen OpenAI Models terminan en menos de 2 segundos. Los modelos de vídeo y 3D son varias veces más rápidos que las alternativas autoalojadas.

99,9 % de disponibilidad

Conmutación por error multirregión y reintentos automáticos mantienen tu tráfico de producción en línea — incluso durante caídas del proveedor.

Preguntas frecuentes

¿Cuánto cuesta la API de OpenAI Models?+

Cada modelo tiene su propio precio por llamada listado en su página. Cobramos por generación exitosa, sin cuotas de suscripción ni mínimos.

¿Qué tan rápidos son los modelos OpenAI Models en WaveSpeedAI?+

Los modelos de imagen de esta colección suelen completarse en menos de 2 segundos. Los modelos de vídeo y 3D dependen de la duración y la resolución, pero suelen ser varias veces más rápidos que las ejecuciones autoalojadas.

¿Puedo probar la API sin tarjeta de crédito?+

Las cuentas nuevas que cumplan los requisitos pueden recibir $1 en créditos promocionales para probar modelos OpenAI Models sin tarjeta de crédito. No se garantizan créditos de prueba en cada registro; consulta tu saldo antes de generar.

¿Hay límites de tasa?+

Las cuentas estándar tienen límites generosos de trabajos concurrentes. Los planes Enterprise ofrecen RPM personalizado, mayor concurrencia y capacidad dedicada — contacta con ventas para más detalles.