Seedance 2.5 è online | Provalo nel generatore video →
Qwen AI Models

Qwen AI Models

Qwen multimodal models for image and video generation

Qwen multimodal models for image and video generation

Tutti i modelli

32 modelli
alibaba/qwen-image-3.0-pro/edit
image-to-image

alibaba/qwen-image-3.0-pro/edit

Qwen Image 3.0 Pro Edit is a professional-grade image editing model that transforms existing images with natural-language instructions, delivering advanced instruction understanding, superior visual quality, and up to 2K output for creative and production workflows. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

alibaba/qwen-image-3.0/text-to-image
text-to-image

alibaba/qwen-image-3.0/text-to-image

Qwen Image 3.0 Text to Image is a high-quality image generation model that creates high-quality images from text prompts, with advanced prompt understanding, strong visual quality, and up to 2K output for creative and production workflows. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

alibaba/qwen-image-3.0/edit
image-to-image

alibaba/qwen-image-3.0/edit

Qwen Image 3.0 Edit is a high-quality image editing model that transforms existing images with natural-language instructions, delivering advanced instruction understanding, superior visual quality, and up to 2K output for creative and production workflows. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

alibaba/qwen-image-3.0-pro/text-to-image
text-to-image

alibaba/qwen-image-3.0-pro/text-to-image

Qwen Image 3.0 Pro Text-to-Image is a professional-grade image generation model that creates high-quality images from text prompts, with advanced prompt understanding, strong visual quality, and up to 2K output for creative and production workflows. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

wavespeed-ai/qwen-image-2512-lora-trainer
training

wavespeed-ai/qwen-image-2512-lora-trainer

Qwen-Image-2512 LoRA Trainer lets you train custom LoRA models 10x faster with style, character, and object training. From concept to model in minutes, not hours—upload a ZIP file containing images to start. Ready-to-use REST inference API, best performance, no cold starts, affordable pricing.

wavespeed-ai/qwen-image/edit-multiple-angles
image-to-image

wavespeed-ai/qwen-image/edit-multiple-angles

Generate specific camera angles from a single image using a 96-pose camera system. Control horizontal rotation, vertical tilt, and zoom to create front, side, back views and more. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

wavespeed-ai/qwen-image/edit-plus-lora
lora-support

wavespeed-ai/qwen-image/edit-plus-lora

Qwen-Image-Edit-Plus (2509) is 20B MMDiT image-to-image editor supporting multi-image edits, single-image consistency, and native ControlNet. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

alibaba/qwen-image/translate
image-to-image

alibaba/qwen-image/translate

Qwen Vision Translate offers OCR-based image understanding and multilingual in-image text translation for context-aware results. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

alibaba/qwen3-tts-flash
text-to-audio

alibaba/qwen3-tts-flash

Qwen3 TTS Flash: Low-latency Text-to-Speech for English and Chinese with multiple voices, ideal for real-time dialogue. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

wavespeed-ai/jib-mix-qwen-image/text-to-image
text-to-image

wavespeed-ai/jib-mix-qwen-image/text-to-image

Jib Mix Qwen is a next-gen Text-to-Image model optimized for producing natural, pretty faces with improved Asian facial rendering. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

wavespeed-ai/jib-mix-qwen-image/text-to-image-lora
lora-support

wavespeed-ai/jib-mix-qwen-image/text-to-image-lora

Jib Mix Qwen LoRA specializes in producing more natural, attractive faces and is particularly strong at rendering Asian facial features for next-gen text-to-image generation with LoRA support. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

wavespeed-ai/qwen-image/layered
image-to-image

wavespeed-ai/qwen-image/layered

Qwen-Image Layered is a unified image-layer decomposition model for prompt-guided compositing. Provide points, boxes, or rough masks to isolate subjects and regions, and the model splits a single image into multiple RGBA layers with clean alpha, soft edges, and correct occlusion order. Ready-to-use REST inference API with fast response, no cold starts, and affordable pricing.

wavespeed-ai/qwen-image/edit-2511
image-to-image

wavespeed-ai/qwen-image/edit-2511

Qwen Image Edit 2511 is a major upgrade over 2509 for real-world image editing and design. It delivers stronger edit consistency, robust multi-person identity/pose consistency, built-in LoRA styles, enhanced industrial/product design, and improved geometric reasoning for structure-preserving edits. Built for stable production use with a ready-to-use REST API, no cold starts, and predictable pricing.

wavespeed-ai/qwen-image/edit-2511-lora
lora-support

wavespeed-ai/qwen-image/edit-2511-lora

Qwen Image Edit 2511 LoRA is an enhanced version with custom LoRA support for personalized styles. It delivers stronger edit consistency, robust multi-person identity/pose consistency, custom LoRA styles, enhanced industrial/product design, and improved geometric reasoning for structure-preserving edits. Built for stable production use with a ready-to-use REST API, no cold starts, and predictable pricing.

wavespeed-ai/qwen-image/text-to-image-2512
text-to-image

wavespeed-ai/qwen-image/text-to-image-2512

Qwen Image 2512 is Qwen's latest text-to-image model with enhanced prompt understanding, superior text rendering, and versatile aspect ratio support. Ready-to-use REST inference API, best performance, no cold starts, affordable pricing.

wavespeed-ai/qwen-image/text-to-image-2512-lora
lora-support

wavespeed-ai/qwen-image/text-to-image-2512-lora

Qwen-Image-2512 LoRA is an enhanced 20B MMDiT text-to-image model with LoRA support for fast customization and refined image generation. Ready-to-use REST inference API, best performance, no cold starts, affordable pricing.

wavespeed-ai/qwen-image/text-to-image
text-to-image

wavespeed-ai/qwen-image/text-to-image

Qwen-Image is a 20B MMDiT next-gen text-to-image model that generates images from text prompts. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

wavespeed-ai/qwen-image/edit-plus
image-to-image

wavespeed-ai/qwen-image/edit-plus

Qwen-Image-Edit-Plus (2509) is a 20B MMDiT image editor with multi-image editing, single-image consistency and native ControlNet support. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

wavespeed-ai/qwen3-tts/text-to-speech
text-to-audio

wavespeed-ai/qwen3-tts/text-to-speech

Qwen3 TTS: Multi-language, multi-voice text-to-speech synthesis with style control. Supports 11 languages and 9 voice characters. Ready-to-use REST inference API, best performance, no cold starts, affordable pricing.

wavespeed-ai/qwen3-tts/voice-clone
audio-to-audio

wavespeed-ai/qwen3-tts/voice-clone

Qwen3 TTS Voice Clone: Clone any voice from a reference audio and generate speech in that voice. Ready-to-use REST inference API, best performance, no cold starts, affordable pricing.

wavespeed-ai/qwen3-tts/voice-design
text-to-audio

wavespeed-ai/qwen3-tts/voice-design

Qwen3 TTS Voice Design: Generate speech with custom voice characteristics described in natural language. Ready-to-use REST inference API, best performance, no cold starts, affordable pricing.

wavespeed-ai/qwen-image-max/text-to-image
text-to-image

wavespeed-ai/qwen-image-max/text-to-image

Qwen Image Max is a text-to-image model with high-quality image generation supporting Chinese and English prompts. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

wavespeed-ai/qwen-image-max/edit
image-to-image

wavespeed-ai/qwen-image-max/edit

Qwen Image Max Edit is an AI model for image editing with text prompts, supporting both Chinese and English languages. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

wavespeed-ai/qwen-image/edit-2509-multiple-angles
image-to-image

wavespeed-ai/qwen-image/edit-2509-multiple-angles

Qwen Image Edit 2509 Multiple Angles is an AI image editing model that generates multiple-angle views of objects or scenes from a single image. Transform perspectives and create diverse viewpoints with text prompts. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

wavespeed-ai/qwen-image-2.0-pro/text-to-image
text-to-image

wavespeed-ai/qwen-image-2.0-pro/text-to-image

Qwen Image 2.0 Pro is a professional-grade text-to-image model with superior quality and advanced prompt understanding. Up to 2k. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

wavespeed-ai/qwen-image-2.0/text-to-image
text-to-image

wavespeed-ai/qwen-image-2.0/text-to-image

Qwen Image 2.0 is an advanced text-to-image model with enhanced image quality and improved prompt understanding. Up to 2k. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

wavespeed-ai/qwen-image-2.0-pro/edit
image-to-image

wavespeed-ai/qwen-image-2.0-pro/edit

Qwen Image 2.0 Pro Edit is a professional-grade image editing model with superior quality and advanced instruction understanding. Up to 2k. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

wavespeed-ai/qwen-image-2.0/edit
image-to-image

wavespeed-ai/qwen-image-2.0/edit

Qwen Image 2.0 Edit is an advanced image-editing model with improved quality and better understanding of instructions. Up to 2k. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

wavespeed-ai/qwen-image/edit-lora
lora-support

wavespeed-ai/qwen-image/edit-lora

Qwen-Image-Edit LoRA (20B) enables bilingual Chinese/English image-to-image editing with style preservation and semantic and appearance edits. Ready-to-use REST API, best performance, no coldstarts, affordable pricing.

wavespeed-ai/qwen-image/edit
image-to-image

wavespeed-ai/qwen-image/edit

Qwen-Image-Edit is a 20B MMDiT image-to-image model offering precise bilingual (Chinese & English) text edits while preserving style. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

wavespeed-ai/qwen-image-lora-trainer
training

wavespeed-ai/qwen-image-lora-trainer

Train custom Qwen-Image LoRA models 10x faster. Style training, character training, object training. From concept to model in minutes, not hours. Upload a ZIP file containing images to start!

wavespeed-ai/qwen-image/text-to-image-lora
lora-support

wavespeed-ai/qwen-image/text-to-image-lora

Qwen-Image LoRA is a 20B MMDiT next-gen text-to-image model with LoRA support for fast customization and refined image generation. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

Qwen AI Models

Qwen multimodal models developed by Alibaba Cloud offer advanced capabilities in image and video generation. These models excel at creating high-quality visual content from text descriptions with a strong understanding of both Chinese and English prompts.

Qwen Image 3.0 — Standard & Pro Generation and Editing

  • alibaba/qwen-image-3.0/text-to-image: High-quality text-to-image generation model with strong instruction understanding, detailed rendering, coherent compositions, and flexible creative control.
  • alibaba/qwen-image-3.0-pro/text-to-image: Professional-grade text-to-image model with enhanced detail, advanced prompt understanding, superior visual quality, and up to 2K output.
  • alibaba/qwen-image-3.0/edit: High-quality image editing model for transforming existing images with natural-language instructions while preserving visual consistency and subject identity.
  • alibaba/qwen-image-3.0-pro/edit: Advanced image editing model with stronger instruction understanding, superior visual quality, precise control, and up to 2K output for professional workflows.

Qwen Image 2.0 — Standard & Pro Generation and Editing

  • wavespeed-ai/qwen-image-2.0/text-to-image: Fast, high-quality text-to-image model with strong prompt fidelity, detailed rendering, and balanced performance for everyday creative tasks.
  • wavespeed-ai/qwen-image-2.0-pro/text-to-image: Premium text-to-image model with enhanced detail, superior aesthetic quality, and finer control over complex multi-subject compositions.
  • wavespeed-ai/qwen-image-2.0/edit: Intelligent image editing model for quick modifications, style adjustments, targeted content changes, and everyday creative workflows.
  • wavespeed-ai/qwen-image-2.0-pro/edit: Advanced image editing model with higher precision, better context awareness, and production-grade output for professional retouching and creative transformation.

LoRA-ready Image Editing & Generation

  • qwen-image/edit-plus-lora:  Advanced image editing model with LoRA support, enabling precise style transfer, character customization, and high-fidelity local edits driven by text prompts.
  • qwen-image/edit-lora: Lightweight edit model for LoRA-based style and character control, ideal for quick retouching, outfit changes, and consistent persona updates.
  • qwen-image/text-to-image-lora: LoRA-enabled text-to-image generation that supports custom styles and characters while keeping strong prompt adherence and clean composition.
  • jib-mix-qwen-image/text-to-image-lora: Mixed-style LoRA T2I model tuned for vivid anime and illustration aesthetics, combining sharp linework with rich color and expressive characters.
  • qwen-image-lora-trainer: Training endpoint for building your own Qwen Image LoRA adapters from reference images, enabling personalized styles and characters across all LoRA-capable Qwen models.

Base Image Editing

  • qwen-image/edit-plus: Enhanced image editing model for high-quality global and local edits, improving lighting, realism, and detail while preserving subject identity.
  • qwen-image/edit: General-purpose edit model for everyday photo and artwork adjustments—ideal for quick fixes, background tweaks, and light retouching.
  • qwen-image/edit-2511: High-consistency image editing model for reliable multi-subject, identity-preserving edits, delivering reduced drift, stronger geometric control, and cleaner, product-grade results for iterative, production workflows.
  • qwen-image/edit-2511-edit-lora: LoRA-enhanced editing model built on the 2511 backbone—enables style injection, character customization, and fine-tuned aesthetic control while preserving the core stability of production-grade edits.
  • qwen-image-max/edit: Advanced image editing model offering precise object manipulation, seamless background replacement, and intelligent style transfer, while preserving high-fidelity details and natural lighting.

Base Text-to-Image Generation

  • qwen-image/text-to-image: Core T2I model that generates clean, realistic images from text prompts, suitable for product shots, portraits, and general creative use.
  • jib-mix-qwen-image/text-to-image: Stylized T2I variant blending anime and illustration styles, producing vibrant, character-focused art with strong visual appeal.
  • qwen-image/text-to-image-2512: Next-generation text-to-image model with enhanced prompt adherence, refined detail rendering, and improved compositional accuracy—engineered for photorealistic outputs and complex multi-element scene generation.
  • qwen-image-max/text-to-image: Premium text-to-image model delivering exceptional detail, superior photorealism, and complex scene coherence. Designed for professional-grade generation with advanced lighting, texture rendering, and precise compositional control.

Utilities & Audio

  • qwen-image/translate: Image translation utility that reads charts, UI screenshots, and text-heavy graphics, then outputs translated content while preserving layout semantics.
  • qwen3-tts family: Fast text-to-speech model for natural-sounding voice previews, optimized for low latency in assistants, demos, and real-time applications.

API Qwen AI Models — prezzi e prestazioni

Esegui qualsiasi modello della collezione Qwen AI Models tramite una singola API REST. Paga a generazione — senza abbonamenti né minimi — con latenza ai vertici del settore su un'infrastruttura con uptime del 99,9%.

Perché eseguire Qwen AI Models su WaveSpeedAI

Prezzi trasparenti

Prezzo per chiamata per ogni modello Qwen AI Models. Il prezzo è indicato nella pagina di ogni modello — senza costi di piattaforma aggiuntivi.

Ottimizzato per bassa latenza

La maggior parte dei modelli immagine Qwen AI Models si completa in meno di 2 secondi. I modelli video e 3D sono diverse volte più veloci delle alternative self-hosted.

Uptime 99,9%

Failover multi-regione e tentativi automatici tengono online il tuo traffico di produzione — anche durante interruzioni del provider.

Domande frequenti

Quanto costa l'API di Qwen AI Models?+

Ogni modello ha il proprio prezzo per chiamata indicato nella pagina del modello. Fatturiamo per generazione riuscita, senza abbonamenti né minimi.

Quanto sono veloci i modelli Qwen AI Models su WaveSpeedAI?+

I modelli immagine di questa collezione tipicamente si completano in meno di 2 secondi. I modelli video e 3D dipendono da durata e risoluzione, ma sono di solito diverse volte più veloci delle esecuzioni self-hosted.

Posso provare l'API senza carta di credito?+

Sì — ogni account riceve $1 di crediti gratuiti alla registrazione, sufficienti per provare la maggior parte dei modelli Qwen AI Models senza carta di credito.

Ci sono limiti di velocità?+

Gli account standard hanno limiti generosi di job concorrenti. I piani Enterprise offrono RPM personalizzato, concurrency più alta e capacità dedicata — contatta il commerciale per i dettagli.