Nano Banana 2.1 公開中 — Google 最新モデル | 今すぐ試す →
Qwen AI Models

Qwen AI Models

Qwen multimodal models for image and video generation

Qwen multimodal models for image and video generation

すべてのモデル

32 モデル
alibaba/qwen-image-3.0-pro/edit
image-to-image$0.0400

alibaba/qwen-image-3.0-pro/edit

Qwen Image 3.0 Pro Edit is a professional-grade image editing model that transforms existing images with natural-language instructions, delivering advanced instruction understanding, superior visual quality, and up to 2K output for creative and production workflows. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

alibaba/qwen-image-3.0/text-to-image
text-to-image$0.0300

alibaba/qwen-image-3.0/text-to-image

Qwen Image 3.0 Text to Image is a high-quality image generation model that creates high-quality images from text prompts, with advanced prompt understanding, strong visual quality, and up to 2K output for creative and production workflows. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

alibaba/qwen-image-3.0/edit
image-to-image$0.0300

alibaba/qwen-image-3.0/edit

Qwen Image 3.0 Edit is a high-quality image editing model that transforms existing images with natural-language instructions, delivering advanced instruction understanding, superior visual quality, and up to 2K output for creative and production workflows. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

alibaba/qwen-image-3.0-pro/text-to-image
text-to-image$0.0400

alibaba/qwen-image-3.0-pro/text-to-image

Qwen Image 3.0 Pro Text-to-Image is a professional-grade image generation model that creates high-quality images from text prompts, with advanced prompt understanding, strong visual quality, and up to 2K output for creative and production workflows. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

wavespeed-ai/qwen-image-2512-lora-trainer
training$1.0000

wavespeed-ai/qwen-image-2512-lora-trainer

Qwen-Image-2512 LoRA Trainer lets you train custom LoRA models 10x faster with style, character, and object training. From concept to model in minutes, not hours—upload a ZIP file containing images to start. Ready-to-use REST inference API, best performance, no cold starts, affordable pricing.

wavespeed-ai/qwen-image/edit-multiple-angles
image-to-image$0.0250

wavespeed-ai/qwen-image/edit-multiple-angles

Generate specific camera angles from a single image using a 96-pose camera system. Control horizontal rotation, vertical tilt, and zoom to create front, side, back views and more. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

wavespeed-ai/qwen-image/edit-plus-lora
image-to-image$0.0250

wavespeed-ai/qwen-image/edit-plus-lora

Qwen-Image-Edit-Plus (2509) is 20B MMDiT image-to-image editor supporting multi-image edits, single-image consistency, and native ControlNet. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

alibaba/qwen-image/translate
image-to-image$0.0120

alibaba/qwen-image/translate

Qwen Vision Translate offers OCR-based image understanding and multilingual in-image text translation for context-aware results. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

alibaba/qwen3-tts-flash
text-to-speech$0.0500

alibaba/qwen3-tts-flash

Qwen3 TTS Flash: Low-latency Text-to-Speech for English and Chinese with multiple voices, ideal for real-time dialogue. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

wavespeed-ai/jib-mix-qwen-image/text-to-image
text-to-image$0.0200

wavespeed-ai/jib-mix-qwen-image/text-to-image

Jib Mix Qwen is a next-gen Text-to-Image model optimized for producing natural, pretty faces with improved Asian facial rendering. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

wavespeed-ai/jib-mix-qwen-image/text-to-image-lora
text-to-image$0.0250

wavespeed-ai/jib-mix-qwen-image/text-to-image-lora

Jib Mix Qwen LoRA specializes in producing more natural, attractive faces and is particularly strong at rendering Asian facial features for next-gen text-to-image generation with LoRA support. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

wavespeed-ai/qwen-image/layered
image-to-image$0.0250

wavespeed-ai/qwen-image/layered

Qwen-Image Layered is a unified image-layer decomposition model for prompt-guided compositing. Provide points, boxes, or rough masks to isolate subjects and regions, and the model splits a single image into multiple RGBA layers with clean alpha, soft edges, and correct occlusion order. Ready-to-use REST inference API with fast response, no cold starts, and affordable pricing.

wavespeed-ai/qwen-image/edit-2511
image-to-image$0.0200

wavespeed-ai/qwen-image/edit-2511

Qwen Image Edit 2511 is a major upgrade over 2509 for real-world image editing and design. It delivers stronger edit consistency, robust multi-person identity/pose consistency, built-in LoRA styles, enhanced industrial/product design, and improved geometric reasoning for structure-preserving edits. Built for stable production use with a ready-to-use REST API, no cold starts, and predictable pricing.

wavespeed-ai/qwen-image/edit-2511-lora
image-to-image$0.0250

wavespeed-ai/qwen-image/edit-2511-lora

Qwen Image Edit 2511 LoRA is an enhanced version with custom LoRA support for personalized styles. It delivers stronger edit consistency, robust multi-person identity/pose consistency, custom LoRA styles, enhanced industrial/product design, and improved geometric reasoning for structure-preserving edits. Built for stable production use with a ready-to-use REST API, no cold starts, and predictable pricing.

wavespeed-ai/qwen-image/text-to-image-2512
text-to-image$0.0200

wavespeed-ai/qwen-image/text-to-image-2512

Qwen Image 2512 is Qwen's latest text-to-image model with enhanced prompt understanding, superior text rendering, and versatile aspect ratio support. Ready-to-use REST inference API, best performance, no cold starts, affordable pricing.

wavespeed-ai/qwen-image/text-to-image-2512-lora
text-to-image$0.0250

wavespeed-ai/qwen-image/text-to-image-2512-lora

Qwen-Image-2512 LoRA is an enhanced 20B MMDiT text-to-image model with LoRA support for fast customization and refined image generation. Ready-to-use REST inference API, best performance, no cold starts, affordable pricing.

wavespeed-ai/qwen-image/text-to-image
text-to-image$0.0200

wavespeed-ai/qwen-image/text-to-image

Qwen-Image is a 20B MMDiT next-gen text-to-image model that generates images from text prompts. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

wavespeed-ai/qwen-image/edit-plus
image-to-image$0.0200

wavespeed-ai/qwen-image/edit-plus

Qwen-Image-Edit-Plus (2509) is a 20B MMDiT image editor with multi-image editing, single-image consistency and native ControlNet support. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

wavespeed-ai/qwen3-tts/text-to-speech
text-to-speech$0.0050

wavespeed-ai/qwen3-tts/text-to-speech

Qwen3 TTS: Multi-language, multi-voice text-to-speech synthesis with style control. Supports 11 languages and 9 voice characters. Ready-to-use REST inference API, best performance, no cold starts, affordable pricing.

wavespeed-ai/qwen3-tts/voice-clone
text-to-speech$0.0050

wavespeed-ai/qwen3-tts/voice-clone

Qwen3 TTS Voice Clone: Clone any voice from a reference audio and generate speech in that voice. Ready-to-use REST inference API, best performance, no cold starts, affordable pricing.

wavespeed-ai/qwen3-tts/voice-design
text-to-speech$0.0050

wavespeed-ai/qwen3-tts/voice-design

Qwen3 TTS Voice Design: Generate speech with custom voice characteristics described in natural language. Ready-to-use REST inference API, best performance, no cold starts, affordable pricing.

wavespeed-ai/qwen-image-max/text-to-image
text-to-image$0.0700

wavespeed-ai/qwen-image-max/text-to-image

Qwen Image Max is a text-to-image model with high-quality image generation supporting Chinese and English prompts. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

wavespeed-ai/qwen-image-max/edit
image-to-image$0.0700

wavespeed-ai/qwen-image-max/edit

Qwen Image Max Edit is an AI model for image editing with text prompts, supporting both Chinese and English languages. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

wavespeed-ai/qwen-image/edit-2509-multiple-angles
image-to-image$0.0250

wavespeed-ai/qwen-image/edit-2509-multiple-angles

Qwen Image Edit 2509 Multiple Angles is an AI image editing model that generates multiple-angle views of objects or scenes from a single image. Transform perspectives and create diverse viewpoints with text prompts. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

wavespeed-ai/qwen-image-2.0-pro/text-to-image
text-to-image$0.0700

wavespeed-ai/qwen-image-2.0-pro/text-to-image

Qwen Image 2.0 Pro is a professional-grade text-to-image model with superior quality and advanced prompt understanding. Up to 2k. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

wavespeed-ai/qwen-image-2.0/text-to-image
text-to-image$0.0300

wavespeed-ai/qwen-image-2.0/text-to-image

Qwen Image 2.0 is an advanced text-to-image model with enhanced image quality and improved prompt understanding. Up to 2k. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

wavespeed-ai/qwen-image-2.0-pro/edit
image-to-image$0.0700

wavespeed-ai/qwen-image-2.0-pro/edit

Qwen Image 2.0 Pro Edit is a professional-grade image editing model with superior quality and advanced instruction understanding. Up to 2k. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

wavespeed-ai/qwen-image-2.0/edit
image-to-image$0.0300

wavespeed-ai/qwen-image-2.0/edit

Qwen Image 2.0 Edit is an advanced image-editing model with improved quality and better understanding of instructions. Up to 2k. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

wavespeed-ai/qwen-image/edit-lora
image-to-image$0.0250

wavespeed-ai/qwen-image/edit-lora

Qwen-Image-Edit LoRA (20B) enables bilingual Chinese/English image-to-image editing with style preservation and semantic and appearance edits. Ready-to-use REST API, best performance, no coldstarts, affordable pricing.

wavespeed-ai/qwen-image/edit
image-to-image$0.0200

wavespeed-ai/qwen-image/edit

Qwen-Image-Edit is a 20B MMDiT image-to-image model offering precise bilingual (Chinese & English) text edits while preserving style. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

wavespeed-ai/qwen-image-lora-trainer
training$1.0000

wavespeed-ai/qwen-image-lora-trainer

Train custom Qwen-Image LoRA models 10x faster. Style training, character training, object training. From concept to model in minutes, not hours. Upload a ZIP file containing images to start!

wavespeed-ai/qwen-image/text-to-image-lora
text-to-image$0.0250

wavespeed-ai/qwen-image/text-to-image-lora

Qwen-Image LoRA is a 20B MMDiT next-gen text-to-image model with LoRA support for fast customization and refined image generation. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

Qwen AI Models

Qwen Image Models provide Alibaba’s advanced AI image generation and editing suite, with Qwen Image 3.0 as the core highlight. The collection supports high-quality text-to-image generation, natural-language image editing, pro-level visual production, LoRA-based customization, image translation, and related creative utilities.

Qwen Image 3.0 is the latest generation in the Qwen Image family, designed for professional creative workflows that require stronger instruction understanding, detailed rendering, coherent composition, multilingual prompt support, and up to 2K output. It includes both standard and pro tiers for text-to-image and image editing, giving creators and developers flexible options for everyday generation, advanced editing, and production-grade visual creation.

Core Model Capabilities

Qwen Image 3.0 Text-to-Image:

Generate high-quality images from natural-language prompts with strong instruction understanding, detailed visual rendering, coherent composition, and flexible creative control.

Qwen Image 3.0 Pro Text-to-Image:

Use the pro tier for higher-fidelity image generation, enhanced detail quality, stronger prompt adherence, superior visual output, and up to 2K resolution for professional workflows.

Qwen Image 3.0 Image Editing:

Edit and transform existing images with natural-language instructions while preserving subject identity, composition, lighting, and visual consistency.

Qwen Image 3.0 Pro Editing:

Apply more precise edits with stronger instruction following, higher-quality results, better detail preservation, and up to 2K output for demanding creative and commercial use cases.

Qwen Image 2.0 and Base Models:

Use earlier Qwen Image 2.0, Qwen Image Max, 2511/2512, and base models for additional generation and editing workflows, including fast image creation, product visuals, portraits, illustration, and photorealistic content production.

LoRA-Ready Generation and Editing:

Use Qwen Image LoRA models to support custom styles, consistent characters, personalized aesthetics, and trainable adapters for repeatable brand or identity-focused workflows.

Image Translation and Utility Workflows:

Use Qwen Image utilities for image translation, layout-aware text handling, and supporting creative tasks around image understanding and multilingual content adaptation.

Audio and Speech Support:

Qwen3 TTS models extend the collection with low-latency text-to-speech capabilities for natural voice previews, assistants, demos, and real-time applications.

Qwen Image Models on WaveSpeedAI give creators and developers fast access to Alibaba’s latest image generation and editing tools, centered on Qwen Image 3.0, with scalable APIs, flexible pricing, multilingual support, LoRA customization, and production-ready visual quality.

Qwen AI Models API — 料金とパフォーマンス

Qwen AI Models コレクションのすべてのモデルを単一の REST API で実行できます。生成ごとに課金 — サブスクなし、最低料金なし — で、稼働率 99.9% のインフラ上の業界トップクラスのレイテンシを提供します。

WaveSpeedAI で Qwen AI Models を使う理由

透明な料金体系

各 Qwen AI Models モデルにコールごとの料金が設定されています。料金は各モデルのページに表示され、プラットフォーム手数料はかかりません。

低レイテンシに最適化

ほとんどの Qwen AI Models 画像モデルは 2 秒以内に完了します。動画や 3D モデルはセルフホスト構成より数倍高速です。

稼働率 99.9%

マルチリージョンのフェイルオーバーと自動リトライで、プロバイダー障害時にも本番トラフィックを維持します。

よくある質問

Qwen AI Models API の料金はいくらですか?+

各モデルにはモデルページ上にコール単価が記載されています。成功した生成ごとに課金され、サブスクリプション料金や最低料金はありません。

WaveSpeedAI 上の Qwen AI Models モデルはどのくらい高速ですか?+

このコレクションの画像モデルは通常 2 秒以内に完了します。動画や 3D モデルは長さや解像度に依存しますが、セルフホスト実行より数倍高速なことが多いです。

クレジットカードなしで API を試せますか?+

条件を満たす新規アカウントは、クレジットカードなしで Qwen AI Models モデルを試すためのプロモーションクレジット 1 ドル分を受け取れる場合があります。すべての登録で付与されるわけではありません。生成前にアカウント残高をご確認ください。

レート制限はありますか?+

標準アカウントには十分な同時実行ジョブ枠があります。Enterprise プランではカスタム RPM、より高い同時実行性、専用キャパシティを提供します — 詳細は営業へお問い合わせください。