Nano Banana 2.1 RA MẮT — Mới nhất từ Google | Dùng thử →
AlibabaAPI ảnhTừ $0.02/lượt chạy

Qwen Image API

Alibaba Qwen-Image — bộ công cụ text-to-image và chỉnh sửa MMDiT 20B thế hệ mới với hỗ trợ song ngữ Trung/Anh, chỉnh sửa nhiều ảnh, tùy chỉnh LoRA, ghép lớp và hệ thống góc máy 96 tư thế.

Text-to-image ở các biến thể cơ bản, 2512 nâng cao và 2.0-pro. Các endpoint chỉnh sửa gồm Edit, Edit-Plus (nhiều ảnh, ControlNet), Edit-LoRA, Edit-Multiple-Angles (hệ thống máy quay 96 tư thế) và Layered (phân tách theo prompt). Các biến thể dòng Qwen Image 2.0 nằm chung tiền tố.

Kết quả mẫu của Qwen Image

Tổng quan

Giới thiệu về Qwen Image API

Qwen Image làm được gì, nằm ở đâu trong dòng mô hình của Alibaba, và vì sao các đội ngũ chọn nó.

Qwen Image là mô hình tạo và chỉnh sửa ảnh từ Alibaba, có sẵn qua REST API của WaveSpeedAI. Alibaba Qwen-Image — bộ công cụ text-to-image và chỉnh sửa MMDiT 20B thế hệ mới với hỗ trợ song ngữ Trung/Anh, chỉnh sửa nhiều ảnh, tùy chỉnh LoRA, ghép lớp và hệ thống góc máy 96 tư thế.

Text-to-image ở các biến thể cơ bản, 2512 nâng cao và 2.0-pro. Các endpoint chỉnh sửa gồm Edit, Edit-Plus (nhiều ảnh, ControlNet), Edit-LoRA, Edit-Multiple-Angles (hệ thống máy quay 96 tư thế) và Layered (phân tách theo prompt). Các biến thể dòng Qwen Image 2.0 nằm chung tiền tố.

Họ Qwen Image trên WaveSpeedAI cung cấp 21 endpoint REST bao gồm 3 quy trình Image-To-Image, Text-To-Image, Training. Mỗi biến thể có giá, các tham số điều chỉnh và kết quả mẫu riêng — hãy chọn biến thể phù hợp với dạng đầu vào và ràng buộc sản xuất của bạn, hoặc gọi nhiều biến thể từ cùng một API key để ghép các pipeline nhiều bước.

Chạy Qwen Image bằng cùng API key, tài khoản thanh toán và hạn mức tốc độ bạn dùng cho hơn 1.000 mô hình AI khác trên WaveSpeedAI. Không cần thiết lập nhà cung cấp riêng, không cần SDK cho từng nhà cung cấp, không có hạn mức tốc độ riêng của từng hãng — một lần tích hợp bao quát mọi thứ, từ text-to-image và text-to-video đến tổng hợp âm thanh, tạo 3D, nâng cấp và chỉnh sửa.

Endpoint

Tất cả endpoint API Qwen Image

21 endpoint Qwen Image hiện có sẵn trên WaveSpeedAI — hãy chọn biến thể phù hợp với quy trình của bạn.

image-to-imageQwen Image 2.0 Edit — bản xem trước image-to-image Qwen Image từ Alibaba

Qwen Image 2.0 Edit

Qwen Image 2.0 Edit is an advanced image-editing model with improved quality and better understanding of instructions. Up to 2k. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

từ $0.03
image-to-imageQwen Image 2.0 Pro Edit — bản xem trước image-to-image Qwen Image từ Alibaba

Qwen Image 2.0 Pro Edit

Qwen Image 2.0 Pro Edit is a professional-grade image editing model with superior quality and advanced instruction understanding. Up to 2k. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

từ $0.07
text-to-imageQwen Image 2.0 Text To Image — bản xem trước text-to-image Qwen Image từ Alibaba

Qwen Image 2.0 Text To Image

Qwen Image 2.0 is an advanced text-to-image model with enhanced image quality and improved prompt understanding. Up to 2k. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

từ $0.03
text-to-imageQwen Image 2.0 Pro Text To Image — bản xem trước text-to-image Qwen Image từ Alibaba

Qwen Image 2.0 Pro Text To Image

Qwen Image 2.0 Pro is a professional-grade text-to-image model with superior quality and advanced prompt understanding. Up to 2k. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

từ $0.07
image-to-imageQwen Image Edit 2509 Multiple Angles — bản xem trước image-to-image Qwen Image từ Alibaba

Qwen Image Edit 2509 Multiple Angles

Qwen Image Edit 2509 Multiple Angles is an AI image editing model that generates multiple-angle views of objects or scenes from a single image. Transform perspectives and create diverse viewpoints with text prompts. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

từ $0.025
image-to-imageQwen Image Max Edit — bản xem trước image-to-image Qwen Image từ Alibaba

Qwen Image Max Edit

Qwen Image Max Edit is an AI model for image editing with text prompts, supporting both Chinese and English languages. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

từ $0.07
text-to-imageQwen Image Max Text To Image — bản xem trước text-to-image Qwen Image từ Alibaba

Qwen Image Max Text To Image

Qwen Image Max is a text-to-image model with high-quality image generation supporting Chinese and English prompts. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

từ $0.07
image-to-imageQwen Image Edit Multiple Angles — bản xem trước image-to-image Qwen Image từ Alibaba

Qwen Image Edit Multiple Angles

Generate specific camera angles from a single image using a 96-pose camera system. Control horizontal rotation, vertical tilt, and zoom to create front, side, back views and more. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

từ $0.025
trainingQwen Image 2512 Lora Trainer — bản xem trước training Qwen Image từ Alibaba

Qwen Image 2512 Lora Trainer

Qwen-Image-2512 LoRA Trainer lets you train custom LoRA models 10x faster with style, character, and object training. From concept to model in minutes, not hours—upload a ZIP file containing images to start. Ready-to-use REST inference API, best performance, no cold starts, affordable pricing.

từ $1.00
text-to-imageQwen Image Text To Image 2512 Lora — bản xem trước text-to-image Qwen Image từ Alibaba

Qwen Image Text To Image 2512 Lora

Qwen-Image-2512 LoRA is an enhanced 20B MMDiT text-to-image model with LoRA support for fast customization and refined image generation. Ready-to-use REST inference API, best performance, no cold starts, affordable pricing.

từ $0.025
text-to-imageQwen Image Text To Image 2512 — bản xem trước text-to-image Qwen Image từ Alibaba

Qwen Image Text To Image 2512

Qwen Image 2512 is Qwen's latest text-to-image model with enhanced prompt understanding, superior text rendering, and versatile aspect ratio support. Ready-to-use REST inference API, best performance, no cold starts, affordable pricing.

từ $0.02
image-to-imageQwen Image Edit 2511 Lora — bản xem trước image-to-image Qwen Image từ Alibaba

Qwen Image Edit 2511 Lora

Qwen Image Edit 2511 LoRA is an enhanced version with custom LoRA support for personalized styles. It delivers stronger edit consistency, robust multi-person identity/pose consistency, custom LoRA styles, enhanced industrial/product design, and improved geometric reasoning for structure-preserving edits. Built for stable production use with a ready-to-use REST API, no cold starts, and predictable pricing.

từ $0.025
image-to-imageQwen Image Edit 2511 — bản xem trước image-to-image Qwen Image từ Alibaba

Qwen Image Edit 2511

Qwen Image Edit 2511 is a major upgrade over 2509 for real-world image editing and design. It delivers stronger edit consistency, robust multi-person identity/pose consistency, built-in LoRA styles, enhanced industrial/product design, and improved geometric reasoning for structure-preserving edits. Built for stable production use with a ready-to-use REST API, no cold starts, and predictable pricing.

từ $0.02
image-to-imageQwen Image Layered — bản xem trước image-to-image Qwen Image từ Alibaba

Qwen Image Layered

Qwen-Image Layered is a unified image-layer decomposition model for prompt-guided compositing. Provide points, boxes, or rough masks to isolate subjects and regions, and the model splits a single image into multiple RGBA layers with clean alpha, soft edges, and correct occlusion order. Ready-to-use REST inference API with fast response, no cold starts, and affordable pricing.

từ $0.025
image-to-imageQwen Image Edit Plus Lora — bản xem trước image-to-image Qwen Image từ Alibaba

Qwen Image Edit Plus Lora

Qwen-Image-Edit-Plus (2509) is 20B MMDiT image-to-image editor supporting multi-image edits, single-image consistency, and native ControlNet. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

từ $0.025
image-to-imageQwen Image Edit Plus — bản xem trước image-to-image Qwen Image từ Alibaba

Qwen Image Edit Plus

Qwen-Image-Edit-Plus (2509) is a 20B MMDiT image editor with multi-image editing, single-image consistency and native ControlNet support. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

từ $0.02
image-to-imageQwen Image Edit Lora — bản xem trước image-to-image Qwen Image từ Alibaba

Qwen Image Edit Lora

Qwen-Image-Edit LoRA (20B) enables bilingual Chinese/English image-to-image editing with style preservation and semantic and appearance edits. Ready-to-use REST API, best performance, no coldstarts, affordable pricing.

từ $0.025
image-to-imageQwen Image Edit — bản xem trước image-to-image Qwen Image từ Alibaba

Qwen Image Edit

Qwen-Image-Edit is a 20B MMDiT image-to-image model offering precise bilingual (Chinese & English) text edits while preserving style. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

từ $0.02
trainingQwen Image Lora Trainer — bản xem trước training Qwen Image từ Alibaba

Qwen Image Lora Trainer

Train custom Qwen-Image LoRA models 10x faster. Style training, character training, object training. From concept to model in minutes, not hours. Upload a ZIP file containing images to start!

từ $1.00
text-to-imageQwen Image Text To Image Lora — bản xem trước text-to-image Qwen Image từ Alibaba

Qwen Image Text To Image Lora

Qwen-Image LoRA is a 20B MMDiT next-gen text-to-image model with LoRA support for fast customization and refined image generation. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

từ $0.025
text-to-imageQwen Image Text To Image — bản xem trước text-to-image Qwen Image từ Alibaba

Qwen Image Text To Image

Qwen-Image is a 20B MMDiT next-gen text-to-image model that generates images from text prompts. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

từ $0.02

Ví dụ

Xem Qwen Image hoạt động

Kết quả thật do API Qwen Image tạo ra. Rê chuột lên video để xem trước, nhấp để mở trình xem kích thước đầy đủ.

Hướng dẫn

Cách dùng API Qwen Image

Bốn bước từ đăng ký đến khi có kết quả. Ví dụ đầy đủ bằng Python, Node.js và cURL nằm ở phần API bên dưới.

  1. 01

    Lấy API key

    Đăng ký tài khoản WaveSpeedAI và sao chép API key từ bảng điều khiển. Tài khoản mới được tặng credit dùng thử miễn phí — đủ để chạy playground vài chục lần trước khi bắt đầu tính phí.

  2. 02

    Gửi một prediction

    POST đầu vào của bạn dưới dạng JSON tới https://api.wavespeed.ai/api/v3/wavespeed-ai/qwen-image/text-to-image. Endpoint trả về prediction id ngay lập tức — việc tạo là bất đồng bộ nên bạn không phải giữ kết nối mở trong lúc suy luận.

  3. 03

    Poll để chờ hoàn tất

    GET https://api.wavespeed.ai/api/v3/predictions/{request_id}/result. Trả về kết quả khi completed; dừng và báo lỗi khi failed, cancelled, timeout hoặc deleted; tiếp tục poll với mọi trạng thái khác.

  4. 04

    Đọc URL kết quả

    Khi status là "completed", đọc URL từ data.outputs[0]. URL trỏ tới media bạn đã tạo trên CDN của WaveSpeedAI — ảnh, video, âm thanh hoặc tệp 3D tùy biến thể Qwen Image bạn đã gọi.

Trường hợp sử dụng

Bạn có thể xây dựng gì với Qwen Image

Các quy trình phổ biến mà nhà phát triển và người sáng tạo dùng API Qwen Image cho.

01

Text-to-image MMDiT 20B

wavespeed-ai/qwen-image/text-to-image là mô hình text-to-image MMDiT 20B thế hệ mới tạo ảnh từ prompt văn bản — endpoint tạo ảnh cơ bản trong dòng Qwen Image.

text-to-image20bmmdit
02

2512 nâng cao với kết xuất chữ vượt trội

wavespeed-ai/qwen-image/text-to-image-2512 là mô hình text-to-image mới nhất của Qwen với khả năng hiểu prompt nâng cao, kết xuất chữ vượt trội và hỗ trợ nhiều tỷ lệ khung hình theo danh mục.

2512typographyaspect-ratio
03

Chỉnh sửa nhiều ảnh với Edit-Plus

wavespeed-ai/qwen-image/edit-plus là trình chỉnh sửa MMDiT 20B với chỉnh sửa nhiều ảnh, nhất quán trên một ảnh và hỗ trợ ControlNet gốc — hữu ích cho các chỉnh sửa phức tạp tham chiếu nhiều ảnh nguồn.

edit-plusmulti-imagecontrolnet
04

Điều khiển góc máy 96 tư thế

wavespeed-ai/qwen-image/edit-multiple-angles tạo các góc máy cụ thể từ một ảnh bằng hệ thống máy quay 96 tư thế — điều khiển xoay ngang, nghiêng dọc và thu phóng để có góc chính diện, bên hông, phía sau và hơn thế.

camera-angles96-pose3d-views
05

Phân tách ghép lớp

wavespeed-ai/qwen-image/layered là mô hình phân tách lớp ảnh hợp nhất cho ghép lớp theo prompt — cung cấp điểm, khung hoặc mặt nạ thô để tách chủ thể và chia một ảnh thành nhiều lớp.

layeredcompositingdecomposition
06

Tùy chỉnh LoRA

Các biến thể hỗ trợ LoRA (text-to-image-2512-lora, edit-lora, edit-plus-lora) cho phép tùy chỉnh nhanh và tạo ảnh tinh chỉnh — huấn luyện hoặc áp dụng checkpoint LoRA để nhất quán phong cách, nhân vật hoặc thương hiệu.

loracustomizationconsistency

Mẹo

Mẹo viết prompt cho Qwen Image

Lời khuyên thực tế để có kết quả tốt hơn từ Qwen Image — rút ra từ các mẫu hiệu quả trên các mô hình ảnh trong các pipeline sản xuất.

  1. 01

    Dùng 2512 cho kết xuất chữ mới nhất

    text-to-image-2512 là biến thể nâng cao với kết xuất chữ và khả năng hiểu prompt vượt trội — hãy chọn nó thay cho text-to-image cơ bản khi typography quan trọng.

  2. 02

    Edit-Multiple-Angles cho các góc nhìn sản phẩm

    edit-multiple-angles tạo góc chính diện, bên hông, phía sau và các góc máy tùy chỉnh từ một ảnh sản phẩm — hữu ích cho danh mục thương mại điện tử mà không cần buổi chụp nhiều máy.

  3. 03

    Edit-Plus cho tham chiếu nhiều ảnh

    Khi chỉnh sửa phụ thuộc vào nhiều ảnh nguồn, hãy dùng edit-plus với hỗ trợ ControlNet gốc thay vì chỉnh sửa một ảnh.

  4. 04

    Layered cho quy trình ghép lớp

    Dùng layered để phân tách ảnh thành các lớp theo prompt — tách chủ thể bằng điểm, khung hoặc mặt nạ thô cho khâu ghép lớp phía sau.

  5. 05

    Biến thể LoRA để nhất quán thương hiệu

    Áp dụng checkpoint LoRA đã huấn luyện qua text-to-image-2512-lora hoặc edit-plus-lora để giữ phong cách, nhân vật hoặc danh tính thương hiệu lặp lại giữa các lần tạo.

  6. 06

    Chỉnh sửa song ngữ Trung/Anh

    Edit và Edit-LoRA hỗ trợ chỉnh sửa image-to-image song ngữ Trung/Anh với giữ nguyên phong cách — hữu ích cho quy trình địa phương hóa.

Giá

Giá API Qwen Image

Giá tính theo từng kết quả đầu ra. Khoản phí cuối cùng thay đổi theo các tham số bạn đặt trong playground của từng biến thể (độ phân giải, thời lượng, số kết quả, tham chiếu).

API

Gọi API Qwen Image

Đăng ký API key tại wavespeed.ai/accesskey, rồi gửi prediction qua REST. Playground tạo sẵn mã mẫu có thể dán ngay cho mọi tổ hợp đầu vào.

POSThttps://api.wavespeed.ai/api/v3/wavespeed-ai/qwen-image/text-to-image

# 1. Submit the prediction.
SUBMIT_RESPONSE=$(curl --silent --show-error --fail-with-body \
  -X POST "https://api.wavespeed.ai/api/v3/wavespeed-ai/qwen-image/text-to-image" \
  -H "Content-Type: application/json" \
  -H "Authorization: Bearer $WAVESPEED_API_KEY" \
  -d '{
    "prompt": "A cinematic shot of a city at sunset, soft golden light",
    "size": "1024*1024",
    "output_format": "jpeg"
}')

TASK=$(printf '%s' "$SUBMIT_RESPONSE" | jq 'if has("data") then .data else . end')
PREDICTION_ID=$(printf '%s' "$TASK" | jq -r '.id')
if [ -z "$PREDICTION_ID" ] || [ "$PREDICTION_ID" = "null" ]; then
  printf 'Submission response did not contain a prediction id
' >&2
  exit 1
fi
RESULT_URL="https://api.wavespeed.ai/api/v3/predictions/$PREDICTION_ID/result"

# 2. Poll until the prediction finishes.
while true; do
  RESPONSE=$(curl --silent --show-error --fail-with-body "$RESULT_URL" \
    -H "Authorization: Bearer $WAVESPEED_API_KEY")
  RESULT=$(printf '%s' "$RESPONSE" | jq 'if has("data") then .data else . end')
  STATUS=$(printf '%s' "$RESULT" | jq -r '.status')
  case "$STATUS" in
    completed) printf '%s\n' "$RESULT" | jq '.outputs'; break ;;
    failed|cancelled|timeout|deleted) printf '%s\n' "$RESULT" | jq . >&2; exit 1 ;;
    *) sleep 2 ;;
  esac
done

So sánh

Qwen Image so với các lựa chọn khác

Khi nào nên chọn Qwen Image thay vì các mô hình tương tự trên WaveSpeedAI.

Qwen Image so với Seedream 4.5

Seedream 4.5 nhấn mạnh typography và có các biến thể Sequential để khóa danh tính nhiều ảnh. Qwen Image bao quát phạm vi chỉnh sửa rộng hơn — Edit-Plus, multiple-angles, ghép lớp và chỉnh sửa song ngữ Trung/Anh.

Qwen Image so với GPT Image 2

GPT Image 2 có các bậc chất lượng rõ ràng và quy trình chỉnh sửa bằng ảnh tham chiếu. Qwen Image có hỗ trợ ControlNet gốc, góc máy 96 tư thế và phân tách lớp — các công cụ chỉnh sửa khác biệt với chi phí mỗi lần gọi thấp hơn.

Qwen Image so với Nano Banana 2

Nano Banana 2 có tính nhất quán nhiều nhân vật (tối đa 5) và căn cứ tìm kiếm web. Qwen Image thắng ở chiều sâu chỉnh sửa — Edit-Plus nhiều ảnh, tạo góc máy và phân tách lớp theo prompt.

Câu hỏi thường gặp

Qwen Image API — Câu hỏi thường gặp

Giá, giấy phép, tích hợp — những câu hỏi phổ biến về việc chạy Qwen Image trên WaveSpeedAI.

API Qwen Image là gì?

Qwen Image là mô hình tạo ảnh của Alibaba, được cung cấp dưới dạng REST API trên WaveSpeedAI. Alibaba Qwen-Image — bộ công cụ text-to-image và chỉnh sửa MMDiT 20B thế hệ mới với hỗ trợ song ngữ Trung/Anh, chỉnh sửa nhiều ảnh, tùy chỉnh LoRA, ghép lớp và hệ thống góc máy 96 tư thế. Bạn có thể gọi bằng lập trình hoặc thử từ playground ở liên kết phía trên.

Làm sao để gọi API Qwen Image?

Đăng ký tài khoản WaveSpeedAI, sao chép API key từ /accesskey, rồi POST tới https://api.wavespeed.ai/api/v3/wavespeed-ai/qwen-image/text-to-image với đầu vào dưới dạng JSON. Endpoint trả về prediction id. Hãy poll endpoint kết quả khoảng mỗi 2 giây, tăng khoảng cách với các tác vụ chạy lâu, và dừng ở bất kỳ trạng thái kết thúc nào. Ví dụ Python / Node.js / cURL hướng tới môi trường production ở phía trên.

API Qwen Image có giá bao nhiêu?

Qwen Image bắt đầu từ $0.02 mỗi lượt chạy. Chi phí chính xác thay đổi theo các tham số bạn đặt (độ phân giải, thời lượng, số kết quả, tham chiếu). Phần xem trước chi phí trực tiếp cạnh nút Tạo trong playground hiển thị giá chính xác cho đầu vào hiện tại của bạn.

Có những biến thể Qwen Image nào?

WaveSpeedAI cung cấp 21 endpoint Qwen Image đang hoạt động: wavespeed-ai/qwen-image-2.0/edit, wavespeed-ai/qwen-image-2.0-pro/edit, wavespeed-ai/qwen-image-2.0/text-to-image, wavespeed-ai/qwen-image-2.0-pro/text-to-image, wavespeed-ai/qwen-image/edit-2509-multiple-angles, wavespeed-ai/qwen-image-max/edit, wavespeed-ai/qwen-image-max/text-to-image, wavespeed-ai/qwen-image/edit-multiple-angles, và nhiều hơn nữa. Mỗi biến thể có trang playground và giá riêng.

Tôi có thể dùng kết quả của Qwen Image cho mục đích thương mại không?

Quyền sử dụng thương mại tuân theo giấy phép mô hình của Alibaba. Hầu hết mô hình của Alibaba cho phép dùng kết quả đầu ra cho mục đích thương mại; hãy xem trang playground của từng mô hình để biết tóm tắt giấy phép cụ thể, và Điều khoản dịch vụ của WaveSpeedAI cho các điều kiện cấp nền tảng.

Tại sao dùng Qwen Image trên WaveSpeedAI thay vì gọi trực tiếp?

Một API key + một tài khoản thanh toán cho Qwen Image VÀ hơn 1.000 mô hình AI khác từ các nhà cung cấp khác. Không cần thiết lập SDK cho từng hãng, không có hạn mức tốc độ riêng, không phải viết lại mã tích hợp cho từng hãng. Giá thường ngang bằng hoặc thấp hơn API trực tiếp của Alibaba.

Nhà cung cấp

Giới thiệu về Alibaba

Đội ngũ đứng sau Qwen Image và dòng mô hình Alibaba rộng hơn trên WaveSpeedAI.

Tongyi Lab của Alibaba tạo ra dòng mô hình video Wan và dòng LLM Qwen. Wan nổi bật nhờ được phát hành với trọng số mở, phủ nhiều biến thể (text-to-video, image-to-video, reference-to-video, video-edit, video-extend, image-edit, text-to-image) và luôn mạnh về độ ổn định chuyển động cũng như khả năng bám sát prompt với các prompt đa ngôn ngữ.

Bắt đầu xây dựng với Qwen Image trên WaveSpeedAI

Tặng credit dùng thử miễn phí khi đăng ký. Một API key cho hơn 1.000 mô hình AI từ Alibaba và mọi nhà cung cấp khác.