Qwen3 VL 30b A3b Thinking | Qwen LLM API

Name: Qwen3 Vl 30b A3b Thinking API
Brand: qwen
Price: 0.2 USD
Availability: InStock

Użycie API

Użyj poniższych przykładów kodu, aby zintegrować się z naszym API:

from openai import OpenAI

client = OpenAI(
    api_key="YOUR_API_KEY",
    base_url="https://llm.wavespeed.ai/v1"
)

response = client.chat.completions.create(
    model="qwen/qwen3-vl-30b-a3b-thinking",
    messages=[
        {"role": "user", "content": "Hello!"}
    ]
)

print(response.choices[0].message.content)

Wprowadzenie do modelu

Qwen qwen3-vl-30b-a3b-thinking

Qwen3-VL-30B-A3B-Thinking is a multimodal model that unifies strong text generation with visual understanding for images and videos

Qwen3-VL-30B-A3B-Thinking is a multimodal model that unifies strong text generation with visual understanding for images and videos. Its Thinking variant enhances reasoning in STEM, math, and complex tasks. It excels in perception of real-world/synthetic categories, 2D/3D spatial grounding, and long-form visual comprehension, achieving competitive multimodal benchmark results. For agentic use, it handles multi-image multi-turn instructions, video timeline alignments, GUI automation, and visual coding from sketches to debugged UI. Text performance matches flagship Qwen3 models, suiting document AI, OCR, UI assistance, spatial tasks, and agent research.

Why It Looks Great

Large Language Model architecture for efficient processing
131072 context window for long document handling
Competitive pricing at $0.2/$1.1 per million tokens

Key Features

Context Window: 131072 tokens
Max Output: 32768 tokens
Vision: Supported
Function Calling: Supported

Specifications

Specification	Value
Provider	Qwen
Model Type	Large Language Model (LLM)
Architecture	N/A
Context Window	131072 tokens
Max Output	32768 tokens
Input	Text
Output	Text
Vision	Supported
Function Calling	Supported

Pricing

Token Type	Cost per Million Tokens
Input	$0.2
Output	$1.1

How to Use

Write your prompt — describe the task, provide context, and specify desired output format.
Submit — the model processes your request and returns the response.

API Integration

Base URL: https://llm.wavespeed.ai/v1 API Endpoint: chat/completions Model ID: qwen/qwen3-vl-30b-a3b-thinking

API Usage

Python SDK

from openai import OpenAI

client = OpenAI(
    api_key="YOUR_API_KEY",
    base_url="https://llm.wavespeed.ai/v1"
)

response = client.chat.completions.create(
    model="qwen/qwen3-vl-30b-a3b-thinking",
    messages=[
        {"role": "user", "content": "Hello!"}
    ]
)

print(response.choices[0].message.content)

cURL

curl https://llm.wavespeed.ai/v1/chat/completions \
  -H "Content-Type: application/json" \
  -H "Authorization: Bearer YOUR_API_KEY" \
  -d '{
    "model": "qwen/qwen3-vl-30b-a3b-thinking",
    "messages": [{"role": "user", "content": "Hello!"}]
  }'

Notes

Model: qwen/qwen3-vl-30b-a3b-thinking
Provider: Qwen

Najczęstsze pytania o Qwen3 Vl 30b A3b Thinking

Ile kosztuje API Qwen3 Vl 30b A3b Thinking?+

Cennik na WaveSpeedAI: $0.20 za milion tokenów wejściowych i $1.00 za milion tokenów wyjściowych. Prompt caching i przetwarzanie wsadowe są rozliczane oddzielnie i obniżają efektywny koszt długich, powtarzalnych obciążeń.

Jakie jest okno kontekstu Qwen3 Vl 30b A3b Thinking?+

Qwen3 Vl 30b A3b Thinking obsługuje do 131K tokenów kontekstu i do 33K tokenów wyjściowych na zapytanie.

Czy Qwen3 Vl 30b A3b Thinking jest kompatybilny z OpenAI?+

Tak. WaveSpeedAI udostępnia Qwen3 Vl 30b A3b Thinking przez endpoint kompatybilny z OpenAI pod https://llm.wavespeed.ai/v1. Skieruj oficjalny OpenAI SDK na ten base URL ze swoim kluczem API WaveSpeedAI — bez innych zmian w kodzie.

Jak zacząć z Qwen3 Vl 30b A3b Thinking?+

Zaloguj się do WaveSpeedAI, utwórz klucz API w Access Keys, a następnie wyślij żądanie na https://llm.wavespeed.ai/v1/chat/completions z id modelu pokazanym powyżej. Nowe konta otrzymują darmowe kredyty na ocenę Qwen3 Vl 30b A3b Thinking.

Cennik

Wypróbuj model