Seedream 5.0 Pro è LIVE | Prova nel Generatore di immagini →
alibaba
qwen/qwen3-235b-a22b-thinking-2507

qwen/qwen3-235b-a22b-thinking-2507

Data di rilascio: 2025-07-25

131,072 context · $0.15/M input tokens · $1.50/M output tokens

Qwen3-235B-A22B-Thinking-2507 is a high-performance, open-weight Mixture-of-Experts (MoE) language model optimized for complex reasoning tasks. It activates 22B of its 235B parameters per forward pass and natively supports up to 262,144...

In dismissione il 14 luglio 2026

Prezzi

Pay-per-use

Nessun costo iniziale, paga solo per ciò che usi

Input$0.15 / M Tokens
Output$1.50 / M Tokens

Prova il modello

qwen/qwen3-235b-a22b-thinking-2507
Online
alibaba
Ciao! Sono un assistente IA utile. Come posso aiutarti?
Pronto a usare questo modello in un coding agent locale?Setup agent

Utilizzo API

Usa i seguenti esempi di codice per integrare la nostra API:

from openai import OpenAI

client = OpenAI(
    api_key="YOUR_API_KEY",
    base_url="https://llm.wavespeed.ai/v1"
)

response = client.chat.completions.create(
    model="qwen/qwen3-235b-a22b-thinking-2507",
    messages=[
        {"role": "user", "content": "Hello!"}
    ]
)

print(response.choices[0].message.content)

Introduzione al modello

Qwen qwen3-235b-a22b-thinking-2507

qwen qwen3-235b-a22b-thinking-2507


Why It Looks Great

  • Large Language Model architecture for efficient processing
  • 262144 context window for long document handling
  • Competitive pricing at $0.1/$0.7 per million tokens

Key Features

  • Context Window: 262144 tokens
  • Max Output: 262144 tokens
  • Vision: Supported
  • Function Calling: Supported

Specifications

SpecificationValue
ProviderQwen
Model TypeLarge Language Model (LLM)
ArchitectureN/A
Context Window262144 tokens
Max Output262144 tokens
InputText
OutputText
VisionSupported
Function CallingSupported

Pricing

Token TypeCost per Million Tokens
Input$0.1
Output$0.7

How to Use

  1. Write your prompt — describe the task, provide context, and specify desired output format.
  2. Submit — the model processes your request and returns the response.

API Integration

Base URL: https://llm.wavespeed.ai/v1 API Endpoint: chat/completions Model ID: qwen/qwen3-235b-a22b-thinking-2507


API Usage

Python SDK

from openai import OpenAI

client = OpenAI(
    api_key="YOUR_API_KEY",
    base_url="https://llm.wavespeed.ai/v1"
)

response = client.chat.completions.create(
    model="qwen/qwen3-235b-a22b-thinking-2507",
    messages=[
        {"role": "user", "content": "Hello!"}
    ]
)

print(response.choices[0].message.content)

cURL

curl https://llm.wavespeed.ai/v1/chat/completions \
  -H "Content-Type: application/json" \
  -H "Authorization: Bearer YOUR_API_KEY" \
  -d '{
    "model": "qwen/qwen3-235b-a22b-thinking-2507",
    "messages": [{"role": "user", "content": "Hello!"}]
  }'

Notes

  • Model: qwen/qwen3-235b-a22b-thinking-2507
  • Provider: Qwen

Info

Provideralibaba
Tipollm

Funzionalità supportate

Input
Testo
Output
Testo
Contesto131,072
Output massimo262,144
Vision-
Function Calling✓ Supportato

Guida all'accesso API

Base URLhttps://llm.wavespeed.ai/v1
API Endpointchat/completions
ID modelloqwen/qwen3-235b-a22b-thinking-2507

Qwen3 235b A22b Thinking 2507 API

qwen/qwen3-235b-a22b-thinking-2507

Qwen3-235B-A22B-Thinking-2507 is a high-performance, open-weight Mixture-of-Experts (MoE) language model optimized for complex reasoning tasks. It activates 22B of its 235B parameters per forward pass and natively supports up to 262,144...

Input

$0.1495 /M

Output

$1.495 /M

Contesto

131K

Output max

262K

Uso strumenti

Supportato

Prova Qwen3 235b A22b Thinking 2507 su WaveSpeedAI

Accedi a Qwen3 235b A22b Thinking 2507 tramite la nostra API unificata — compatibile con OpenAI, senza cold start, prezzi trasparenti.

Domande frequenti su Qwen3 235b A22b Thinking 2507

Quanto costa Qwen3 235b A22b Thinking 2507 via API?+

Prezzi su WaveSpeedAI: $0.15 per milione di token in input e $1.50 per milione di token in output. Prompt caching e batch processing sono fatturati separatamente e riducono il costo effettivo su carichi lunghi e ripetitivi.

Qual è la context window di Qwen3 235b A22b Thinking 2507?+

Qwen3 235b A22b Thinking 2507 supporta fino a 131K token di contesto e fino a 262K token di output per richiesta.

Qwen3 235b A22b Thinking 2507 è compatibile con OpenAI?+

Sì. WaveSpeedAI espone Qwen3 235b A22b Thinking 2507 tramite un endpoint compatibile con OpenAI all'indirizzo https://llm.wavespeed.ai/v1. Punta l'SDK ufficiale di OpenAI a questa base URL con la tua API key WaveSpeedAI — senza altre modifiche al codice.

Come si inizia con Qwen3 235b A22b Thinking 2507?+

Accedi a WaveSpeedAI, crea una API key in Access Keys, poi invia una richiesta a https://llm.wavespeed.ai/v1/chat/completions con il model id mostrato sopra. I nuovi account ricevono crediti gratuiti per testare Qwen3 235b A22b Thinking 2507.

API LLM correlate

Qwen3 235b A22b Thinking 2507 | Qwen LLM API | WaveSpeedAI