Seedream 5.0 Pro ist LIVE | Jetzt im Bildgenerator testen →
alibaba
qwen/qwen3-vl-8b-instruct

qwen/qwen3-vl-8b-instruct

Veröffentlichungsdatum: 2025-10-15

131,072 context · $0.08/M input tokens · $0.50/M output tokens

Qwen3-VL-8B-Instruct is a multimodal vision-language model from the Qwen3-VL series, built for high-fidelity understanding and reasoning across text, images, and video. It features improved multimodal fusion with Interleaved-MRoPE for long-horizon...

Preise

Pay-per-Use

Keine Vorabkosten, zahlen Sie nur, was Sie nutzen

Eingabe$0.08 / M Tokens
Ausgabe$0.50 / M Tokens

Modell ausprobieren

qwen/qwen3-vl-8b-instruct
Online
alibaba
Hallo! Ich bin ein hilfreicher KI-Assistent. Womit kann ich helfen?
Bereit, dieses Modell in einem lokalen Coding-Agent zu verwenden?Agent-Setup

API-Nutzung

Verwenden Sie die folgenden Codebeispiele zur Integration mit unserer API:

import OpenAI from 'openai';

if (!process.env.WAVESPEED_API_KEY) throw new Error('Set WAVESPEED_API_KEY');
const client = new OpenAI({
  apiKey: process.env.WAVESPEED_API_KEY,
  baseURL: 'https://llm.wavespeed.ai/v1',
  timeout: 120_000,
  maxRetries: 2,
});

try {
  const response = await client.chat.completions.create({
    model: 'qwen/qwen3-vl-8b-instruct',
    messages: [{ role: 'user', content: 'Hello!' }],
  });
  console.log(response.choices[0]?.message?.content ?? '');
} catch (error) {
  console.error('LLM request failed:', error);
  process.exitCode = 1;
}

Modelleinführung

Qwen qwen3-vl-8b-instruct

**Qwen3-VL-8B-Instruct is a multimodal vision-language model from the Qwen3-VL series, built for high-fidelity understanding and reasoning across text, **

Qwen3-VL-8B-Instruct is a multimodal vision-language model from the Qwen3-VL series, built for high-fidelity understanding and reasoning across text, images, and video. It features improved multimodal fusion with Interleaved-MRoPE for long-horizon temporal reasoning, DeepStack for fine-grained visual-text alignment, and text-timestamp alignment for precise event localization.

The model supports a native 256K-token context window, extensible to 1M tokens, and handles both static and dynamic media inputs for tasks like document parsing, visual question answering, spatial reasoning, and GUI control. It achieves text understanding comparable to leading LLMs while expanding OCR coverage to 32 languages and enhancing robustness under varied visual conditions.


Why It Looks Great

  • Large Language Model architecture for efficient processing
  • 131072 context window for long document handling
  • Competitive pricing at $0.1/$0.5 per million tokens

Key Features

  • Context Window: 131072 tokens
  • Max Output: 32768 tokens
  • Vision: Supported
  • Function Calling: Supported

Specifications

SpecificationValue
ProviderQwen
Model TypeLarge Language Model (LLM)
ArchitectureN/A
Context Window131072 tokens
Max Output32768 tokens
InputText
OutputText
VisionSupported
Function CallingSupported

Pricing

Token TypeCost per Million Tokens
Input$0.1
Output$0.5

How to Use

  1. Write your prompt — describe the task, provide context, and specify desired output format.
  2. Submit — the model processes your request and returns the response.

API Integration

Base URL: https://llm.wavespeed.ai/v1 API Endpoint: chat/completions Model ID: qwen/qwen3-vl-8b-instruct


API Usage

Python SDK

from openai import OpenAI

client = OpenAI(
    api_key="YOUR_API_KEY",
    base_url="https://llm.wavespeed.ai/v1"
)

response = client.chat.completions.create(
    model="qwen/qwen3-vl-8b-instruct",
    messages=[
        {"role": "user", "content": "Hello!"}
    ]
)

print(response.choices[0].message.content)

cURL

curl https://llm.wavespeed.ai/v1/chat/completions \
  -H "Content-Type: application/json" \
  -H "Authorization: Bearer YOUR_API_KEY" \
  -d '{
    "model": "qwen/qwen3-vl-8b-instruct",
    "messages": [{"role": "user", "content": "Hello!"}]
  }'

Notes

  • Model: qwen/qwen3-vl-8b-instruct
  • Provider: Qwen

Info

Anbieteralibaba
Typllm

Unterstützte Funktionen

Eingabe
TextBild
Ausgabe
Text
Kontext131,072
Max. Ausgabe32,768
Vision✓ Unterstützt
Function Calling✓ Unterstützt

API-Zugriffsanleitung

Base URLhttps://llm.wavespeed.ai/v1
API-Endpunktchat/completions
Modell-IDqwen/qwen3-vl-8b-instruct

Qwen3 Vl 8b Instruct API

qwen/qwen3-vl-8b-instruct

Qwen3-VL-8B-Instruct is a multimodal vision-language model from the Qwen3-VL series, built for high-fidelity understanding and reasoning across text, images, and video. It features improved multimodal fusion with Interleaved-MRoPE for long-horizon...

Eingabe

$0.08 /M

Ausgabe

$0.5 /M

Kontext

131K

Max. Ausgabe

33K

Vision

Unterstützt

Tool-Nutzung

Unterstützt

Qwen3 Vl 8b Instruct auf WaveSpeedAI testen

Zugriff auf Qwen3 Vl 8b Instruct über unsere einheitliche API — OpenAI-kompatibel, keine Kaltstarts, transparente Preise.

Häufige Fragen zu Qwen3 Vl 8b Instruct

Wie viel kostet die Qwen3 Vl 8b Instruct-API?+

Preise auf WaveSpeedAI: $0.08 pro Million Input-Tokens und $0.50 pro Million Output-Tokens. Prompt-Caching und Batch-Verarbeitung werden separat berechnet und reduzieren die effektiven Kosten bei langen, sich wiederholenden Workloads.

Wie groß ist das Kontextfenster von Qwen3 Vl 8b Instruct?+

Qwen3 Vl 8b Instruct unterstützt bis zu 131K Kontext-Tokens und bis zu 33K Output-Tokens pro Anfrage.

Ist Qwen3 Vl 8b Instruct OpenAI-kompatibel?+

WaveSpeedAI stellt Qwen3 Vl 8b Instruct unter https://llm.wavespeed.ai/v1 über die OpenAI-kompatible Chat-Completions-Schnittstelle bereit. Bei den meisten OpenAI-SDK-Clients reichen Base-URL und API-Schlüssel; optionale Felder hängen vom Modell ab.

Wie starte ich mit Qwen3 Vl 8b Instruct?+

Melden Sie sich bei WaveSpeedAI an, erstellen Sie unter Access Keys einen API-Schlüssel und senden Sie eine Anfrage mit der oben gezeigten Modell-ID an https://llm.wavespeed.ai/v1/chat/completions. Verfügbarkeit, Fähigkeiten und Preise finden Sie im aktuellen Modellkatalog.

Verwandte LLM-APIs

Qwen3 VL 8b Instruct | Qwen LLM API | WaveSpeedAI