openai/gpt-oss-120b
प्रकाशन तिथि: 2025-08-06
131,072 context · $0.15/M input tokens · $0.60/M output tokens
gpt-oss-120b is an open-weight, 117B-parameter Mixture-of-Experts (MoE) language model from OpenAI designed for high-reasoning, agentic, and general-purpose production use cases. It activates 5.1B parameters per forward pass and is optimized...
उपयोग के अनुसार भुगतान
कोई अग्रिम लागत नहीं, केवल उतना ही भुगतान करें जितना आप उपयोग करते हैं
हमारे API के साथ एकीकृत करने के लिए निम्नलिखित कोड उदाहरणों का उपयोग करें:
import OpenAI from 'openai';
if (!process.env.WAVESPEED_API_KEY) throw new Error('Set WAVESPEED_API_KEY');
const client = new OpenAI({
apiKey: process.env.WAVESPEED_API_KEY,
baseURL: 'https://llm.wavespeed.ai/v1',
timeout: 120_000,
maxRetries: 2,
});
try {
const response = await client.chat.completions.create({
model: 'openai/gpt-oss-120b',
messages: [{ role: 'user', content: 'Hello!' }],
});
console.log(response.choices[0]?.message?.content ?? '');
} catch (error) {
console.error('LLM request failed:', error);
process.exitCode = 1;
}import OpenAI from 'openai';
if (!process.env.WAVESPEED_API_KEY) throw new Error('Set WAVESPEED_API_KEY');
const client = new OpenAI({
apiKey: process.env.WAVESPEED_API_KEY,
baseURL: 'https://llm.wavespeed.ai/v1',
timeout: 120_000,
maxRetries: 2,
});
try {
const response = await client.chat.completions.create({
model: 'openai/gpt-oss-120b',
messages: [{ role: 'user', content: 'Hello!' }],
});
console.log(response.choices[0]?.message?.content ?? '');
} catch (error) {
console.error('LLM request failed:', error);
process.exitCode = 1;
}gpt-oss-120b is an open-weight, 117B-parameter Mixture-of-Experts (MoE) language model from OpenAI designed for high-reasoning, agentic, and general-p
gpt-oss-120b is an open-weight, 117B-parameter Mixture-of-Experts (MoE) language model from OpenAI designed for high-reasoning, agentic, and general-purpose production use cases. It activates 5.1B parameters per forward pass and is optimized to run on a single H100 GPU with native MXFP4 quantization. The model supports configurable reasoning depth, full chain-of-thought access, and native tool use, including function calling, browsing, and structured output generation.
| Specification | Value |
|---|---|
| Provider | Openai |
| Model Type | Large Language Model (LLM) |
| Architecture | N/A |
| Context Window | 131072 tokens |
| Max Output | tokens |
| Input | Text |
| Output | Text |
| Vision | Supported |
| Function Calling | Supported |
| Token Type | Cost per Million Tokens |
|---|---|
| Input | $0.0 |
| Output | $0.2 |
Base URL: https://llm.wavespeed.ai/v1 API Endpoint: chat/completions Model ID: openai/gpt-oss-120b
from openai import OpenAI
client = OpenAI(
api_key="YOUR_API_KEY",
base_url="https://llm.wavespeed.ai/v1"
)
response = client.chat.completions.create(
model="openai/gpt-oss-120b",
messages=[
{"role": "user", "content": "Hello!"}
]
)
print(response.choices[0].message.content)
curl https://llm.wavespeed.ai/v1/chat/completions \
-H "Content-Type: application/json" \
-H "Authorization: Bearer YOUR_API_KEY" \
-d '{
"model": "openai/gpt-oss-120b",
"messages": [{"role": "user", "content": "Hello!"}]
}'
openai/gpt-oss-120b
gpt-oss-120b is an open-weight, 117B-parameter Mixture-of-Experts (MoE) language model from OpenAI designed for high-reasoning, agentic, and general-purpose production use cases. It activates 5.1B parameters per forward pass and is optimized...
इनपुट
$0.15 /M
आउटपुट
$0.6 /M
कॉन्टेक्स्ट
131K
टूल उपयोग
समर्थित
हमारे एकीकृत API के ज़रिए GPT Oss 120b तक पहुंचें — OpenAI-कम्पैटिबल, कोई कोल्ड स्टार्ट नहीं, पारदर्शी मूल्य निर्धारण।
WaveSpeedAI पर मूल्य निर्धारण: प्रति मिलियन इनपुट टोकन $0.15 और प्रति मिलियन आउटपुट टोकन $0.60। प्रॉम्प्ट कैशिंग और बैच प्रोसेसिंग की बिलिंग अलग से होती है और लंबे, दोहराव वाले वर्कलोड पर प्रभावी लागत कम करती है।
GPT Oss 120b प्रति अनुरोध — टोकन तक के आउटपुट के साथ 131K टोकन तक के कॉन्टेक्स्ट को सपोर्ट करता है।
WaveSpeedAI GPT Oss 120b को https://llm.wavespeed.ai/v1 पर OpenAI-कम्पैटिबल Chat Completions इंटरफ़ेस के ज़रिए उपलब्ध कराता है। अधिकांश OpenAI SDK क्लाइंट बेस URL और API कुंजी बदलकर काम करते हैं; वैकल्पिक फ़ील्ड चयनित मॉडल पर निर्भर करते हैं।
WaveSpeedAI में साइन इन करें, Access Keys में एक API कुंजी बनाएं, फिर ऊपर दिखाए गए मॉडल id के साथ https://llm.wavespeed.ai/v1/chat/completions पर एक अनुरोध भेजें। उपलब्धता, क्षमताओं और मूल्य निर्धारण के लिए वर्तमान मॉडल कैटलॉग देखें।