Seedream 5.0 Pro अब लाइव है | Image Generator में आज़माएं →
साइन इन
X
xiaomi/mimo-v2.5-pro

xiaomi/mimo-v2.5-pro

प्रकाशन तिथि: 2026-04-23

1,048,576 context · $1.00/M input tokens · $3.00/M output tokens

MiMo-V2.5-Pro is Xiaomi’s flagship open model for advanced agentic workflows, complex software engineering, and long-horizon task execution. Built on a sparse Mixture-of-Experts architecture with 1.02T total parameters and 42B active parameters, it supports a 1M-token context window and is optimized for autonomous coding agents, large codebase reasoning, tool-use workflows, and multi-step problem solving. It delivers strong performance on agentic and software engineering benchmarks such as ClawEval, GDPVal, and SWE-bench Pro, with an emphasis on token-efficient long-context execution.

मूल्य निर्धारण

उपयोग के अनुसार भुगतान

कोई अग्रिम लागत नहीं, केवल उतना ही भुगतान करें जितना आप उपयोग करते हैं

इनपुट
256K $1.00 / M Tokens
> 256K $2.00 / M Tokens
आउटपुट
256K $3.00 / M Tokens
> 256K $6.00 / M Tokens
Cache Read
256K $0.20 / M Tokens
> 256K $0.40 / M Tokens

मॉडल आज़माएं

xiaomi/mimo-v2.5-pro
ऑनलाइन
X
नमस्ते! मैं एक सहायक AI असिस्टेंट हूं। मैं आपके लिए क्या कर सकता हूं?
क्या इस मॉडल को लोकल कोडिंग एजेंट में इस्तेमाल करने के लिए तैयार हैं?एजेंट सेटअप

API उपयोग

हमारे API के साथ एकीकृत करने के लिए निम्नलिखित कोड उदाहरणों का उपयोग करें:

import OpenAI from 'openai';

if (!process.env.WAVESPEED_API_KEY) throw new Error('Set WAVESPEED_API_KEY');
const client = new OpenAI({
  apiKey: process.env.WAVESPEED_API_KEY,
  baseURL: 'https://llm.wavespeed.ai/v1',
  timeout: 120_000,
  maxRetries: 2,
});

try {
  const response = await client.chat.completions.create({
    model: 'xiaomi/mimo-v2.5-pro',
    messages: [{ role: 'user', content: 'Hello!' }],
  });
  console.log(response.choices[0]?.message?.content ?? '');
} catch (error) {
  console.error('LLM request failed:', error);
  process.exitCode = 1;
}

मॉडल परिचय

Xiaomi: MiMo-V2.5-Pro

MiMo-V2.5-Pro is Xiaomi’s flagship open model for advanced agentic workflows, complex software engineering, and long-horizon task execution. Built on a sparse Mixture-of-Experts architecture with 1.02T total parameters and 42B active parameters, it is optimized for autonomous coding agents, large codebase reasoning, tool use, and multi-step problem solving.


Why It Looks Great

  • Flagship Xiaomi MiMo model for complex agentic and software engineering workloads
  • Sparse Mixture-of-Experts architecture with 1.02T total parameters and 42B active parameters
  • 1M-token context window for long prompts, large codebases, documents, and multi-turn workflows
  • Strong fit for autonomous coding agents, long-horizon task execution, and tool-heavy workflows
  • Competitive performance on benchmarks such as ClawEval, GDPVal, and SWE-bench Pro
  • Designed for token-efficient agent trajectories and extended multi-step execution
  • Function calling and tool-use support for agentic application workflows
  • Structured output support for JSON responses and schema-constrained generation
  • Reasoning controls for tuning latency, quality, and cost per request

Key Features

  • Architecture: Sparse Mixture-of-Experts
  • Total Parameters: 1.02T
  • Active Parameters: 42B
  • Context Window: 1,048,576 tokens
  • Max Input: 1,032,192 tokens
  • Max Output: 16,384 tokens
  • Input: Text
  • Output: Text
  • Vision: Not listed
  • Function Calling: Supported
  • Structured Outputs: Supported
  • Thinking Mode: Supported
  • Image Generation: Not listed
  • Audio Input: Not listed
  • Supported Parameters: frequency_penalty, include_reasoning, logit_bias, max_tokens, min_p, presence_penalty, reasoning, repetition_penalty, response_format, seed, stop, structured_outputs, temperature, tool_choice, tools, top_k, top_p

Specifications

SpecificationValue
Providerxiaomi
Model TypeChat Completions model
ArchitectureSparse Mixture-of-Experts
Parameters1.02T total / 42B active
Context Window1,048,576 tokens
Max Input1,032,192 tokens
Max Output16,384 tokens
InputText
OutputText
VisionNot listed
Function CallingSupported
Structured OutputsSupported
Primary Use CasesAgentic coding, complex software engineering, long-horizon tasks, tool use

Pricing

Token TypeCost
Input$1.00 per million tokens
Output$3.00 per million tokens
Cached Input$0.20 per million tokens

How to Use

  1. Write your prompt - describe the task, provide context, and specify the desired output format.
  2. Submit - the model processes your request and returns the response.

API Integration

Base URL: https://llm.wavespeed.ai/v1
API Endpoint: chat/completions
Model ID: xiaomi/mimo-v2.5-pro


API Usage

Python SDK

from openai import OpenAI

client = OpenAI(
    api_key="YOUR_API_KEY",
    base_url="https://llm.wavespeed.ai/v1"
)

response = client.chat.completions.create(
    model="xiaomi/mimo-v2.5-pro",
    messages=[{"role": "user", "content": "Hello!"}]
)

print(response.choices[0].message.content)

cURL

curl https://llm.wavespeed.ai/v1/chat/completions \
  -H "Content-Type: application/json" \
  -H "Authorization: Bearer YOUR_API_KEY" \
  -d '{
    "model": "xiaomi/mimo-v2.5-pro",
    "messages": [{"role": "user", "content": "Hello!"}]
  }'

Notes

  • Model: xiaomi/mimo-v2.5-pro
  • Provider: xiaomi
  • Best suited for autonomous coding, complex software engineering, long-horizon reasoning, tool-heavy agent workflows, and large-context text tasks

जानकारी

प्रदाताxiaomi
प्रकारllm

समर्थित कार्यक्षमता

इनपुट
टेक्स्ट
आउटपुट
टेक्स्ट
कॉन्टेक्स्ट1,048,576
अधिकतम आउटपुट16,384
विज़न-
फ़ंक्शन कॉलिंग✓ समर्थित

API एक्सेस गाइड

बेस URLhttps://llm.wavespeed.ai/v1
API एंडपॉइंटchat/completions
मॉडल IDxiaomi/mimo-v2.5-pro

Mimo V2.5 Pro API

xiaomi/mimo-v2.5-pro

MiMo-V2.5-Pro is Xiaomi’s flagship open model for advanced agentic workflows, complex software engineering, and long-horizon task execution. Built on a sparse Mixture-of-Experts architecture with 1.02T total parameters and 42B active parameters, it supports a 1M-token context window and is optimized for autonomous coding agents, large codebase reasoning, tool-use workflows, and multi-step problem solving. It delivers strong performance on agentic and software engineering benchmarks such as ClawEval, GDPVal, and SWE-bench Pro, with an emphasis on token-efficient long-context execution.

इनपुट

$1 /M

आउटपुट

$3 /M

कॉन्टेक्स्ट

1049K

अधिकतम आउटपुट

16K

टूल उपयोग

समर्थित

WaveSpeedAI पर Mimo V2.5 Pro आज़माएं

हमारे एकीकृत API के ज़रिए Mimo V2.5 Pro तक पहुंचें — OpenAI-कम्पैटिबल, कोई कोल्ड स्टार्ट नहीं, पारदर्शी मूल्य निर्धारण।

Mimo V2.5 Pro के बारे में अक्सर पूछे जाने वाले सवाल

API के ज़रिए Mimo V2.5 Pro की लागत कितनी है?+

WaveSpeedAI पर मूल्य निर्धारण: प्रति मिलियन इनपुट टोकन $1.00 और प्रति मिलियन आउटपुट टोकन $3.00। प्रॉम्प्ट कैशिंग और बैच प्रोसेसिंग की बिलिंग अलग से होती है और लंबे, दोहराव वाले वर्कलोड पर प्रभावी लागत कम करती है।

Mimo V2.5 Pro की कॉन्टेक्स्ट विंडो क्या है?+

Mimo V2.5 Pro प्रति अनुरोध 16K टोकन तक के आउटपुट के साथ 1049K टोकन तक के कॉन्टेक्स्ट को सपोर्ट करता है।

क्या Mimo V2.5 Pro OpenAI-कम्पैटिबल है?+

WaveSpeedAI Mimo V2.5 Pro को https://llm.wavespeed.ai/v1 पर OpenAI-कम्पैटिबल Chat Completions इंटरफ़ेस के ज़रिए उपलब्ध कराता है। अधिकांश OpenAI SDK क्लाइंट बेस URL और API कुंजी बदलकर काम करते हैं; वैकल्पिक फ़ील्ड चयनित मॉडल पर निर्भर करते हैं।

मैं Mimo V2.5 Pro के साथ कैसे शुरू करूं?+

WaveSpeedAI में साइन इन करें, Access Keys में एक API कुंजी बनाएं, फिर ऊपर दिखाए गए मॉडल id के साथ https://llm.wavespeed.ai/v1/chat/completions पर एक अनुरोध भेजें। उपलब्धता, क्षमताओं और मूल्य निर्धारण के लिए वर्तमान मॉडल कैटलॉग देखें।

संबंधित LLM API

MiMo-V2.5-Pro | Xiaomi LLM API Pricing & Performance | WaveSpeedAI