GPT Image 2.5 is LIVE — Flare & Sunburst | Try in Image Generator →
xai
x-ai/grok-4.6

x-ai/grok-4.6

Release date: 2026-08-12

500,000 context · $2.00/M input tokens · $6.00/M output tokens

Grok 4.6 is SpaceXAI's smartest model with frontier performance on coding, knowledge work, and STEM.

Pricing

Pay-per-use

No upfront costs, pay only for what you use

Input
> 200K $4.00 / M Tokens
Output
> 200K $12.00 / M Tokens
Cache Read
> 200K $1.00 / M Tokens

Try the model

x-ai/grok-4.6
Online
xai
Hi! I am a helpful AI assistant. What can I do for you?
Ready to use this model in a local coding agent?Agent setup

API Usage

Use the following code examples to integrate with our API:

import OpenAI from 'openai';

if (!process.env.WAVESPEED_API_KEY) throw new Error('Set WAVESPEED_API_KEY');
const client = new OpenAI({
  apiKey: process.env.WAVESPEED_API_KEY,
  baseURL: 'https://llm.wavespeed.ai/v1',
  timeout: 120_000,
  maxRetries: 2,
});

try {
  const response = await client.chat.completions.create({
    model: 'x-ai/grok-4.6',
    messages: [{ role: 'user', content: 'Hello!' }],
  });
  console.log(response.choices[0]?.message?.content ?? '');
} catch (error) {
  console.error('LLM request failed:', error);
  process.exitCode = 1;
}

Model Introduction

SpaceXAI: Grok 4.6

Grok 4.6 is SpaceXAI's smartest model with frontier performance on coding, knowledge work, and STEM.

This model is imported from OpenRouter metadata and exposed through the WaveSpeed AI OpenAI-compatible API for chat completions and compatible application workflows.


Why It Looks Great

  • text+image+file->text architecture for Text, Image, file to Text workloads
  • 500000 context window for long prompts, document analysis, and multi-turn workflows
  • Competitive pricing at $2/$6 per million tokens
  • Vision input support for image understanding and multimodal tasks
  • Function calling and tool-use support for agentic application workflows
  • Structured output support for JSON responses and schema-constrained generation
  • Reasoning controls available through supported OpenRouter parameters

Key Features

  • Context Window: 500000 tokens
  • Max Input: Not listed
  • Max Output: Not listed
  • Vision: Supported
  • Function Calling: Supported
  • Structured Outputs: Supported
  • Image Generation: Not listed
  • Audio Input: Not listed
  • Supported Parameters: frequency_penalty, include_reasoning, logprobs, max_tokens, presence_penalty, reasoning, reasoning_effort, response_format, seed, stop, structured_outputs, temperature, tool_choice, tools, top_logprobs, top_p

Specifications

SpecificationValue
Providerxai
Model TypeChat Completions model
Architecturetext+image+file->text
Context Window500000 tokens
Max InputNot listed
Max OutputNot listed
InputText, Image, file
OutputText
VisionSupported
Function CallingSupported
Structured OutputsSupported
OpenRouter CreatedAugust 12, 2026

Pricing

Token TypeCost
Input$2 per million tokens
Output$6 per million tokens
Cached Input$0.5 per million tokens
Web Search$5000 per million tokens

Note: Pricing is generated from OpenRouter model metadata. If multiple upstream providers expose different endpoint prices, review and adjust the price before publishing.


How to Use

  1. Write your prompt - describe the task, provide context, and specify the desired output format.
  2. Submit - the model processes your request and returns the response.

API Integration

Base URL: https://llm.wavespeed.ai/v1 API Endpoint: chat/completions Model ID: x-ai/grok-4.6


API Usage

Python SDK

from openai import OpenAI

client = OpenAI(
    api_key="YOUR_API_KEY",
    base_url="https://llm.wavespeed.ai/v1"
)

response = client.chat.completions.create(
    model="x-ai/grok-4.6",
    messages=[{"role": "user", "content": "Hello!"}]
)

print(response.choices[0].message.content)

cURL

curl https://llm.wavespeed.ai/v1/chat/completions \
  -H "Content-Type: application/json" \
  -H "Authorization: Bearer YOUR_API_KEY" \
  -d '{
    "model": "x-ai/grok-4.6",
    "messages": [{"role": "user", "content": "Hello!"}]
  }'

Notes

  • Model: x-ai/grok-4.6
  • Provider: xai
  • Endpoint availability and provider-specific routing may change on OpenRouter

Sources: OpenRouter model metadata.

Info

Providerxai
Typellm

Supported Functionality

Input
TextImage
Output
Text
Context500,000
Max Output-
Vision✓ Supported
Function Calling✓ Supported

API Access Guide

Base URLhttps://llm.wavespeed.ai/v1
API Endpointchat/completions
Model IDx-ai/grok-4.6

Grok 4.6 API

x-ai/grok-4.6

Grok 4.6 is SpaceXAI's smartest model with frontier performance on coding, knowledge work, and STEM.

Input

$2 /M

Output

$6 /M

Context

500K

Vision

Supported

Tool Use

Supported

Try Grok 4.6 on WaveSpeedAI

Access Grok 4.6 through our unified API — OpenAI-compatible, no cold starts, transparent pricing.

Frequently Asked Questions about Grok 4.6

How much does Grok 4.6 cost via the API?+

Pricing on WaveSpeedAI: $2.00 per million input tokens and $6.00 per million output tokens. Prompt caching and batch processing are billed separately and reduce effective cost on long, repetitive workloads.

What is the context window of Grok 4.6?+

Grok 4.6 supports up to 500K tokens of context with up to — tokens of output per request.

Is Grok 4.6 OpenAI-compatible?+

WaveSpeedAI exposes Grok 4.6 at https://llm.wavespeed.ai/v1 through the OpenAI-compatible Chat Completions interface. Most OpenAI SDK clients work by changing the base URL and API key; optional fields depend on the selected model.

How do I get started with Grok 4.6?+

Sign in to WaveSpeedAI, create an API key in Access Keys, then send a request to https://llm.wavespeed.ai/v1/chat/completions with the model id shown above. Check the current model catalog for availability, capabilities, and pricing.

Related LLM APIs