openai/gpt-5.4-mini
发布时间: 2026-03-17
400,000 context · $0.75/M input tokens · $4.50/M output tokens
GPT-5.4 mini brings the core capabilities of GPT-5.4 to a faster, more efficient model optimized for high-throughput workloads. It supports text and image inputs with strong performance across reasoning, coding, and tool use, while reducing latency and cost for large-scale deployments. The model is designed for production environments that require a balance of capability and efficiency, making it well suited for chat applications, coding assistants, and agent workflows that operate at scale. GPT-5.4 mini delivers reliable instruction following, solid multi-step reasoning, and consistent performance across diverse tasks with improved cost efficiency.
按量付费
无需预付费用,仅按实际使用量付费
使用以下代码示例接入我们的 API:
import os
from openai import OpenAI
client = OpenAI(
api_key=os.environ["WAVESPEED_API_KEY"],
base_url="https://llm.wavespeed.ai/v1",
timeout=120.0,
max_retries=2,
)
try:
response = client.chat.completions.create(
model="openai/gpt-5.4-mini",
messages=[{"role": "user", "content": "Hello!"}],
)
print(response.choices[0].message.content or "")
except Exception as exc:
raise SystemExit(f"LLM request failed: {exc}") from excopenai/gpt-5.4-mini
GPT-5.4 mini brings the core capabilities of GPT-5.4 to a faster, more efficient model optimized for high-throughput workloads. It supports text and image inputs with strong performance across reasoning, coding, and tool use, while reducing latency and cost for large-scale deployments. The model is designed for production environments that require a balance of capability and efficiency, making it well suited for chat applications, coding assistants, and agent workflows that operate at scale. GPT-5.4 mini delivers reliable instruction following, solid multi-step reasoning, and consistent performance across diverse tasks with improved cost efficiency.
输入
$0.75 /M
输出
$4.5 /M
上下文
400K
最大输出
128K
Vision
支持
工具调用
支持
WaveSpeedAI 定价:输入每百万 token $0.75,输出每百万 token $4.50。Prompt 缓存和批处理单独计费,可显著降低长上下文、高重复任务的实际成本。
GPT 5.4 Mini 单次请求最多支持 400K 上下文 token,输出最多 128K token。
WaveSpeedAI 通过 https://llm.wavespeed.ai/v1 的 OpenAI 兼容 Chat Completions 接口提供 GPT 5.4 Mini。大多数 OpenAI SDK 客户端只需更换 base URL 和 API Key;可选字段取决于具体模型。
登录 WaveSpeedAI,在 Access Keys 中创建 API Key,然后使用上方显示的 model id 向 https://llm.wavespeed.ai/v1/chat/completions 发送请求。模型可用性、能力和价格请以当前模型目录为准。