跳至内容

模型详情

gpt-5.4-mini

由 OpenAI 提供
按量付费

GPT-5.4 mini brings the core capabilities of GPT-5.4 to a faster, more efficient model optimized for high-throughput workloads. It supports text and image inputs with strong performance across reasoning, coding, and tool use, while reducing latency and cost for large-scale deployments. The model is designed for production environments that require a balance of capability and efficiency, making it well suited for chat applications, coding assistants, and agent workflows that operate at scale. GPT-5.4 mini delivers reliable instruction following, solid multi-step reasoning, and consistent performance across diverse tasks with improved cost efficiency.

模型规格

上下文长度
400K
最大输出
128K
输入/输出模态
文本 / 图像
发布日期
2026-03

02

API 端点

单一 MixRoute 网关,OpenAI 兼容

  • OpenAI-compatible /v1/chat/completions POST

03

价格

固定费率,单位:/1M Tokens

输入

$0.7500 /1M Tokens

补全

$4.5000 /1M Tokens

缓存读取

$0.0750 /1M Tokens

05

选型摘要

快速判断是否适合你的工作负载

What is this model good for?

Start with the documented use cases for gpt-5.4-mini

GPT-5.4 mini brings the core capabilities of GPT-5.4 to a faster, more efficient model optimized for high-throughput workloads. It supports text and image inputs with strong performance across reasoning, coding, and tool use, while reducing latency and cost for large-scale deployments. The model is designed for production environments that require a balance of capability and efficiency, making it well suited for chat applications, coding assistants, and agent workflows that operate at scale. GPT-5.4 mini delivers reliable instruction following, solid multi-step reasoning, and consistent performance across diverse tasks with improved cost efficiency.

What should you check before using it?

Validate limits, pricing, and a representative workload

Review the current model limits and pricing record before production use.

How do you call it through MixRoute?

Use the documented endpoint and exact model ID

The model record lists OpenAI-compatible access via POST /v1/chat/completions with model ID gpt-5.4-mini.

08

Token 成本估算

基于本页价格的实时估算,非实际账单

同一前缀被重复读取的输入占比,最高 100%

单一端点,可验证的决策

使用相同的请求格式测试此模型和备用路由

从真实工作负载开始,再由质量、总成本和故障条件决定是否接入生产流量。