コンテンツへスキップ

モデル詳細

gpt-5.4-mini

OpenAI 提供
従量課金

GPT-5.4 mini brings the core capabilities of GPT-5.4 to a faster, more efficient model optimized for high-throughput workloads. It supports text and image inputs with strong performance across reasoning, coding, and tool use, while reducing latency and cost for large-scale deployments. The model is designed for production environments that require a balance of capability and efficiency, making it well suited for chat applications, coding assistants, and agent workflows that operate at scale. GPT-5.4 mini delivers reliable instruction following, solid multi-step reasoning, and consistent performance across diverse tasks with improved cost efficiency.

モデル仕様

コンテキスト長
400K
最大出力
128K
入出力モダリティ
テキスト / 画像
リリース
2026-03

02

API エンドポイント

MixRoute 単一ゲートウェイ、OpenAI 互換

  • OpenAI-compatible /v1/chat/completions POST

03

料金

固定料金、単位: /1M Tokens

入力

$0.7500 /1M Tokens

生成

$4.5000 /1M Tokens

キャッシュ読み取り

$0.0750 /1M Tokens

05

選定サマリー

ワークロード適合性を素早く判断

What is this model good for?

Start with the documented use cases for gpt-5.4-mini

GPT-5.4 mini brings the core capabilities of GPT-5.4 to a faster, more efficient model optimized for high-throughput workloads. It supports text and image inputs with strong performance across reasoning, coding, and tool use, while reducing latency and cost for large-scale deployments. The model is designed for production environments that require a balance of capability and efficiency, making it well suited for chat applications, coding assistants, and agent workflows that operate at scale. GPT-5.4 mini delivers reliable instruction following, solid multi-step reasoning, and consistent performance across diverse tasks with improved cost efficiency.

What should you check before using it?

Validate limits, pricing, and a representative workload

Review the current model limits and pricing record before production use.

How do you call it through MixRoute?

Use the documented endpoint and exact model ID

The model record lists OpenAI-compatible access via POST /v1/chat/completions with model ID gpt-5.4-mini.

08

トークンコスト見積もり

このページの料金に基づく見積もりで、実請求ではありません

同一プレフィックスが繰り返し読み取られる入力の割合(最大100%)

単一エンドポイントで検証可能な判断

同じリクエスト形式でこのモデルと代替ルートをテスト

実際のワークロードから始め、品質、総コスト、障害条件を基に本番トラフィックを送るか判断します。