コンテンツへスキップ

モデル詳細

o4-mini-2025-04-16

OpenAI 提供
従量課金

OpenAI o4-mini is a compact reasoning model in the o-series, optimized for fast, cost-efficient performance while retaining strong multimodal and agentic capabilities. It supports tool use and demonstrates competitive reasoning and coding performance across benchmarks like AIME (99.5% with Python) and SWE-bench, outperforming its predecessor o3-mini and even approaching o3 in some domains. Despite its smaller size, o4-mini exhibits high accuracy in STEM tasks, visual problem solving (e.g., MathVista, MMMU), and code editing. It is especially well-suited for high-throughput scenarios where latency or cost is critical. Thanks to its efficient architecture and refined reinforcement learning training, o4-mini can chain tools, generate structured outputs, and solve multi-step tasks with minimal delay—often in under a minute.

モデル仕様

コンテキスト長
200K
最大出力
100K
入出力モダリティ
テキスト / 画像
リリース
2025-04

02

API エンドポイント

MixRoute 単一ゲートウェイ、OpenAI 互換

  • OpenAI-compatible /v1/chat/completions POST

03

料金

固定料金、単位: /1M Tokens

入力

$1.1000 /1M Tokens

生成

$1.1000 /1M Tokens

05

選定サマリー

ワークロード適合性を素早く判断

What is this model good for?

Fast, cost-efficient reasoning with tool use

Use o4-mini for high-throughput reasoning tasks where latency and cost are critical. It supports tool use, structured output, and multimodal inputs, making it suitable for agentic workflows at scale.

What should you check before using it?

Confirm tool use and structured output capabilities

Test your specific tool definitions, structured output schemas, and multimodal inputs. o4-mini is optimized for speed—verify it meets your accuracy requirements for your most demanding tasks.

Why use it through MixRoute?

Use the compatible endpoint with stable model ID

Use the confirmed compatible endpoint with model ID o4-mini through MixRoute, then run the same integration tests used for the current client.

08

トークンコスト見積もり

このページの料金に基づく見積もりで、実請求ではありません

このモデルにはキャッシュ読み取り価格がないため、キャッシュ料金は計算されません

単一エンドポイントで検証可能な判断

同じリクエスト形式でこのモデルと代替ルートをテスト

実際のワークロードから始め、品質、総コスト、障害条件を基に本番トラフィックを送るか判断します。