跳至主要內容

模型詳情

o3-mini

由 OpenAI 提供
隨用隨付

OpenAI o3-mini is a cost-efficient language model optimized for STEM reasoning tasks, particularly excelling in science, mathematics, and coding. This model supports the `reasoning_effort` parameter, which can be set to "high", "medium", or "low" to control the thinking time of the model. The default is "medium". OpenRouter also offers the model slug `openai/o3-mini-high` to default the parameter to "high". The model features three adjustable reasoning effort levels and supports key developer capabilities including function calling, structured outputs, and streaming, though it does not include vision processing capabilities. The model demonstrates significant improvements over its predecessor, with expert testers preferring its responses 56% of the time and noting a 39% reduction in major errors on complex questions. With medium reasoning effort settings, o3-mini matches the performance of the larger o1 model on challenging reasoning evaluations like AIME and GPQA, while maintaining lower latency and cost.

模型規格

上下文長度
200K
最大輸出
100K
輸入/輸出模態
文字
發布日期
2025-01

02

API 端點

單一 MixRoute 閘道,OpenAI 相容

  • OpenAI-compatible /v1/chat/completions POST

03

價格

固定費率,單位:/1M Tokens

輸入

$1.1000 /1M Tokens

補全

$4.4000 /1M Tokens

快取讀取

$0.5500 /1M Tokens

05

選型摘要

快速判斷是否適合你的工作負載

What is this model good for?

Cost-efficient STEM reasoning with adjustable effort

Use o3-mini for science, math, and coding tasks where you need strong reasoning at lower cost. It supports adjustable reasoning_effort (low/medium/high) and matches o1 performance on AIME and GPQA at medium effort.

What should you check before using it?

Tune reasoning_effort for your task

Set reasoning_effort to low, medium, or high based on task complexity. Test function calling and structured outputs. Note that o3-mini does not support vision processing.

Why use it through MixRoute?

Use the compatible endpoint with stable model ID

Use the confirmed compatible endpoint with model ID o3-mini through MixRoute, then run the same integration tests used for the current client.

08

Token 成本估算

基於本頁價格的即時估算,非實際帳單

同一前綴被重複讀取的輸入占比,最高 100%

單一端點,可驗證的決策

使用相同的請求格式測試此模型與替代路由

從真實工作負載開始,再由品質、總成本與故障條件決定是否導入正式環境流量。