模型详情
o3-mini-2025-01-31
OpenAI o3-mini is a cost-efficient language model optimized for STEM reasoning tasks, particularly excelling in science, mathematics, and coding. This model supports the `reasoning_effort` parameter, which can be set to "high", "medium", or "low" to control the thinking time of the model. The default is "medium". OpenRouter also offers the model slug `openai/o3-mini-high` to default the parameter to "high". The model features three adjustable reasoning effort levels and supports key developer capabilities including function calling, structured outputs, and streaming, though it does not include vision processing capabilities. The model demonstrates significant improvements over its predecessor, with expert testers preferring its responses 56% of the time and noting a 39% reduction in major errors on complex questions. With medium reasoning effort settings, o3-mini matches the performance of the larger o1 model on challenging reasoning evaluations like AIME and GPQA, while maintaining lower latency and cost.
模型规格
- 上下文长度
- 200K
- 最大输出
- 100K
- 输入/输出模态
- 文本
- 发布日期
- 2025-01
02
API 端点
单一 MixRoute 网关,OpenAI 兼容
-
OpenAI-compatible
/v1/chat/completionsPOST
03
价格
固定费率,单位:/1M Tokens
输入
$1.1000 /1M Tokens
补全
$4.4000 /1M Tokens
缓存读取
$0.5500 /1M Tokens
05
选型摘要
快速判断是否适合你的工作负载
What is this model good for?
Cost-efficient STEM reasoning with adjustable effort
Use o3-mini for science, math, and coding tasks where you need strong reasoning at lower cost. It supports adjustable reasoning_effort (low/medium/high) and matches o1 performance on AIME and GPQA at medium effort.
What should you check before using it?
Tune reasoning_effort for your task
Set reasoning_effort to low, medium, or high based on task complexity. Test function calling and structured outputs. Note that o3-mini does not support vision processing.
Why use it through MixRoute?
Use the compatible endpoint with stable model ID
Use the confirmed compatible endpoint with model ID o3-mini through MixRoute, then run the same integration tests used for the current client.
08
Token 成本估算
基于本页价格的实时估算,非实际账单
同一前缀被重复读取的输入占比,最高 100%
单一端点,可验证的决策
使用相同的请求格式测试此模型和备用路由
从真实工作负载开始,再由质量、总成本和故障条件决定是否接入生产流量。