Skip to content

Model detail

o3-mini-2025-01-31

Provided by OpenAI
Pay-as-you-go

OpenAI o3-mini is a cost-efficient language model optimized for STEM reasoning tasks, particularly excelling in science, mathematics, and coding. This model supports the `reasoning_effort` parameter, which can be set to "high", "medium", or "low" to control the thinking time of the model. The default is "medium". OpenRouter also offers the model slug `openai/o3-mini-high` to default the parameter to "high". The model features three adjustable reasoning effort levels and supports key developer capabilities including function calling, structured outputs, and streaming, though it does not include vision processing capabilities. The model demonstrates significant improvements over its predecessor, with expert testers preferring its responses 56% of the time and noting a 39% reduction in major errors on complex questions. With medium reasoning effort settings, o3-mini matches the performance of the larger o1 model on challenging reasoning evaluations like AIME and GPQA, while maintaining lower latency and cost.

Model specs

Context length
200K
Max output
100K
I/O modalities
Text
Released
2025-01

02

API endpoints

One MixRoute gateway, OpenAI-compatible

  • OpenAI-compatible /v1/chat/completions POST

03

Pricing

Flat rate, unit: /1M Tokens

Input

$1.1000 /1M Tokens

Completion

$4.4000 /1M Tokens

Cache read

$0.5500 /1M Tokens

05

Selection summary

Quickly judge whether it fits your workload

What is this model good for?

Cost-efficient STEM reasoning with adjustable effort

Use o3-mini for science, math, and coding tasks where you need strong reasoning at lower cost. It supports adjustable reasoning_effort (low/medium/high) and matches o1 performance on AIME and GPQA at medium effort.

What should you check before using it?

Tune reasoning_effort for your task

Set reasoning_effort to low, medium, or high based on task complexity. Test function calling and structured outputs. Note that o3-mini does not support vision processing.

Why use it through MixRoute?

Use the compatible endpoint with stable model ID

Use the confirmed compatible endpoint with model ID o3-mini through MixRoute, then run the same integration tests used for the current client.

08

Token cost estimator

Live estimate from this page's pricing, not an actual bill

Share of the same prefix read repeatedly, up to 100%

09

FAQ

o3-mini is a cost-efficient language model optimized for STEM reasoning, particularly science, math, and coding. It supports adjustable reasoning effort levels (low/medium/high) and key developer features including function calling, structured outputs, and streaming.

One endpoint, a testable decision

Test this model and alternate routes with the same request format

Start from a real workload, then let quality, total cost, and failure conditions decide whether to send production traffic.