模型詳情
gpt-4o-mini-2024-07-18
GPT-4o mini is OpenAI's newest model after [GPT-4 Omni](/models/openai/gpt-4o), supporting both text and image inputs with text outputs. As their most advanced small model, it is many multiples more affordable than other recent frontier models, and more than 60% cheaper than [GPT-3.5 Turbo](/models/openai/gpt-3.5-turbo). It maintains SOTA intelligence, while being significantly more cost-effective. GPT-4o mini achieves an 82% score on MMLU and presently ranks higher than GPT-4 on chat preferences [common leaderboards](https://arena.lmsys.org/). Check out the [launch announcement](https://openai.com/index/gpt-4o-mini-advancing-cost-efficient-intelligence/) to learn more. #multimodal
模型規格
- 上下文長度
- 128K
- 最大輸出
- 16.384K
- 輸入/輸出模態
- 文字 / 圖片
- 發布日期
- 2024-07
02
API 端點
單一 MixRoute 閘道,OpenAI 相容
-
OpenAI-compatible
/v1/chat/completionsPOST
03
價格
固定費率,單位:/1M Tokens
輸入
$0.1500 /1M Tokens
補全
$0.6000 /1M Tokens
快取讀取
$0.0750 /1M Tokens
05
選型摘要
快速判斷是否適合你的工作負載
What is this model good for?
Cost-effective multimodal tasks at scale
Use GPT-4o mini for high-volume multimodal tasks where cost efficiency is critical. It supports text and image inputs, achieving strong benchmark scores while being over 60% cheaper than GPT-3.5 Turbo.
What should you check before using it?
Confirm performance on your specific task class
GPT-4o mini is a small model optimized for cost. Test your specific multimodal tasks to confirm the accuracy meets your requirements compared to larger models like GPT-4o.
Why use it through MixRoute?
Use the OpenAI-compatible endpoint with stable model ID
Use the confirmed compatible endpoint with model ID gpt-4o-mini through MixRoute, then run the same integration tests used for the current client.
08
Token 成本估算
基於本頁價格的即時估算,非實際帳單
同一前綴被重複讀取的輸入占比,最高 100%
單一端點,可驗證的決策
使用相同的請求格式測試此模型與替代路由
從真實工作負載開始,再由品質、總成本與故障條件決定是否導入正式環境流量。