跳至主要內容

模型詳情

gpt-4o-mini

由 OpenAI 提供
隨用隨付

GPT-4o mini is OpenAI's newest model after [GPT-4 Omni](/models/openai/gpt-4o), supporting both text and image inputs with text outputs. As their most advanced small model, it is many multiples more affordable than other recent frontier models, and more than 60% cheaper than [GPT-3.5 Turbo](/models/openai/gpt-3.5-turbo). It maintains SOTA intelligence, while being significantly more cost-effective. GPT-4o mini achieves an 82% score on MMLU and presently ranks higher than GPT-4 on chat preferences [common leaderboards](https://arena.lmsys.org/). Check out the [launch announcement](https://openai.com/index/gpt-4o-mini-advancing-cost-efficient-intelligence/) to learn more. #multimodal

模型規格

上下文長度
128K
最大輸出
16.384K
輸入/輸出模態
文字 / 圖片
發布日期
2024-07

02

API 端點

單一 MixRoute 閘道,OpenAI 相容

  • OpenAI-compatible /v1/chat/completions POST

03

價格

固定費率,單位:/1M Tokens

輸入

$0.1500 /1M Tokens

補全

$0.6000 /1M Tokens

快取讀取

$0.0750 /1M Tokens

05

選型摘要

快速判斷是否適合你的工作負載

What is this model good for?

Cost-effective multimodal tasks at scale

Use GPT-4o mini for high-volume multimodal tasks where cost efficiency is critical. It supports text and image inputs, achieving strong benchmark scores while being over 60% cheaper than GPT-3.5 Turbo.

What should you check before using it?

Confirm performance on your specific task class

GPT-4o mini is a small model optimized for cost. Test your specific multimodal tasks to confirm the accuracy meets your requirements compared to larger models like GPT-4o.

Why use it through MixRoute?

Use the OpenAI-compatible endpoint with stable model ID

Use the confirmed compatible endpoint with model ID gpt-4o-mini through MixRoute, then run the same integration tests used for the current client.

08

Token 成本估算

基於本頁價格的即時估算,非實際帳單

同一前綴被重複讀取的輸入占比,最高 100%

單一端點,可驗證的決策

使用相同的請求格式測試此模型與替代路由

從真實工作負載開始,再由品質、總成本與故障條件決定是否導入正式環境流量。