跳至主要內容

模型詳情

gpt-4.1

由 OpenAI 提供
隨用隨付

GPT-4.1 is a flagship large language model optimized for advanced instruction following, real-world software engineering, and long-context reasoning. It supports a 1 million token context window and outperforms GPT-4o and GPT-4.5 across coding (54.6% SWE-bench Verified), instruction compliance (87.4% IFEval), and multimodal understanding benchmarks. It is tuned for precise code diffs, agent reliability, and high recall in large document contexts, making it ideal for agents, IDE tooling, and enterprise knowledge retrieval.

模型規格

上下文長度
1.047576M
最大輸出
32.768K
輸入/輸出模態
文字 / 圖片
發布日期
2025-04

02

API 端點

單一 MixRoute 閘道,OpenAI 相容

  • OpenAI-compatible /v1/chat/completions POST

03

價格

固定費率,單位:/1M Tokens

輸入

$2.0000 /1M Tokens

補全

$4.0000 /1M Tokens

快取讀取

$0.5000 /1M Tokens

05

選型摘要

快速判斷是否適合你的工作負載

What is this model good for?

Long-context reasoning and software engineering

Use GPT-4.1 for tasks requiring a 1M token context window, including long-document analysis, complex codebases, and enterprise knowledge retrieval. It excels at instruction following and agent reliability.

What should you check before using it?

Plan for the 1M context window and instruction-following needs

Confirm your workload benefits from the 1M token context window. Test instruction-following accuracy on your specific prompts and verify agent reliability for your most demanding workflows.

Why use it through MixRoute?

Use the OpenAI-compatible endpoint with stable model ID

Use the confirmed compatible endpoint with model ID gpt-4.1 through MixRoute, then run the same integration tests used for the current client.

08

Token 成本估算

基於本頁價格的即時估算,非實際帳單

同一前綴被重複讀取的輸入占比,最高 100%

單一端點,可驗證的決策

使用相同的請求格式測試此模型與替代路由

從真實工作負載開始,再由品質、總成本與故障條件決定是否導入正式環境流量。