模型详情
gpt-4.1-2025-04-14
GPT-4.1 is a flagship large language model optimized for advanced instruction following, real-world software engineering, and long-context reasoning. It supports a 1 million token context window and outperforms GPT-4o and GPT-4.5 across coding (54.6% SWE-bench Verified), instruction compliance (87.4% IFEval), and multimodal understanding benchmarks. It is tuned for precise code diffs, agent reliability, and high recall in large document contexts, making it ideal for agents, IDE tooling, and enterprise knowledge retrieval.
模型规格
- 上下文长度
- 1.047576M
- 最大输出
- 32.768K
- 输入/输出模态
- 文本 / 图像
- 发布日期
- 2025-04
02
API 端点
单一 MixRoute 网关,OpenAI 兼容
-
OpenAI-compatible
/v1/chat/completionsPOST
03
价格
固定费率,单位:/1M Tokens
输入
$2.0000 /1M Tokens
补全
$4.0000 /1M Tokens
05
选型摘要
快速判断是否适合你的工作负载
What is this model good for?
Long-context reasoning and software engineering
Use GPT-4.1 for tasks requiring a 1M token context window, including long-document analysis, complex codebases, and enterprise knowledge retrieval. It excels at instruction following and agent reliability.
What should you check before using it?
Plan for the 1M context window and instruction-following needs
Confirm your workload benefits from the 1M token context window. Test instruction-following accuracy on your specific prompts and verify agent reliability for your most demanding workflows.
Why use it through MixRoute?
Use the OpenAI-compatible endpoint with stable model ID
Use the confirmed compatible endpoint with model ID gpt-4.1 through MixRoute, then run the same integration tests used for the current client.
08
Token 成本估算
基于本页价格的实时估算,非实际账单
此模型没有缓存读取价格,不计算缓存费用
单一端点,可验证的决策
使用相同的请求格式测试此模型和备用路由
从真实工作负载开始,再由质量、总成本和故障条件决定是否接入生产流量。