模型详情
gpt-5.3-codex
GPT-5.3-Codex is OpenAI’s most advanced agentic coding model, combining the frontier software engineering performance of GPT-5.2-Codex with the broader reasoning and professional knowledge capabilities of GPT-5.2. It achieves state-of-the-art results on SWE-Bench Pro and strong performance on Terminal-Bench 2.0 and OSWorld-Verified, reflecting improved multi-language coding, terminal proficiency, and real-world computer-use skills. The model is optimized for long-running, tool-using workflows and supports interactive steering during execution, making it suitable for complex development tasks, debugging, deployment, and iterative product work. Beyond coding, GPT-5.3-Codex performs strongly on structured knowledge-work benchmarks such as GDPval, supporting tasks like document drafting, spreadsheet analysis, slide creation, and operational research across domains. It is trained with enhanced cybersecurity awareness, including vulnerability identification capabilities, and deployed with additional safeguards for high-risk use cases. Compared to prior Codex models, it is more token-efficient and approximately 25% faster, targeting professional end-to-end workflows that span reasoning, execution, and computer interaction.
模型规格
- 上下文长度
- 200K
- 最大输出
- 100K
- 输入/输出模态
- 文本
- 发布日期
- 2026-02
02
API 端点
单一 MixRoute 网关,OpenAI 兼容
-
OpenAI-compatible
/v1/chat/completionsPOST
03
价格
固定费率,单位:/1M Tokens
输入
$1.7500 /1M Tokens
补全
$14.0000 /1M Tokens
05
选型摘要
快速判断是否适合你的工作负载
What is this model good for?
Long-running agentic software engineering and computer-use work
Use gpt-5.3-codex for complex development, debugging, deployment, terminal-oriented workflows, and iterative work that uses tools.
What should you check before using it?
Evaluate tool control and completion quality end to end
Test instruction adherence, tool permissions, code review quality, test execution, and human oversight on a representative engineering task.
How should you start?
Pilot with one bounded repository workflow
Call model ID gpt-5.3-codex in a sandboxed development workflow before allowing it to act on broader engineering tasks.
08
Token 成本估算
基于本页价格的实时估算,非实际账单
此模型没有缓存读取价格,不计算缓存费用
单一端点,可验证的决策
使用相同的请求格式测试此模型和备用路由
从真实工作负载开始,再由质量、总成本和故障条件决定是否接入生产流量。