模型詳情
gpt-5.3-codex
GPT-5.3-Codex is OpenAI’s most advanced agentic coding model, combining the frontier software engineering performance of GPT-5.2-Codex with the broader reasoning and professional knowledge capabilities of GPT-5.2. It achieves state-of-the-art results on SWE-Bench Pro and strong performance on Terminal-Bench 2.0 and OSWorld-Verified, reflecting improved multi-language coding, terminal proficiency, and real-world computer-use skills. The model is optimized for long-running, tool-using workflows and supports interactive steering during execution, making it suitable for complex development tasks, debugging, deployment, and iterative product work. Beyond coding, GPT-5.3-Codex performs strongly on structured knowledge-work benchmarks such as GDPval, supporting tasks like document drafting, spreadsheet analysis, slide creation, and operational research across domains. It is trained with enhanced cybersecurity awareness, including vulnerability identification capabilities, and deployed with additional safeguards for high-risk use cases. Compared to prior Codex models, it is more token-efficient and approximately 25% faster, targeting professional end-to-end workflows that span reasoning, execution, and computer interaction.
模型規格
- 上下文長度
- 200K
- 最大輸出
- 100K
- 輸入/輸出模態
- 文字
- 發布日期
- 2026-02
02
API 端點
單一 MixRoute 閘道,OpenAI 相容
-
OpenAI-compatible
/v1/chat/completionsPOST
03
價格
固定費率,單位:/1M Tokens
輸入
$1.7500 /1M Tokens
補全
$14.0000 /1M Tokens
05
選型摘要
快速判斷是否適合你的工作負載
What is this model good for?
Long-running agentic software engineering and computer-use work
Use gpt-5.3-codex for complex development, debugging, deployment, terminal-oriented workflows, and iterative work that uses tools.
What should you check before using it?
Evaluate tool control and completion quality end to end
Test instruction adherence, tool permissions, code review quality, test execution, and human oversight on a representative engineering task.
How should you start?
Pilot with one bounded repository workflow
Call model ID gpt-5.3-codex in a sandboxed development workflow before allowing it to act on broader engineering tasks.
08
Token 成本估算
基於本頁價格的即時估算,非實際帳單
此模型沒有快取讀取價格,不計算快取費用
單一端點,可驗證的決策
使用相同的請求格式測試此模型與替代路由
從真實工作負載開始,再由品質、總成本與故障條件決定是否導入正式環境流量。