Model detail
gpt-5.3-codex
GPT-5.3-Codex is OpenAI’s most advanced agentic coding model, combining the frontier software engineering performance of GPT-5.2-Codex with the broader reasoning and professional knowledge capabilities of GPT-5.2. It achieves state-of-the-art results on SWE-Bench Pro and strong performance on Terminal-Bench 2.0 and OSWorld-Verified, reflecting improved multi-language coding, terminal proficiency, and real-world computer-use skills. The model is optimized for long-running, tool-using workflows and supports interactive steering during execution, making it suitable for complex development tasks, debugging, deployment, and iterative product work. Beyond coding, GPT-5.3-Codex performs strongly on structured knowledge-work benchmarks such as GDPval, supporting tasks like document drafting, spreadsheet analysis, slide creation, and operational research across domains. It is trained with enhanced cybersecurity awareness, including vulnerability identification capabilities, and deployed with additional safeguards for high-risk use cases. Compared to prior Codex models, it is more token-efficient and approximately 25% faster, targeting professional end-to-end workflows that span reasoning, execution, and computer interaction.
Model specs
- Context length
- 200K
- Max output
- 100K
- I/O modalities
- Text
- Released
- 2026-02
02
API endpoints
One MixRoute gateway, OpenAI-compatible
-
OpenAI-compatible
/v1/chat/completionsPOST
03
Pricing
Flat rate, unit: /1M Tokens
Input
$1.7500 /1M Tokens
Completion
$14.0000 /1M Tokens
05
Selection summary
Quickly judge whether it fits your workload
What is this model good for?
Long-running agentic software engineering and computer-use work
Use gpt-5.3-codex for complex development, debugging, deployment, terminal-oriented workflows, and iterative work that uses tools.
What should you check before using it?
Evaluate tool control and completion quality end to end
Test instruction adherence, tool permissions, code review quality, test execution, and human oversight on a representative engineering task.
How should you start?
Pilot with one bounded repository workflow
Call model ID gpt-5.3-codex in a sandboxed development workflow before allowing it to act on broader engineering tasks.
08
Token cost estimator
Live estimate from this page's pricing, not an actual bill
This model has no cache-read price; caching is not counted
09
FAQ
gpt-5.3-codex is intended for long-running agentic software engineering, debugging, deployment, terminal-oriented workflows, and iterative computer-use tasks.
One endpoint, a testable decision
Test this model and alternate routes with the same request format
Start from a real workload, then let quality, total cost, and failure conditions decide whether to send production traffic.