Skip to content

Model detail

gpt-5.3-codex

Provided by OpenAI
Pay-as-you-go

GPT-5.3-Codex is OpenAI’s most advanced agentic coding model, combining the frontier software engineering performance of GPT-5.2-Codex with the broader reasoning and professional knowledge capabilities of GPT-5.2. It achieves state-of-the-art results on SWE-Bench Pro and strong performance on Terminal-Bench 2.0 and OSWorld-Verified, reflecting improved multi-language coding, terminal proficiency, and real-world computer-use skills. The model is optimized for long-running, tool-using workflows and supports interactive steering during execution, making it suitable for complex development tasks, debugging, deployment, and iterative product work. Beyond coding, GPT-5.3-Codex performs strongly on structured knowledge-work benchmarks such as GDPval, supporting tasks like document drafting, spreadsheet analysis, slide creation, and operational research across domains. It is trained with enhanced cybersecurity awareness, including vulnerability identification capabilities, and deployed with additional safeguards for high-risk use cases. Compared to prior Codex models, it is more token-efficient and approximately 25% faster, targeting professional end-to-end workflows that span reasoning, execution, and computer interaction.

Model specs

Context length
200K
Max output
100K
I/O modalities
Text
Released
2026-02

02

API endpoints

One MixRoute gateway, OpenAI-compatible

  • OpenAI-compatible /v1/chat/completions POST

03

Pricing

Flat rate, unit: /1M Tokens

Input

$1.7500 /1M Tokens

Completion

$14.0000 /1M Tokens

05

Selection summary

Quickly judge whether it fits your workload

What is this model good for?

Long-running agentic software engineering and computer-use work

Use gpt-5.3-codex for complex development, debugging, deployment, terminal-oriented workflows, and iterative work that uses tools.

What should you check before using it?

Evaluate tool control and completion quality end to end

Test instruction adherence, tool permissions, code review quality, test execution, and human oversight on a representative engineering task.

How should you start?

Pilot with one bounded repository workflow

Call model ID gpt-5.3-codex in a sandboxed development workflow before allowing it to act on broader engineering tasks.

08

Token cost estimator

Live estimate from this page's pricing, not an actual bill

This model has no cache-read price; caching is not counted

09

FAQ

gpt-5.3-codex is intended for long-running agentic software engineering, debugging, deployment, terminal-oriented workflows, and iterative computer-use tasks.

One endpoint, a testable decision

Test this model and alternate routes with the same request format

Start from a real workload, then let quality, total cost, and failure conditions decide whether to send production traffic.