Skip to content

Model detail

gpt-5-2025-08-07

Provided by OpenAI
Pay-as-you-go

GPT-5 is OpenAI’s most advanced model, offering major improvements in reasoning, code quality, and user experience. It is optimized for complex tasks that require step-by-step reasoning, instruction following, and accuracy in high-stakes use cases. It supports test-time routing features and advanced prompt understanding, including user-specified intent like "think hard about this." Improvements include reductions in hallucination, sycophancy, and better performance in coding, writing, and health-related tasks.

Model specs

Context length
400K
Max output
128K
I/O modalities
Text / Image
Released
2025-08

02

API endpoints

One MixRoute gateway, OpenAI-compatible

  • OpenAI-compatible /v1/chat/completions POST

03

Pricing

Flat rate, unit: /1M Tokens

Input

$1.2500 /1M Tokens

Completion

$10.0000 /1M Tokens

Cache read

$0.1250 /1M Tokens

05

Selection summary

Quickly judge whether it fits your workload

What is this model good for?

Complex reasoning and high-stakes accuracy

Use GPT-5 for tasks requiring step-by-step reasoning, precise instruction following, and high accuracy. It excels at coding, writing, and multi-step problem solving with reduced hallucination.

What should you check before using it?

Test routing features and reasoning depth

GPT-5 supports test-time routing and user-specified reasoning intent. Test your prompts with appropriate reasoning directives and confirm the output quality meets your requirements.

Why use it through MixRoute?

Use the OpenAI-compatible endpoint with stable model ID

Use the confirmed compatible endpoint with model ID gpt-5 through MixRoute, then run the same integration tests used for the current client.

08

Token cost estimator

Live estimate from this page's pricing, not an actual bill

Share of the same prefix read repeatedly, up to 100%

09

FAQ

GPT-5 is OpenAI’s most advanced model, offering major improvements in reasoning, code quality, and user experience. It is optimized for complex tasks requiring step-by-step reasoning, instruction following, and accuracy, with reduced hallucination and better performance across coding, writing, and analysis.

One endpoint, a testable decision

Test this model and alternate routes with the same request format

Start from a real workload, then let quality, total cost, and failure conditions decide whether to send production traffic.