Skip to content

Model detail

grok-4-20-reasoning

Provided by xAI
Pay-as-you-go Dynamic pricing 2 tiers

Model specs

Context length
Max output
I/O modalities
Released

02

API endpoints

One MixRoute gateway, OpenAI-compatible

  • OpenAI-compatible /v1/chat/completions POST

03

Tiered pricing

Unit: /1M Tokens

Tier Input /1M Tokens Output /1M Tokens Cache read /1M Tokens
standard Length ≤ 200K $1.2500 $2.5000 $0.2000
long_context Length > 200K $2.5000 $5.0000 $0.4000

08

Token cost estimator

Live estimate from this page's pricing, not an actual bill

Share of the same prefix read repeatedly, up to 100%

One endpoint, a testable decision

Test this model and alternate routes with the same request format

Start from a real workload, then let quality, total cost, and failure conditions decide whether to send production traffic.