Model detail
gemini-2.5-pro
Gemini 2.5 Pro is Google’s state-of-the-art AI model designed for advanced reasoning, coding, mathematics, and scientific tasks. It employs “thinking” capabilities, enabling it to reason through responses with enhanced accuracy and nuanced context handling. Gemini 2.5 Pro achieves top-tier performance on multiple benchmarks, including first-place positioning on the LMArena leaderboard, reflecting superior human-preference alignment and complex problem-solving abilities.
Model specs
- Context length
- –
- Max output
- –
- I/O modalities
- –
- Released
- 2025-06
02
API endpoints
One MixRoute gateway, OpenAI-compatible
-
Gemini
/v1beta/models/{model}:generateContentPOST -
OpenAI-compatible
/v1/chat/completionsPOST
03
Pricing
Flat rate, unit: /1M Tokens
Input
$1.2500 /1M Tokens
Completion
$10.0000 /1M Tokens
Cache read
$0.1250 /1M Tokens
08
Token cost estimator
Live estimate from this page's pricing, not an actual bill
Share of the same prefix read repeatedly, up to 100%
One endpoint, a testable decision
Test this model and alternate routes with the same request format
Start from a real workload, then let quality, total cost, and failure conditions decide whether to send production traffic.