Model detail
gpt-3.5-turbo-0125
The latest GPT-3.5 Turbo model with improved instruction following, JSON mode, reproducible outputs, parallel function calling, and more. Training data: up to Sep 2021. This version has a higher accuracy at responding in requested formats and a fix for a bug which caused a text encoding issue for non-English language function calls.
Model specs
- Context length
- 16.385K
- Max output
- 4.096K
- I/O modalities
- Text
- Released
- 2023-05
02
API endpoints
One MixRoute gateway, OpenAI-compatible
-
OpenAI-compatible
/v1/chat/completionsPOST
03
Pricing
Flat rate, unit: /1M Tokens
Input
$0.5000 /1M Tokens
Completion
$1.0000 /1M Tokens
05
Selection summary
Quickly judge whether it fits your workload
What is this model good for?
Latest GPT-3.5 Turbo with improved instruction following
Use gpt-3.5-turbo-0125 for tasks requiring JSON mode, reproducible outputs, and parallel function calling. It is the latest GPT-3.5 Turbo variant with improved instruction following.
What should you check before using it?
Consider GPT-4o Mini for better performance
GPT-4o Mini offers significantly better intelligence at a comparable cost. Test your tasks to determine if GPT-3.5 Turbo still meets your needs.
Why use it through MixRoute?
Use the OpenAI-compatible endpoint with stable model ID
Use the confirmed compatible endpoint with model ID gpt-3.5-turbo-0125 through MixRoute.
08
Token cost estimator
Live estimate from this page's pricing, not an actual bill
This model has no cache-read price; caching is not counted
09
FAQ
gpt-3.5-turbo-0125 is the latest GPT-3.5 Turbo model with improved instruction following, JSON mode, reproducible outputs, and parallel function calling.
One endpoint, a testable decision
Test this model and alternate routes with the same request format
Start from a real workload, then let quality, total cost, and failure conditions decide whether to send production traffic.