Skip to content

Model detail

gpt-3.5-turbo-0125

Provided by OpenAI
Pay-as-you-go

The latest GPT-3.5 Turbo model with improved instruction following, JSON mode, reproducible outputs, parallel function calling, and more. Training data: up to Sep 2021. This version has a higher accuracy at responding in requested formats and a fix for a bug which caused a text encoding issue for non-English language function calls.

Model specs

Context length
16.385K
Max output
4.096K
I/O modalities
Text
Released
2023-05

02

API endpoints

One MixRoute gateway, OpenAI-compatible

  • OpenAI-compatible /v1/chat/completions POST

03

Pricing

Flat rate, unit: /1M Tokens

Input

$0.5000 /1M Tokens

Completion

$1.0000 /1M Tokens

05

Selection summary

Quickly judge whether it fits your workload

What is this model good for?

Latest GPT-3.5 Turbo with improved instruction following

Use gpt-3.5-turbo-0125 for tasks requiring JSON mode, reproducible outputs, and parallel function calling. It is the latest GPT-3.5 Turbo variant with improved instruction following.

What should you check before using it?

Consider GPT-4o Mini for better performance

GPT-4o Mini offers significantly better intelligence at a comparable cost. Test your tasks to determine if GPT-3.5 Turbo still meets your needs.

Why use it through MixRoute?

Use the OpenAI-compatible endpoint with stable model ID

Use the confirmed compatible endpoint with model ID gpt-3.5-turbo-0125 through MixRoute.

08

Token cost estimator

Live estimate from this page's pricing, not an actual bill

This model has no cache-read price; caching is not counted

09

FAQ

gpt-3.5-turbo-0125 is the latest GPT-3.5 Turbo model with improved instruction following, JSON mode, reproducible outputs, and parallel function calling.

One endpoint, a testable decision

Test this model and alternate routes with the same request format

Start from a real workload, then let quality, total cost, and failure conditions decide whether to send production traffic.