Skip to content

Model detail

gpt-4o-mini-tts

Provided by OpenAI
Pay-as-you-go

Model specs

Context length
4.096K
Max output
I/O modalities
Audio
Released

02

API endpoints

One MixRoute gateway, OpenAI-compatible

  • OpenAI-compatible /v1/chat/completions POST

03

Pricing

Flat rate, unit: /1M Tokens

Input

$0.6000 /1M Tokens

Completion

$0.0000 /1M Tokens

Audio input

$0.6000 /1M Tokens

Audio completion

$12.0000 /1M Tokens

05

Selection summary

Quickly judge whether it fits your workload

What is this model good for?

Natural text-to-speech with GPT-4o Mini

Use GPT-4o Mini TTS to convert text to natural-sounding spoken audio. Built on GPT-4o Mini, it supports up to 2000 input tokens for fast, cost-efficient speech synthesis.

What should you check before using it?

Confirm voice options and input limits

The maximum input is 2000 tokens. Check available voice options, supported languages, and output audio formats. Test with your specific text inputs.

Why use it through MixRoute?

Use the compatible endpoint with stable model ID

Use the confirmed compatible endpoint with model ID gpt-4o-mini-tts through MixRoute.

08

Token cost estimator

Live estimate from this page's pricing, not an actual bill

This model has no cache-read price; caching is not counted

09

FAQ

GPT-4o Mini TTS is a text-to-speech model built on GPT-4o Mini. It converts text to natural-sounding spoken audio with a maximum of 2000 input tokens.

One endpoint, a testable decision

Test this model and alternate routes with the same request format

Start from a real workload, then let quality, total cost, and failure conditions decide whether to send production traffic.