Model detail
gpt-4o-mini-tts
Model specs
- Context length
- 4.096K
- Max output
- –
- I/O modalities
- Audio
- Released
- –
02
API endpoints
One MixRoute gateway, OpenAI-compatible
-
OpenAI-compatible
/v1/chat/completionsPOST
03
Pricing
Flat rate, unit: /1M Tokens
Input
$0.6000 /1M Tokens
Completion
$0.0000 /1M Tokens
Audio input
$0.6000 /1M Tokens
Audio completion
$12.0000 /1M Tokens
05
Selection summary
Quickly judge whether it fits your workload
What is this model good for?
Natural text-to-speech with GPT-4o Mini
Use GPT-4o Mini TTS to convert text to natural-sounding spoken audio. Built on GPT-4o Mini, it supports up to 2000 input tokens for fast, cost-efficient speech synthesis.
What should you check before using it?
Confirm voice options and input limits
The maximum input is 2000 tokens. Check available voice options, supported languages, and output audio formats. Test with your specific text inputs.
Why use it through MixRoute?
Use the compatible endpoint with stable model ID
Use the confirmed compatible endpoint with model ID gpt-4o-mini-tts through MixRoute.
08
Token cost estimator
Live estimate from this page's pricing, not an actual bill
This model has no cache-read price; caching is not counted
09
FAQ
GPT-4o Mini TTS is a text-to-speech model built on GPT-4o Mini. It converts text to natural-sounding spoken audio with a maximum of 2000 input tokens.
One endpoint, a testable decision
Test this model and alternate routes with the same request format
Start from a real workload, then let quality, total cost, and failure conditions decide whether to send production traffic.