Model detail
gpt-audio
Model specs
- Context length
- 128K
- Max output
- 16.384K
- I/O modalities
- Text
- Released
- –
02
API endpoints
One MixRoute gateway, OpenAI-compatible
-
OpenAI-compatible
/v1/chat/completionsPOST
03
Pricing
Flat rate, unit: /1M Tokens
Input
$2.5000 /1M Tokens
Completion
$10.0000 /1M Tokens
Audio input
$40.0000 /1M Tokens
Audio completion
$5.0000 /1M Tokens
05
Selection summary
Quickly judge whether it fits your workload
What is this model good for?
Audio input and output via Chat Completions API
Use gpt-audio for tasks requiring audio input and output. It is OpenAI's first generally available audio model, accessible through the Chat Completions REST API.
What should you check before using it?
Confirm audio format support and pricing
Check supported audio formats, duration limits, and per-minute or per-token pricing. Test your specific audio tasks before production deployment.
Why use it through MixRoute?
Use the compatible endpoint with stable model ID
Use the confirmed compatible endpoint with model ID gpt-audio through MixRoute.
08
Token cost estimator
Live estimate from this page's pricing, not an actual bill
This model has no cache-read price; caching is not counted
09
FAQ
gpt-audio is OpenAI’s first generally available audio model. It accepts audio inputs and outputs and can be used in the Chat Completions REST API.
One endpoint, a testable decision
Test this model and alternate routes with the same request format
Start from a real workload, then let quality, total cost, and failure conditions decide whether to send production traffic.