Skip to content

Model detail

gpt-audio

Provided by OpenAI
Pay-as-you-go

Model specs

Context length
128K
Max output
16.384K
I/O modalities
Text
Released

02

API endpoints

One MixRoute gateway, OpenAI-compatible

  • OpenAI-compatible /v1/chat/completions POST

03

Pricing

Flat rate, unit: /1M Tokens

Input

$2.5000 /1M Tokens

Completion

$10.0000 /1M Tokens

Audio input

$40.0000 /1M Tokens

Audio completion

$5.0000 /1M Tokens

05

Selection summary

Quickly judge whether it fits your workload

What is this model good for?

Audio input and output via Chat Completions API

Use gpt-audio for tasks requiring audio input and output. It is OpenAI's first generally available audio model, accessible through the Chat Completions REST API.

What should you check before using it?

Confirm audio format support and pricing

Check supported audio formats, duration limits, and per-minute or per-token pricing. Test your specific audio tasks before production deployment.

Why use it through MixRoute?

Use the compatible endpoint with stable model ID

Use the confirmed compatible endpoint with model ID gpt-audio through MixRoute.

08

Token cost estimator

Live estimate from this page's pricing, not an actual bill

This model has no cache-read price; caching is not counted

09

FAQ

gpt-audio is OpenAI’s first generally available audio model. It accepts audio inputs and outputs and can be used in the Chat Completions REST API.

One endpoint, a testable decision

Test this model and alternate routes with the same request format

Start from a real workload, then let quality, total cost, and failure conditions decide whether to send production traffic.