GPT Audio
OpenAI · Proprietary · Released Jan 19, 2026 · openai/gpt-audio
Best $2.50 in / $10.00 out via OpenAI
128K contextToolsStructured outputAudio input
Available from 1 providerUSD / MTOK · PRICES 23 MINUTES AGO · SPEED 3 DAYS AGO
EST. MONTHLY COSTclick a row for detail · click headers to sort
| PROVIDER | INPUT | OUTPUT | CACHE | CONTEXT | QUANT | TTFT | TPS | UPTIME |
|---|---|---|---|---|---|---|---|---|
| $2.50 | $10.00 | — | 128K | — | ✕ | ✕ | 100.0% |
TTFT/TPS: median (P50) over trailing 504h · chat10k profile (~10K in / 200 out) · measured via OpenRouter routing · ✕ = probed, endpoint returned no measurable stream · methodology v1.1 · measured 3 days ago
SERVE THIS MODEL AND NOT LISTED? LIST YOUR INFERENCE →
About
The gpt-audio model is OpenAI's first generally available audio model. The new snapshot features an upgraded decoder for more natural sounding voices and maintains better voice consistency. Audio is priced...
Specifications
Context window128,000
Max output16,384
Modalities intext, audio
Modalities outtext, audio
LicenseProprietary
ReleasedJan 19, 2026
TokenizerGPT
Related
Embed badge
[](https://modelindex.ai/models/openai/gpt-audio)Sources
OpenRouter API23 minutes ago
AGGREGATORLiteLLM dataset23 minutes ago
AGGREGATORModelIndex benchmarks3 days ago
MEASUREDVerified 23 minutes ago · list prices, not negotiated rates.