OpenAI: GPT Audio
The gpt-audio model is OpenAI's first generally available audio model. The new snapshot features an upgraded decoder for more natural sounding voices and maintains better voice consistency. Audio is priced...
Everything else
| Max output | 16K |
|---|---|
| Input / 1M | $2.50 |
| Output / 1M | $10 |
| Cache read / 1M | — |
| Cache write / 1M | — |
| 1M in + 1M out | $12.50 |
| Takes | text, audio |
| Returns | text, audio |
| Tools | yes |
| Reasoning | no |
| Open weights | no |
| Released | 2026-01-19 |
| Knowledge cutoff | — |
| Retires | — |
| Catalog id | openai/gpt-audio |
Using it
In 00, this model is picked per agent — and per tier, so the model that answers your customers need not be the one that writes your code. Get 00.