/ Models

← Media models

OpenAI: GPT Audio

The gpt-audio model is OpenAI's first generally available audio model. The new snapshot features an upgraded decoder for more natural sounding voices and maintains better voice consistency. Audio is priced...

We have not generated an example for this one yet. Everything shown on this site is ours and hosted by us, so where there is no example there is simply nothing to show.

Takes
text, audio
Makes
speech
Release date
2026-01-19
Licence
commercial

What it costs

$2.5 per 1M tokens in · $10 per 1M out

Where it runs

A language model that returns media — billed per token like any other model, on the same key.

Full model page →

← every media model