/ Models

Language models

OpenAI: GPT Audio Mini

Openai · openai/gpt-audio-mini

Runs in the cloud, on our API — not on your own machine.

tools speaks hears

A cost-efficient version of GPT Audio. The new snapshot features an upgraded decoder for more natural sounding voices and maintains better voice consistency. Input is priced at $0.60 per million...

Install on 00

Get the 00 app

Free · macOS

Your agents, your models, your machine — this model installs with one click once the app is on board.

Download free →

What it costs

Input / 1M
$1.20
Output / 1M
$4.80
1M in + 1M out
$6
Audio in / 1M
$1.20
Audio out / 1M
$4.80

Audio is metered separately, per million audio tokens — not added to the text rates above.

Final AI-token prices, per million tokens. Bring your own provider key in the app to pay provider rates directly.

What it can do

Takes
text and audio
Returns
text and audio
Tools
yes
Reasoning
no
Context window
128K tokens
Max output
16K tokens
Open weights
no

Dates

Released
2026-01-19
Knowledge cutoff
—
Retires
—

Using it

In the 00 app — pick it per agent, and per tier, then build your own agents on it. Get 00.

As a drop-in API — the same model at an OpenAI-compatible endpoint. Two settings:

  1. Create a workspace API key in the Overblast console — usage bills that workspace's AI tokens, at the prices on this page.
  2. Point your base URL at https://brain.deployd.network/ai/v1. Chat completions with streaming, embeddings, /images, /audio/transcriptions and /videos.
curl https://brain.deployd.network/ai/v1/chat/completions \
  -H "Authorization: Bearer ob_live_YOUR_KEY" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "openai/gpt-audio-mini",
    "messages": [{ "role": "user", "content": "Hello" }]
  }'

It transcribes too — it takes audio, so the same key reaches it at https://brain.deployd.network/ai/v1/audio/transcriptions. No per-minute rate: a recording is billed as the tokens it becomes, at the input price above.

curl https://brain.deployd.network/ai/v1/audio/transcriptions \
  -H "Authorization: Bearer ob_live_YOUR_KEY" \
  -F file=@audio.mp3 \
  -F model=openai/gpt-audio-mini

# → { "text": "…what was said…" }

It speaks — ask the same chat endpoint for the audio modality; the reply carries a base64 clip. Voices and formats are the model's own.

curl https://brain.deployd.network/ai/v1/chat/completions \
  -H "Authorization: Bearer ob_live_YOUR_KEY" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "openai/gpt-audio-mini",
    "modalities": ["text", "audio"],
    "audio": { "voice": "alloy", "format": "mp3" },
    "messages": [{ "role": "user", "content": "Read this aloud: your order is on its way." }]
  }'

# → base64 audio in choices[0].message.audio.data

← all models · compare it with another · calling image, video and speech models