/ Models

Language models

OpenAI: GPT Astra Latest

Openai · ~openai/gpt-astra-latest

Runs in the cloud, on our API — not on your own machine.

tools always reasons

This model always redirects to the latest model in the GPT Astra family.

What it costs

Input / 1M
$20
Output / 1M
$100
1M in + 1M out
$120
Cache read / 1M
$2
Cache write / 1M
$25
Web search / call
$0.02

Long prompts cost more

Prompt sizeIn / 1MOut / 1MCache read / 1M
Up to 272K$20$100$2
272K and above$40$150$4

The band follows the size of the prompt you send.

Final AI-token prices, per million tokens. Bring your own provider key in the app to pay provider rates directly.

What it can do

Takes
files, images and text
Returns
text
Tools
yes
Reasoning
always on
Context window
1.1M tokens
Max output
128K tokens
Open weights
no

Effort: max, xhigh, high, medium and low — medium by default. More effort means more thinking tokens, billed at the output rate.

Dates

Released
2026-09-11
Knowledge cutoff
Retires

Using it

In the 00 app — pick it per agent, and per tier, then build your own agents on it. Get 00.

As a drop-in API — the same model at an OpenAI-compatible endpoint. Two settings:

  1. Create a workspace API key in the Overblast console — usage bills that workspace's AI tokens, at the prices on this page.
  2. Point your base URL at https://brain.deployd.network/ai/v1. Chat completions with streaming, embeddings, /images, /audio/transcriptions and /videos.
curl https://brain.deployd.network/ai/v1/chat/completions \
  -H "Authorization: Bearer ob_live_YOUR_KEY" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "~openai/gpt-astra-latest",
    "messages": [{ "role": "user", "content": "Hello" }]
  }'

← all models · compare it with another · calling image, video and speech models