Google: Gemini 3.1 Pro Preview
Google · google/gemini-3.1-pro-preview
Runs in the cloud, on our API — not on your own machine.
tools always reasons hears
Gemini 3.1 Pro Preview is Google’s frontier reasoning model, delivering enhanced software engineering performance, improved agentic reliability, and more efficient token usage across complex workflows. Building on the multimodal foundation...
What it costs
Audio is metered separately, per million audio tokens — not added to the text rates above.
Long prompts cost more
| Prompt size | In / 1M | Out / 1M | Cache read / 1M |
|---|---|---|---|
| Up to 200K | $4 | $24 | $0.4 |
| 200K and above | $8 | $36 | $0.8 |
The band follows the size of the prompt you send.
Final AI-token prices, per million tokens. Bring your own provider key in the app to pay provider rates directly.
What it can do
Effort: high, medium and low — medium by default. More effort means more thinking tokens, billed at the output rate.
Benchmarks
| Benchmark | What it measures | Score |
|---|---|---|
| GPQA Diamond | graduate-level science questions | 94.4% ± 1.6 |
| AIME | competition mathematics | 95.6% ± 3.1 |
| FrontierMath | research-level mathematics | 36.9% ± 2.8 |
| SimpleQA | short factual questions — the test for making things up | 73.5% ± 1.4 |
Scores by Epoch AI, CC BY 4.0. The ± is the margin of error: models inside each other's margins are tied, not ranked. Epoch's notes on this row: source “epoch”.
Dates
Using it
In the 00 app — pick it per agent, and per tier, then build your own agents on it. Get 00.
As a drop-in API — the same model at an OpenAI-compatible endpoint. Two settings:
- Create a workspace API key in the Overblast console — usage bills that workspace's AI tokens, at the prices on this page.
- Point your base URL at
https://brain.deployd.network/ai/v1. Chat completions with streaming, embeddings,/images,/audio/transcriptionsand/videos.
curl https://brain.deployd.network/ai/v1/chat/completions \
-H "Authorization: Bearer ob_live_YOUR_KEY" \
-H "Content-Type: application/json" \
-d '{
"model": "google/gemini-3.1-pro-preview",
"messages": [{ "role": "user", "content": "Hello" }]
}'
It transcribes too — it takes audio, so the same key reaches it at
https://brain.deployd.network/ai/v1/audio/transcriptions. No per-minute rate: a recording is billed as the tokens it
becomes, at the input price above.
curl https://brain.deployd.network/ai/v1/audio/transcriptions \
-H "Authorization: Bearer ob_live_YOUR_KEY" \
-F file=@audio.mp3 \
-F model=google/gemini-3.1-pro-preview
# → { "text": "…what was said…" }
← all models · compare it with another · calling image, video and speech models
Subscribe to new models