/ Models

Language models

Google: Gemini 3.1 Pro Preview Custom Tools

Google · google/gemini-3.1-pro-preview-customtools

Runs in the cloud, on our API — not on your own machine.

tools always reasons hears

Gemini 3.1 Pro Preview Custom Tools is a variant of Gemini 3.1 Pro that improves tool selection behavior by preventing overuse of a general bash tool when more efficient third-party...

Install on 00

Get the 00 app

Free · macOS

Your agents, your models, your machine — this model installs with one click once the app is on board.

Download free →

What it costs

Input / 1M
$4
Output / 1M
$24
1M in + 1M out
$28
Cache read / 1M
$0.4
Cache write / 1M
$0.75
Web search / call
$0.028
Audio in / 1M
$4

Audio is metered separately, per million audio tokens — not added to the text rates above.

Long prompts cost more

Prompt sizeIn / 1MOut / 1MCache read / 1M
Up to 200K$4$24$0.4
200K and above$8$36$0.8

The band follows the size of the prompt you send.

Final AI-token prices, per million tokens. Bring your own provider key in the app to pay provider rates directly.

What it can do

Takes
text, audio, images, video and files
Returns
text
Tools
yes
Reasoning
always on
Context window
1M tokens
Max output
66K tokens
Open weights
no

Effort: high, medium and low — medium by default. More effort means more thinking tokens, billed at the output rate.

Benchmarks

BenchmarkWhat it measuresScore
SWE-bench Verifiedreal GitHub issues fixed end to end75.6% ± 1.9

Scores by Epoch AI, CC BY 4.0. The ± is the margin of error: models inside each other's margins are tied, not ranked. Epoch's notes on this row: source “epoch”.

Dates

Released
2026-02-25
Knowledge cutoff
—
Retires
—

Using it

In the 00 app — pick it per agent, and per tier, then build your own agents on it. Get 00.

As a drop-in API — the same model at an OpenAI-compatible endpoint. Two settings:

  1. Create a workspace API key in the Overblast console — usage bills that workspace's AI tokens, at the prices on this page.
  2. Point your base URL at https://brain.deployd.network/ai/v1. Chat completions with streaming, embeddings, /images, /audio/transcriptions and /videos.
curl https://brain.deployd.network/ai/v1/chat/completions \
  -H "Authorization: Bearer ob_live_YOUR_KEY" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "google/gemini-3.1-pro-preview-customtools",
    "messages": [{ "role": "user", "content": "Hello" }]
  }'

It transcribes too — it takes audio, so the same key reaches it at https://brain.deployd.network/ai/v1/audio/transcriptions. No per-minute rate: a recording is billed as the tokens it becomes, at the input price above.

curl https://brain.deployd.network/ai/v1/audio/transcriptions \
  -H "Authorization: Bearer ob_live_YOUR_KEY" \
  -F file=@audio.mp3 \
  -F model=google/gemini-3.1-pro-preview-customtools

# → { "text": "…what was said…" }

← all models · compare it with another · calling image, video and speech models