/ Models

Language models

Google: Gemma 4 26B A4B (free)

Google · google/gemma-4-26b-a4b-it:free

Runs in the cloud on our API and on your own machine — gemma4:26b on the local shelf, free and offline.

open weights free tools reasoning optional

Gemma 4 26B A4B IT is an instruction-tuned Mixture-of-Experts (MoE) model from Google DeepMind. Despite 25.2B total parameters, only 3.8B activate per token during inference — delivering near-31B quality at...

Install on 00

Get the 00 app

Free · macOS

Your agents, your models, your machine — this model installs with one click once the app is on board.

Download free →

What it costs

Input / 1M
free
Output / 1M
free
1M in + 1M out
free

Final AI-token prices, per million tokens. Bring your own provider key in the app to pay provider rates directly.

What it can do

Takes
images, text and video
Returns
text
Tools
yes
Reasoning
optional
Context window
262K tokens
Max output
33K tokens
Open weights
yes

Dates

Released
2026-04-03
Knowledge cutoff
—
Retires
—

Using it

In the 00 app — pick it per agent, and per tier, then build your own agents on it. Get 00.

As a drop-in API — the same model at an OpenAI-compatible endpoint. Two settings:

  1. Create a workspace API key in the Overblast console — usage bills that workspace's AI tokens, at the prices on this page.
  2. Point your base URL at https://brain.deployd.network/ai/v1. Chat completions with streaming, embeddings, /images, /audio/transcriptions and /videos.
curl https://brain.deployd.network/ai/v1/chat/completions \
  -H "Authorization: Bearer ob_live_YOUR_KEY" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "google/gemma-4-26b-a4b-it:free",
    "messages": [{ "role": "user", "content": "Hello" }]
  }'

← all models · compare it with another · calling image, video and speech models