/ Models

Language models

gemma

Runs on your Mac · gemma

Runs on your own machine — free, offline, no key, and no cloud API for it here.

free

Gemma is a family of lightweight, state-of-the-art open models built by Google DeepMind. Updated to version 1.1

Install on 00

Get the 00 app

Free · macOS

Your agents, your models, your machine — this model installs with one click once the app is on board.

Download free →

What it costs

Input / 1M
free
Output / 1M
free
1M in + 1M out
free

Free to run — downloads once from the Ollama library, then runs on your own Mac with no per-token cost. It costs you disk, memory and time instead — none of which the library states.

What it can do

Takes
text
Does
answers in text
Sizes to pull
2b · 7b
Context window
Download size
Released

Blank because the library states none of them. Sizes are parameter counts, not disk size — memory depends on the quantisation you pull.

Run it

In the 00 app — pick it from the local models and the app pulls it; then it answers with no key, no account, no network. Get 00.

In a terminal — with Ollama installed:

ollama pull gemma

Add a size for a specific build — gemma:2b, gemma:7b. Without one you get the default tag.

It runs on your machine: no AI tokens billed, and it does not answer at the hosted endpoints. See it in the Ollama library · ← all models