/ Models

Language models

nemotron-3-super

Runs on your Mac · nemotron-3-super

Runs on your own machine — free, offline, no key — and as a cloud API: NVIDIA: Nemotron 3 Super.

free tools thinking

NVIDIA Nemotron 3 Super is a 120B open MoE model activating just 12B parameters to deliver maximum compute efficiency and accuracy for complex multi-agent applications.

Install on 00

Get the 00 app

Free · macOS

Your agents, your models, your machine — this model installs with one click once the app is on board.

Download free →

What it costs

Input / 1M
free
Output / 1M
free
1M in + 1M out
free

Free to run — downloads once from the Ollama library, then runs on your own Mac with no per-token cost. It costs you disk, memory and time instead — none of which the library states.

What it can do

Takes
text
Does
calls tools and reasons before answering
Sizes to pull
120b
Context window
Download size
Released

Blank because the library states none of them. Sizes are parameter counts, not disk size — memory depends on the quantisation you pull.

Run it

In the 00 app — pick it from the local models and the app pulls it; then it answers with no key, no account, no network. Get 00.

In a terminal — with Ollama installed:

ollama pull nemotron-3-super

It runs on your machine: no AI tokens billed, and it does not answer at the hosted endpoints. See it in the Ollama library · ← all models