NVIDIA: Nemotron 3 Nano Omni (free)
Nvidia · nvidia/nemotron-3-nano-omni-30b-a3b-reasoning:free
Runs in the cloud, on our API — not on your own machine.
open weights free tools reasoning optional hears
NVIDIA Nemotron™ 3 Nano Omni is a 30B-A3B open multimodal model designed to function as a perception and context sub-agent in enterprise agent systems. It accepts text, image, video, and...
What it costs
Audio input has no separate rate: a recording counts against the input price above.
Final AI-token prices, per million tokens. Bring your own provider key in the app to pay provider rates directly.
What it can do
Dates
Using it
In the 00 app — pick it per agent, and per tier, then build your own agents on it. Get 00.
As a drop-in API — the same model at an OpenAI-compatible endpoint. Two settings:
- Create a workspace API key in the Overblast console — usage bills that workspace's AI tokens, at the prices on this page.
- Point your base URL at
https://brain.deployd.network/ai/v1. Chat completions with streaming, embeddings,/images,/audio/transcriptionsand/videos.
curl https://brain.deployd.network/ai/v1/chat/completions \
-H "Authorization: Bearer ob_live_YOUR_KEY" \
-H "Content-Type: application/json" \
-d '{
"model": "nvidia/nemotron-3-nano-omni-30b-a3b-reasoning:free",
"messages": [{ "role": "user", "content": "Hello" }]
}'
It transcribes too — it takes audio, so the same key reaches it at
https://brain.deployd.network/ai/v1/audio/transcriptions. No per-minute rate: a recording is billed as the tokens it
becomes, at the input price above.
curl https://brain.deployd.network/ai/v1/audio/transcriptions \
-H "Authorization: Bearer ob_live_YOUR_KEY" \
-F file=@audio.mp3 \
-F model=nvidia/nemotron-3-nano-omni-30b-a3b-reasoning:free
# → { "text": "…what was said…" }
← all models · compare it with another · calling image, video and speech models
Subscribe to new models