Meta: Llama 3.3 70B Instruct
Meta-llama · meta-llama/llama-3.3-70b-instruct
Runs in the cloud on our API and on your own machine — llama3.3:70b on the local shelf, free and offline.
open weights tools
The Meta Llama 3.3 multilingual large language model (LLM) is a pretrained and instruction tuned generative model in 70B (text in/text out). The Llama 3.3 instruction tuned text only model...
What it costs
Final AI-token prices, per million tokens. Bring your own provider key in the app to pay provider rates directly.
What it can do
Benchmarks
| Benchmark | What it measures | Score |
|---|---|---|
| GPQA Diamond | graduate-level science questions | 47.4% ± 2.9 |
| AIME | competition mathematics | 5.1% ± 2.4 |
| MATH level 5 | the hardest tier of the MATH problem set | 41.6% ± 1.2 |
Scores by Epoch AI, CC BY 4.0. The ± is the margin of error: models inside each other's margins are tied, not ranked. Epoch's notes on this row: source “epoch”.
Dates
Using it
In the 00 app — pick it per agent, and per tier, then build your own agents on it. Get 00.
As a drop-in API — the same model at an OpenAI-compatible endpoint. Two settings:
- Create a workspace API key in the Overblast console — usage bills that workspace's AI tokens, at the prices on this page.
- Point your base URL at
https://brain.deployd.network/ai/v1. Chat completions with streaming, embeddings,/images,/audio/transcriptionsand/videos.
curl https://brain.deployd.network/ai/v1/chat/completions \
-H "Authorization: Bearer ob_live_YOUR_KEY" \
-H "Content-Type: application/json" \
-d '{
"model": "meta-llama/llama-3.3-70b-instruct",
"messages": [{ "role": "user", "content": "Hello" }]
}'
← all models · compare it with another · calling image, video and speech models
Subscribe to new models