/ Models

Language models

Qwen: Qwen3 235B A22B Thinking 2507

Qwen · qwen/qwen3-235b-a22b-thinking-2507

Runs in the cloud on our API and on your own machine — qwen3:235b on the local shelf, free and offline.

open weights tools always reasons

Qwen3-235B-A22B-Thinking-2507 is a high-performance, open-weight Mixture-of-Experts (MoE) language model optimized for complex reasoning tasks. It activates 22B of its 235B parameters per forward pass and natively supports up to 262,144...

Install on 00

Get the 00 app

Free · macOS

Your agents, your models, your machine — this model installs with one click once the app is on board.

Download free →

What it costs

Input / 1M
$0.46
Output / 1M
$4.60
1M in + 1M out
$5.06

Final AI-token prices, per million tokens. Bring your own provider key in the app to pay provider rates directly.

What it can do

Takes
text
Returns
text
Tools
yes
Reasoning
always on
Context window
131K tokens
Max output
118K tokens
Open weights
yes

Benchmarks

BenchmarkWhat it measuresScore
GPQA Diamondgraduate-level science questions80.1% ± 2.6
AIMEcompetition mathematics86.7% ± 5.1
SimpleQAshort factual questions — the test for making things up40.4% ± 1.6

Scores by Epoch AI, CC BY 4.0. The ± is the margin of error: models inside each other's margins are tied, not ranked. Epoch's notes on this row: source “epoch”.

Dates

Released
2025-07-25
Knowledge cutoff
2025-06-30
Retires
—

Using it

In the 00 app — pick it per agent, and per tier, then build your own agents on it. Get 00.

As a drop-in API — the same model at an OpenAI-compatible endpoint. Two settings:

  1. Create a workspace API key in the Overblast console — usage bills that workspace's AI tokens, at the prices on this page.
  2. Point your base URL at https://brain.deployd.network/ai/v1. Chat completions with streaming, embeddings, /images, /audio/transcriptions and /videos.
curl https://brain.deployd.network/ai/v1/chat/completions \
  -H "Authorization: Bearer ob_live_YOUR_KEY" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "qwen/qwen3-235b-a22b-thinking-2507",
    "messages": [{ "role": "user", "content": "Hello" }]
  }'

← all models · compare it with another · calling image, video and speech models