/ Models

Language models

Qwen: Qwen3 Coder Flash

Qwen · qwen/qwen3-coder-flash

Runs in the cloud, on our API — not on your own machine.

tools

Qwen3 Coder Flash is Alibaba's fast and cost efficient version of their proprietary Qwen3 Coder Plus. It is a powerful coding agent model specializing in autonomous programming via tool calling...

Install on 00

Get the 00 app

Free · macOS

Your agents, your models, your machine — this model installs with one click once the app is on board.

Download free →

What it costs

Input / 1M
$0.39
Output / 1M
$1.95
1M in + 1M out
$2.34
Cache read / 1M
$0.078
Cache write / 1M
$0.487

Long prompts cost more

Prompt sizeIn / 1MOut / 1MCache read / 1M
Up to 32K$0.39$1.95$0.078
32K – 128K$0.65$3.25$0.13
128K and above$1.04$5.20$0.208

The band follows the size of the prompt you send.

Final AI-token prices, per million tokens. Bring your own provider key in the app to pay provider rates directly.

What it can do

Takes
text
Returns
text
Tools
yes
Reasoning
no
Context window
1M tokens
Max output
66K tokens
Open weights
no

Dates

Released
2025-09-17
Knowledge cutoff
2025-06-30
Retires
—

Using it

In the 00 app — pick it per agent, and per tier, then build your own agents on it. Get 00.

As a drop-in API — the same model at an OpenAI-compatible endpoint. Two settings:

  1. Create a workspace API key in the Overblast console — usage bills that workspace's AI tokens, at the prices on this page.
  2. Point your base URL at https://brain.deployd.network/ai/v1. Chat completions with streaming, embeddings, /images, /audio/transcriptions and /videos.
curl https://brain.deployd.network/ai/v1/chat/completions \
  -H "Authorization: Bearer ob_live_YOUR_KEY" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "qwen/qwen3-coder-flash",
    "messages": [{ "role": "user", "content": "Hello" }]
  }'

← all models · compare it with another · calling image, video and speech models