Qwen: Qwen3 Coder Flash
Qwen · qwen/qwen3-coder-flash
Runs in the cloud, on our API — not on your own machine.
tools
Qwen3 Coder Flash is Alibaba's fast and cost efficient version of their proprietary Qwen3 Coder Plus. It is a powerful coding agent model specializing in autonomous programming via tool calling...
What it costs
Long prompts cost more
| Prompt size | In / 1M | Out / 1M | Cache read / 1M |
|---|---|---|---|
| Up to 32K | $0.39 | $1.95 | $0.078 |
| 32K – 128K | $0.65 | $3.25 | $0.13 |
| 128K and above | $1.04 | $5.20 | $0.208 |
The band follows the size of the prompt you send.
Final AI-token prices, per million tokens. Bring your own provider key in the app to pay provider rates directly.
What it can do
Dates
Using it
In the 00 app — pick it per agent, and per tier, then build your own agents on it. Get 00.
As a drop-in API — the same model at an OpenAI-compatible endpoint. Two settings:
- Create a workspace API key in the Overblast console — usage bills that workspace's AI tokens, at the prices on this page.
- Point your base URL at
https://brain.deployd.network/ai/v1. Chat completions with streaming, embeddings,/images,/audio/transcriptionsand/videos.
curl https://brain.deployd.network/ai/v1/chat/completions \
-H "Authorization: Bearer ob_live_YOUR_KEY" \
-H "Content-Type: application/json" \
-d '{
"model": "qwen/qwen3-coder-flash",
"messages": [{ "role": "user", "content": "Hello" }]
}'
← all models · compare it with another · calling image, video and speech models
Subscribe to new models