Qwen: Qwen3.6 Flash
Qwen3.6 Flash is a fast, efficient language model from Alibaba's Qwen 3.6 series. It supports text, image, and video input with a 1M token context window. Tiered pricing kicks in...
Everything else
| Max output | 66K |
|---|---|
| Input / 1M | $0.188 |
| Output / 1M | $1.13 |
| Cache read / 1M | — |
| Cache write / 1M | $0.234 |
| 1M in + 1M out | $1.31 |
| Takes | text, image, video |
| Returns | text |
| Tools | yes |
| Reasoning | optional |
| Open weights | no |
| Released | 2026-04-27 |
| Knowledge cutoff | — |
| Retires | — |
| Catalog id | qwen/qwen3.6-flash |
Benchmarks
| source | epoch |
|---|---|
| eci | 144.2 |
| confidence | Unverified |
| gpqa | 0.844 |
| gpqaStderr | 0.0211 |
| frontierMath | 0.103 |
| frontierMathStderr | 0.0179 |
| aime | 0.861 |
| aimeStderr | 0.0474 |
| simpleQA | 0.212 |
| simpleQAStderr | 0.0129 |
Benchmark scores by Epoch AI, CC BY 4.0.
Using it
In 00, this model is picked per agent — and per tier, so the model that answers your customers need not be the one that writes your code. Get 00.