Qwen: Qwen3 VL 30B A3B Instruct
Qwen3-VL-30B-A3B-Instruct is a multimodal model that unifies strong text generation with visual understanding for images and videos. Its Instruct variant optimizes instruction-following for general multimodal tasks. It excels in perception...
Everything else
| Max output | 16K |
|---|---|
| Input / 1M | $0.15 |
| Output / 1M | $0.6 |
| Cache read / 1M | — |
| Cache write / 1M | — |
| 1M in + 1M out | $0.75 |
| Takes | text, image |
| Returns | text |
| Tools | yes |
| Reasoning | no |
| Open weights | yes |
| Released | 2025-10-06 |
| Knowledge cutoff | 2025-03-31 |
| Retires | — |
| Catalog id | qwen/qwen3-vl-30b-a3b-instruct |
Using it
In 00, this model is picked per agent — and per tier, so the model that answers your customers need not be the one that writes your code. Get 00.