/ Models

Language models

Qwen: Qwen3 VL 32B Instruct

Qwen3-VL-32B-Instruct is a large-scale multimodal vision-language model designed for high-precision understanding and reasoning across text, images, and video. With 32 billion parameters, it combines deep visual perception with advanced text...

Maker
Qwen
Context window
131K
Max output
33K
Input / 1M
$0.104
Output / 1M
$0.416
1M in + 1M out
$0.52

Everything else

Max output33K
Input / 1M$0.104
Output / 1M$0.416
Cache read / 1M
Cache write / 1M
1M in + 1M out$0.52
Takestext, image
Returnstext
Toolsyes
Reasoningno
Open weightsyes
Released2025-10-23
Knowledge cutoff
Retires
Catalog idqwen/qwen3-vl-32b-instruct

Using it

In 00, this model is picked per agent — and per tier, so the model that answers your customers need not be the one that writes your code. Get 00.

← all models