/ Models

Language models

Qwen: Qwen3 VL 8B Instruct

Qwen3-VL-8B-Instruct is a multimodal vision-language model from the Qwen3-VL series, built for high-fidelity understanding and reasoning across text, images, and video. It features improved multimodal fusion with Interleaved-MRoPE for long-horizon...

Maker
Qwen
Context window
262K
Max output
33K
Input / 1M
$0.117
Output / 1M
$0.455
1M in + 1M out
$0.572

Everything else

Max output33K
Input / 1M$0.117
Output / 1M$0.455
Cache read / 1M
Cache write / 1M
1M in + 1M out$0.572
Takesimage, text
Returnstext
Toolsyes
Reasoningno
Open weightsyes
Released2025-10-14
Knowledge cutoff
Retires
Catalog idqwen/qwen3-vl-8b-instruct

Using it

In 00, this model is picked per agent — and per tier, so the model that answers your customers need not be the one that writes your code. Get 00.

← all models