/ Models

Language models

Ling-3.0-flash (free)

*Ling-3.0-flash* is a *124B-parameter Mixture-of-Experts (MoE) model*, with approximately *5.1B parameters activated per token*. The model is designed with *token efficiency and production-scale agentic inference* as key priorities, enabling developers...

Maker
Inclusionai
Context window
262K
Max output
33K
Input / 1M
free
Output / 1M
free
1M in + 1M out
free

Everything else

Max output33K
Input / 1Mfree
Output / 1Mfree
Cache read / 1M
Cache write / 1M
1M in + 1M outfree
Takestext
Returnstext
Toolsyes
Reasoningoptional
Open weightsno
Released2026-07-23
Knowledge cutoff
Retires
Catalog idinclusionai/ling-3.0-flash:free

Using it

In 00, this model is picked per agent — and per tier, so the model that answers your customers need not be the one that writes your code. Get 00.

← all models