Ling-3.0-flash (free)
*Ling-3.0-flash* is a *124B-parameter Mixture-of-Experts (MoE) model*, with approximately *5.1B parameters activated per token*. The model is designed with *token efficiency and production-scale agentic inference* as key priorities, enabling developers...
Everything else
| Max output | 33K |
|---|---|
| Input / 1M | free |
| Output / 1M | free |
| Cache read / 1M | — |
| Cache write / 1M | — |
| 1M in + 1M out | free |
| Takes | text |
| Returns | text |
| Tools | yes |
| Reasoning | optional |
| Open weights | no |
| Released | 2026-07-23 |
| Knowledge cutoff | — |
| Retires | — |
| Catalog id | inclusionai/ling-3.0-flash:free |
Using it
In 00, this model is picked per agent — and per tier, so the model that answers your customers need not be the one that writes your code. Get 00.