Z.ai: GLM 4.7 Flash
As a 30B-class SOTA model, GLM-4.7-Flash offers a new option that balances performance and efficiency. It is further optimized for agentic coding use cases, strengthening coding capabilities, long-horizon task planning,...
Everything else
| Max output | 16K |
|---|---|
| Input / 1M | $0.06 |
| Output / 1M | $0.4 |
| Cache read / 1M | $0.01 |
| Cache write / 1M | — |
| 1M in + 1M out | $0.46 |
| Takes | text |
| Returns | text |
| Tools | yes |
| Reasoning | optional |
| Open weights | yes |
| Released | 2026-01-19 |
| Knowledge cutoff | — |
| Retires | — |
| Catalog id | z-ai/glm-4.7-flash |
Using it
In 00, this model is picked per agent — and per tier, so the model that answers your customers need not be the one that writes your code. Get 00.