Google: Gemini 2.5 Flash Lite (batch)
Gemini 2.5 Flash-Lite is a lightweight reasoning model in the Gemini 2.5 family, optimized for ultra-low latency and cost efficiency. It offers improved throughput, faster token generation, and better performance...
Everything else
| Max output | 66K |
|---|---|
| Input / 1M | $0.05 |
| Output / 1M | $0.2 |
| Cache read / 1M | $0.01 |
| Cache write / 1M | — |
| 1M in + 1M out | $0.25 |
| Takes | text, image, file, audio, video |
| Returns | text |
| Tools | yes |
| Reasoning | optional |
| Open weights | no |
| Released | 2025-07-22 |
| Knowledge cutoff | 2025-01-31 |
| Retires | — |
| Catalog id | google/gemini-2.5-flash-lite:batch |
Using it
In 00, this model is picked per agent — and per tier, so the model that answers your customers need not be the one that writes your code. Get 00.