Inception: Mercury 2
Mercury 2 is an extremely fast reasoning LLM, and the first reasoning diffusion LLM (dLLM). Instead of generating tokens sequentially, Mercury 2 produces and refines multiple tokens in parallel, achieving...
Everything else
| Max output | 50K |
|---|---|
| Input / 1M | $0.25 |
| Output / 1M | $0.75 |
| Cache read / 1M | $0.025 |
| Cache write / 1M | — |
| 1M in + 1M out | $1 |
| Takes | text |
| Returns | text |
| Tools | yes |
| Reasoning | optional · high, medium, low, none |
| Open weights | no |
| Released | 2026-03-04 |
| Knowledge cutoff | — |
| Retires | — |
| Catalog id | inception/mercury-2 |
Using it
In 00, this model is picked per agent — and per tier, so the model that answers your customers need not be the one that writes your code. Get 00.