/ Models

Language models

Inception: Mercury 2

Mercury 2 is an extremely fast reasoning LLM, and the first reasoning diffusion LLM (dLLM). Instead of generating tokens sequentially, Mercury 2 produces and refines multiple tokens in parallel, achieving...

Maker
Inception
Context window
128K
Max output
50K
Input / 1M
$0.25
Output / 1M
$0.75
1M in + 1M out
$1

Everything else

Max output50K
Input / 1M$0.25
Output / 1M$0.75
Cache read / 1M$0.025
Cache write / 1M
1M in + 1M out$1
Takestext
Returnstext
Toolsyes
Reasoningoptional · high, medium, low, none
Open weightsno
Released2026-03-04
Knowledge cutoff
Retires
Catalog idinception/mercury-2

Using it

In 00, this model is picked per agent — and per tier, so the model that answers your customers need not be the one that writes your code. Get 00.

← all models