Inception: Mercury 2
inception/mercury-2
Mercury 2 is an extremely fast reasoning LLM, and the first reasoning diffusion LLM (dLLM). Instead of generating tokens sequentially, Mercury 2 produces and refines multiple tokens in parallel, achieving...
Modalities
Text
In / out price
₩439 / ₩1,316
1M
Context
128K
Released
Mar 4, 2026
Providers
Several places host the same model. Requests are spread by priority and health, and a local GPU always goes first when one is available.
| Provider | Context | Quantization | Max output | Priority | State |
|---|---|---|---|---|---|
| OpenRouter | 128K | — | — | 50 | unknown |