Inception: Mercury 2
inception/mercury-2 · InceptionUnmeasurable- Declared family
- Unknown
- Entered catalogue
- 4 March 2026
- Measurement
- Unmeasurable: baseline overhead drifts between calls
- Evidence confidence
- Unknown
| We checked | What we found | Where it came from |
|---|---|---|
| How it counts tokens | Not measured yet | we measured it |
| How its API is set up | A combination of settings no other listing in the catalogue uses | api_d34d4a9e |
| How much it can read and write | Reads up to 128,000 tokens at once, writes up to 50,000 | its own listing |
| Whether it thinks before answering | Reasons when asked to, at high, medium, low, none, medium by default | its own listing |
Mercury 2 is an extremely fast reasoning LLM, and the first reasoning diffusion LLM (dLLM). Instead of generating tokens sequentially, Mercury 2 produces and refines multiple tokens in parallel, achieving...
Supported parameters
include_reasoningmax_tokensreasoningreasoning_effortresponse_formatstopstructured_outputstemperaturetool_choicetools
defaults: {"temperature":0.75,"top_p":null,"top_k":null,"frequency_penalty":null,"presence_penalty":null,"repetition_penalty":null}