Back to Models
inception/mercury-coder
Not Available

Mercury Coder

Mercury Coder is the first diffusion large language model (dLLM). Applying a breakthrough discrete diffusion approach, the model runs 5-10x faster than even speed optimized models like Claude 3.5 Haiku and GPT-4o Mini while matching their performance. Mercury Coder's speed means that developers can stay in the flow while coding, enjoying rapid chat-based iteration and responsive code completion suggestions. On Copilot Arena, Mercury Coder ranks 1st in speed and ties for 2nd in quality. Read more in the [blog post here](https://www.inceptionlabs.ai/blog/introducing-mercury).

4/30/2025
128,000 tokens

Specifications

Modalities

Input
text
Output
text

Supported Parameters

frequency_penalty
max_tokens
presence_penalty
response_format
stop
structured_outputs
temperature
tool_choice
tools
top_k
top_p

Max Output Tokens

16,384

Frequently asked questions

Is Mercury Coder available on FastMetal?
Not at the moment. Mercury 2.5, from the same lab, is available on the FastMetal API today.
What is the context window of Mercury Coder?
128,000 tokens, shared between the prompt and the response.