Inception
Inception: Mercury Coder
Mercury Coder is the first diffusion large language model (dLLM). Applying a breakthrough discrete diffusion approach, the model runs 5-10x faster than even speed optimized models like Claude 3.5 Haiku and GPT-4o Mini while matching their performance. Mercury Coder's speed means that developers can stay in the flow while coding, enjoying rapid chat-based iteration and responsive code completion suggestions. On Copilot Arena, Mercury Coder ranks 1st in speed and ties for 2nd in quality. Read more in the [blog post here](https://www.inceptionlabs.ai/blog/introducing-mercury).
| Specifications | |
|---|---|
| Provider | Inception |
| Input pricing | $0.25/M tokens |
| Output pricing | $0.75/M tokens |
| Context length | 128K tokens |
| Max output | 32K tokens |
| Input modalities | text |
| Output modalities | text |
| Tokenizer | Other |
| Knowledge cutoff | 2025-01-31 |
| Content moderated | No |
Supported Parameters
max_tokensresponse_formatstopstructured_outputstemperaturetool_choicetools