Mercury Coder — Retired model
Mercury Coder is the first diffusion large language model (dLLM). Applying a breakthrough discrete diffusion approach, the model runs 5-10x faster than even speed optimized models like Claude 3.5 Haiku and GPT-4o Mini while matching their performance. Mercury Coder's speed means that developers can stay in the flow while coding, enjoying rapid chat-based iteration and responsive code completion suggestions. On Copilot Arena, Mercury Coder ranks 1st in speed and ties for 2nd in quality. Read more in the [blog post here](https://www.inceptionlabs.ai/blog/introducing-mercury).
Technical specifications
Frequently Asked Questions
Can I still use the retired Mercury Coder model?
No. Mercury Coder has been retired and is retained here only as a historical reference. See the active alternatives linked on this page.
Relay integrity checks
No integrity-check data is available for this model yet.
BazaarLink Probe verifies endpoints that claim to serve this model and flags behavior that differs from the reference baseline.
View the full integrity report →Related models and links
Related models and links