BazaarLinkBazaarLink
Sign in
I

Mercury — Retired model

Retired modelThis model has reached end of life and is no longer available through BazaarLink. Historical specifications and independent benchmarks remain below. Retired on May 1, 2026.

Mercury is the first diffusion large language model (dLLM). Applying a breakthrough discrete diffusion approach, the model runs 5-10x faster than even speed optimized models like GPT-4.1 Nano and Claude 3.5 Haiku while matching their performance. Mercury's speed enables developers to provide responsive user experiences, including with voice agents, search interfaces, and chatbots. Read more in the [blog post] (https://www.inceptionlabs.ai/blog/introducing-mercury) here.

ProviderIinception
ReleasedJune 2025
Model IDinception/mercury

Technical specifications

Context window
128K tokens
Reasoning
Input
Text
Output
Text

Independent benchmarks

Scores below come from Artificial Analysis, an independent third party. BazaarLink does not participate in the testing.

Intelligence 12
Coding 31
Math
MMLU
GPQA 77
This modelCategory leader
Coding ability
#160out of 255 models
LiveCodeBench / SciCode and similar
Overall intelligence
#314out of 638 models
Aggregated across academic benchmarks
Time to first token
#218out of 329 models
AA 全球 p50:3.58 秒
Intelligence
12
Mid-tier
This model
12
Category leader
53
Median
12
Coding
31
Mid-tier
This model
31
Category leader
82
Median
45
GPQA
77%
This model
77%
Category leader
94%
Median
69%
HLE
16%
This model
16%
Category leader
53%
Median
7%
SciCode
39%
This model
39%
Category leader
60%
Median
34%
IFBench
70%
This model
70%
Category leader
83%
Median
44%
τ²-Bench Telecom
71%
This model
71%
Category leader
99%
Median
46%
AA-LCR
36%
This model
36%
Category leader
76%
Median
40%
Terminal-Bench Hard
27%
This model
27%
Category leader
66%
Median
14%

Top 5 — CodingCoding Index leaderboard

2Claude Fable 5.1 (Adaptive Reasoning, Xhigh Effort, Default Fallback)?/claude-fable-5-1-xhigh81$10 / $50
3Claude Fable 5.1 (Adaptive Reasoning, High Effort, Default Fallback)?/claude-fable-5-1-high79$10 / $50
4GPT-5.6 Sol (xhigh)openai/gpt-5-6-sol-xhigh78$4 / $20
AASource: Artificial Analysis · Independent third-party evaluation; not influenced by BazaarLinkView full evaluation on AA →
← All Models

Frequently Asked Questions

Can I still use the retired Mercury model?

No. Mercury has been retired and is retained here only as a historical reference. See the active alternatives linked on this page.

Relay integrity checks

No integrity-check data is available for this model yet.

BazaarLink Probe verifies endpoints that claim to serve this model and flags behavior that differs from the reference baseline.

View the full integrity report →

Related models and links

Related models and links

Qwen3.8 MaxDeepseek V3.2Glm 5.1Deepseek V4 FlashGlm 4.7Minimax M2.7
ModelsAPI documentationAPI latency probe
Support
Support
Hi! How can we help you?
Send a message and we'll get back to you soon.