Skip to content
Back to the index
43CodingProvisional score

Mercury 2

Mercury 2

A diffusion model for text, and it is several times faster than anything built the usual way.

Updated

Mercury 2

The read

Mercury generates text by refining a whole draft at once rather than one token after another, and the result is the highest measured output rate of any model in this index by a wide margin. For code completion and other places where the wait is the whole user experience, that architecture is a genuine advantage rather than a curiosity. Quality is good, not frontier, and the ecosystem around it is tiny.

Where it shines

  • By far the highest measured output speed of anything ranked here
  • Diffusion decoding refines a whole draft rather than one token at a time
  • Strong at code completion specifically

Worth knowing first

  • Quality sits below the frontier on anything demanding
  • Unusual architecture, so tooling support is thin

Index score

73.2

Provisional. This score is read from the product's public capability, not from the one-prompt rebuild the fully reviewed entries went through, so treat it as a placing rather than a verdict.

Craft63
Speed98
Control74
Value88
Rank
43 of 53
Builds
LLMs
Category
Coding
Output
Text and code over an API
Pricing
Paid only
Best for
Code completion where latency is the product

Scores are our own editorial judgement, weighted craft 55, speed 10, control 15 and value 20.