Mercury 2.5 Preview

Inception · inception/mercury-2.5-preview

← Back to leaderboard

Mercury 2.5 is the fastest reasoning LLM, and the latest diffusion LLM (dLLM) from Inception. Instead of generating tokens sequentially, Mercury 2.5 produces and refines multiple tokens in parallel, achieving...

Mercury 2.5 Preview ranks #66 of 348 models in the default AI Benchmark Hub order (quality first, then cost). At $0.10 blended per million tokens, it is cheaper than 93% of models with published API pricing. Its 260,000-token context window is about the catalog median (262,144 tokens).

texttext->text

Context

Max context: 260000
Max output: 65536

Pricing

Input / 1M: 0.04
Output / 1M: 0.15
Blend / 1M: 0.10

Quality

Quality index:

Provider

Provider: Inception
Moderated: no

Monthly cost at this blended rate

Estimated API spend if every token is billed at $0.10per million (the blend of this model's input and output prices). Real bills depend on the input/output mix.

Tokens / monthEst. cost (USD)
1,000,000$0.10
10,000,000$0.95
100,000,000$9.50

Related models

Same lab, similar price, or similar quality — useful next comparisons.

More from Inception

Closest API price