Mercury 2.5
Inception · inception/mercury-2.5
Mercury 2.5 is the fastest reasoning LLM, and the latest diffusion LLM (dLLM) from Inception. Instead of generating tokens sequentially, Mercury 2.5 produces and refines multiple tokens in parallel, achieving...
Mercury 2.5 ranks #60 of 359 models in the default AI Benchmark Hub order (quality first, then cost). At $0.10 blended per million tokens, it is cheaper than 92% of models with published API pricing. Its 260,000-token context window is about the catalog median (262,144 tokens).
Context
Max output: 65536
Pricing
Output / 1M: 0.15
Blend / 1M: 0.10
Quality
Provider
Moderated: no
Monthly cost at this blended rate
Estimated API spend if every token is billed at $0.10per million (the blend of this model's input and output prices). Real bills depend on the input/output mix.
| Tokens / month | Est. cost (USD) |
|---|---|
| 1,000,000 | $0.10 |
| 10,000,000 | $0.95 |
| 100,000,000 | $9.50 |
Related models
Same lab, similar price, or similar quality — useful next comparisons.
More from Inception
- Mercury 2$0.50 / 1M
Closest API price
- MythoMax 13B$0.10 / 1M
- Command R7B (12-2024)$0.09 / 1M
- Reka Edge$0.10 / 1M
- Ministral 3 3B 2512$0.10 / 1M
- Gemma 3 12B$0.10 / 1M