Mercury 2.5 Preview
Inception · inception/mercury-2.5-preview
Mercury 2.5 is the fastest reasoning LLM, and the latest diffusion LLM (dLLM) from Inception. Instead of generating tokens sequentially, Mercury 2.5 produces and refines multiple tokens in parallel, achieving...
Mercury 2.5 Preview ranks #66 of 348 models in the default AI Benchmark Hub order (quality first, then cost). At $0.10 blended per million tokens, it is cheaper than 93% of models with published API pricing. Its 260,000-token context window is about the catalog median (262,144 tokens).
Context
Max output: 65536
Pricing
Output / 1M: 0.15
Blend / 1M: 0.10
Quality
Provider
Moderated: no
Monthly cost at this blended rate
Estimated API spend if every token is billed at $0.10per million (the blend of this model's input and output prices). Real bills depend on the input/output mix.
| Tokens / month | Est. cost (USD) |
|---|---|
| 1,000,000 | $0.10 |
| 10,000,000 | $0.95 |
| 100,000,000 | $9.50 |
Related models
Same lab, similar price, or similar quality — useful next comparisons.
More from Inception
- Mercury 2$0.50 / 1M
Closest API price
- Command R7B (12-2024)$0.09 / 1M
- Reka Edge$0.10 / 1M
- Ministral 3 3B 2512$0.10 / 1M
- Gemma 3 12B$0.10 / 1M
- Laguna XS 2.1$0.09 / 1M