Mercury 2.5 Preview
Mercury 2.5 is the fastest reasoning LLM, and the latest diffusion LLM (dLLM) from Inception.
Specifications
| Lab | Inception |
|---|---|
| Context window | 260K tokens |
| Input price | $0.04/1M |
| Output price | $0.15/1M |
| Release | Aug 2026 |
Strengths
- Parallel token generation targets very low latency
- Reasoning capability at ultra-low token pricing
- 260K-token context window
- Useful for fast coding and structured generation
Weaknesses
- Preview release may change behavior or availability
- Diffusion generation is less established than autoregressive models
Best for
- Latency-sensitive coding assistants
- High-volume structured generation
- Fast reasoning prototypes
In Depth: Mercury 2.5 Preview
Summary
Mercury 2.5 Preview is an AI model from Inception.
Released in Aug 2026. It takes text input and produces text output, with a context window of 260K tokens. Input pricing is $0.04 per 1M tokens and output is $0.15 per 1M tokens on OpenRouter.