Mercury 2.5 Preview

by Inception Coding Reasoning Fast

Mercury 2.5 is the fastest reasoning LLM, and the latest diffusion LLM (dLLM) from Inception.

Choose a model to compare against Mercury 2.5 Preview

Specifications

Specifications for Mercury 2.5 Preview
LabInception
Context window260K tokens
Input price $0.04/1M
Output price $0.15/1M
ReleaseAug 2026

Strengths

  • Parallel token generation targets very low latency
  • Reasoning capability at ultra-low token pricing
  • 260K-token context window
  • Useful for fast coding and structured generation

Weaknesses

  • Preview release may change behavior or availability
  • Diffusion generation is less established than autoregressive models

Best for

  • Latency-sensitive coding assistants
  • High-volume structured generation
  • Fast reasoning prototypes

In Depth: Mercury 2.5 Preview

Summary

Mercury 2.5 Preview is an AI model from Inception.

Released in Aug 2026. It takes text input and produces text output, with a context window of 260K tokens. Input pricing is $0.04 per 1M tokens and output is $0.15 per 1M tokens on OpenRouter.

Sources & Further Reading

Related Models