inception logo

Inception: Mercury 2.5

inception/mercury-2.5

Mercury 2.5 is the fastest reasoning LLM, and the latest diffusion LLM (dLLM) from Inception. Instead of generating tokens sequentially, Mercury 2.5 produces and refines multiple tokens in parallel, achieving...

Modalities

TextText

In / out price

$0.040 / $0.15 per 1M

Context

260K

Released

Sep 8, 2026

Why use Mercury 2.5

  • Step-by-step reasoning mode
  • Tool and function calling (agent-ready)
  • Structured (JSON) outputs

Pricing

Input$0.040 / 1M tokens
Output$0.15 / 1M tokens
Cached input (read)$0.004 / 1M tokens

List prices via OpenRouter, checked daily. Real cost depends on the provider and on reasoning tokens.

How it ranks

Intelligence index12.3

Source: Artificial Analysis (artificialanalysis.ai) via OpenRouter (openrouter.ai/rankings).

Cost calculator

Pick a task, set how often you run it, and compare what it costs on each model.

ModelPer runPer monthPer 1,000 runs
Inception: Mercury 2.5$0.0003$0.332$0.332

Estimates from list prices via OpenRouter; real bills vary by provider and reasoning tokens.

Will it fit? Context window simulator

Each square is about one page (500 words). Colored squares are your content; grey squares are free space in the model's context window.

Your content

  • Novel (90,000 words)

Total: about 120,000 tokens (≈ 90,226 words). Token counts are estimates.

Models

Inception: Mercury 2.5

260,000 tokens ≈ 391 pages

Fits: uses 46% of the window (room for 2× this much).

Write with the best model for the job

WordGPT picks and switches models for you: drafting, editing and files in one place. Free plan, no credit card.

Try it free

More from inception