Mercury 2.5
by Inception
Today's price · per 1M tokens
Input
$0.200
per 1M tokens
Output
$0.750
per 1M tokens
Blended
$0.338
blended $/1M — 3:1 weighted input:output
Cached input
$0.020
10% of input — prompt caching
How this price is scoped: These are Inception's LIST prices, per 1M tokens. Mercury 2.5 is currently on an 80%-off promotion, so the rate you actually pay today is $0.04 input / $0.15 output / $0.004 cached input. Inception has not published an end date for the promotion, so we track the list price — the rate the model reverts to — rather than a discounted number that could expire without notice. Inception's own rate card shows both, listing Mercury 2.5 at $0.20 input / $0.02 cached / $0.75 output struck through, alongside the discounted prices. Inception's launch announcement states the same standard pricing of $0.20 / $0.75 per 1M tokens and describes $0.04 / $0.15 as a launch discount. Verified 2026-09-08.
Price receipt
We read Inception's own pricing page on and found Mercury 2.5 listed with a price on the same row — that read is where the number above comes from, recorded as a new model. We keep the page text we read; its content hash is ce0e61a51b.
Overview
Inception's second-generation diffusion large language model (dLLM), succeeding Mercury 2. Rather than decoding left to right, Mercury 2.5 produces and refines many tokens in parallel by iterative denoising, and adds explicit reasoning-effort control on top of the tool calling and structured outputs Mercury 2 already had. 260K context, 65,536 max output. List price $0.20 / $0.75 per 1M tokens, currently discounted 80% to $0.04 / $0.15 on an undated promotion.
Capabilities
struck through = not supportedSpecifications
- Provider
- Inception
- Context window
- 260K tokens
- Modality
- text
- Parameters
- Proprietary
- Open source
- No — proprietary
- Released
- Status
- Current
- Last updated
- Tags
Availability verified: — listed on Inception's own page