ModelPriceWatch.com
Last scan 2026-09-10 Models tracked 249 Providers 32 Cheapest paid Granite 4.0 H Micro $0.017/Mtok in Every price links to its source

Mercury 2.5

by Inception

Current new · 2d Intro price budget cheap tier

Today's price · per 1M tokens

Input

$0.200

per 1M tokens

Output

$0.750

per 1M tokens

Blended

$0.338

blended $/1M — 3:1 weighted input:output

Cached input

$0.020

10% of input — prompt caching

How this price is scoped: These are Inception's LIST prices, per 1M tokens. Mercury 2.5 is currently on an 80%-off promotion, so the rate you actually pay today is $0.04 input / $0.15 output / $0.004 cached input. Inception has not published an end date for the promotion, so we track the list price — the rate the model reverts to — rather than a discounted number that could expire without notice. Inception's own rate card shows both, listing Mercury 2.5 at $0.20 input / $0.02 cached / $0.75 output struck through, alongside the discounted prices. Inception's launch announcement states the same standard pricing of $0.20 / $0.75 per 1M tokens and describes $0.04 / $0.15 as a launch discount. Verified 2026-09-08.

Source: official Inception pricing · read MODELPRICEWATCH.COM · 2026-09-10

Price receipt

We read Inception's own pricing page on and found Mercury 2.5 listed with a price on the same row — that read is where the number above comes from, recorded as a new model. We keep the page text we read; its content hash is ce0e61a51b.

Overview

Inception's second-generation diffusion large language model (dLLM), succeeding Mercury 2. Rather than decoding left to right, Mercury 2.5 produces and refines many tokens in parallel by iterative denoising, and adds explicit reasoning-effort control on top of the tool calling and structured outputs Mercury 2 already had. 260K context, 65,536 max output. List price $0.20 / $0.75 per 1M tokens, currently discounted 80% to $0.04 / $0.15 on an undated promotion.

Capabilities

struck through = not supported
Input 1/5
Text Image Audio Video PDF
Output 1/5
Text Image Audio Video Embedding
Features 4/9
Prompt caching Reasoning Coding Fast inference Long context Open weights Multimodal Web search Realtime

Specifications

Provider
Inception
Context window
260K tokens
Modality
text
Parameters
Proprietary
Open source
No — proprietary
Released
Status
Current
Last updated
Tags
budgetfastreasoning

Availability verified: listed on Inception's own page