Pricing trends
How LLM API pricing is evolving — price drops, provider competitiveness, and market movements over time. Built from twice-daily snapshots the scan never overwrites.
State of LLM pricing
market overview — computed from the current scanModels tracked
178
current models — deprecated and preview tiers excluded
Providers
30
every price read from the provider's own pricing page
Cheapest paid · blended
$0.041 /Mtok
Granite 4.0 Micro — IBM
Most expensive · blended
$67.50 /Mtok
GPT-5.4 Pro — OpenAI
How prices are moving
movement — aggregated over a 90-day window, ranked by sizeRepricing activity over the 90 days to Aug 1, 2026. For the full chronological log of every launch, price move and context change — plus JSON and Atom feeds — see the Newsroom.
Repricings in 90d
14
tracked list-price changes inside the window
Price drops
7
blended price went down — cheaper than before
Price increases
7
the moves that show up on invoices
Median move size
34.8%
half of the moves in the window were smaller than this
Biggest moves
Ranked by size of the move, not by date — the question the chronological feed can't answer.
| Model | Move | Blended $/1M | In $/1M | Out $/1M | Size | Date |
|---|---|---|---|---|---|---|
| Gemini 2.5 FlashGoogle | increase | $0.131 → $0.850 | $0.075 → $0.300 | $0.300 → $2.50 | 547.6% | Jul 4, 2026 |
| GPT-Image-2OpenAI | increase | $6.25 → $13.50 | $5.00 → $8.00 | $10.00 → $30.00 | 116% | Jun 25, 2026 |
| Gemini 2.5 FlashGoogle | drop | $0.850 → $0.131 | $0.300 → $0.075 | $2.50 → $0.300 | 84.6% | Jun 25, 2026 |
| GPT-5.6 LunaOpenAI | drop | $2.25 → $0.450 | $1.00 → $0.200 | $6.00 → $1.20 | 80% | Aug 1, 2026 |
| MiniMax-M3MiniMax | drop | $1.05 → $0.525 | $0.600 → $0.300 | $2.40 → $1.20 | 50% | Jun 17, 2026 |
| Gemini 3.1 Flash-LiteGoogle | drop | $1.01 → $0.563 | $0.450 → $0.250 | $2.70 → $1.50 | 44.4% | Jun 9, 2026 |
| Mistral Small 3.2 24BMistral | increase | $0.110 → $0.150 | $0.080 → $0.100 | $0.200 → $0.300 | 36.4% | Jul 5, 2026 |
| Pixtral 12BMistral | drop | $0.150 → $0.100 | $0.150 → $0.100 | $0.150 → $0.100 | 33.3% | Jun 25, 2026 |
Who reprices most
Providers by number of tracked price changes in the window — a proxy for how actively each one competes on price.
Provider competitiveness
median blended $/Mtok per providerMedian blended cost per provider — lower means more competitive on price. The count is how many current models each provider lists.
What we're tracking
coverage — how the history is builtModel Price Watch takes twice-daily snapshots of every model's pricing across all major LLM providers. Each snapshot records input price, output price, cached input price, context window, and status — so we capture not just what changed, but when and by how much.
Currently tracking 178 models across 30 providers. Snapshots are captured automatically twice a day by our pricing pipeline. Historical data is never overwritten — each snapshot is append-only, so every recorded price point is preserved.
Dig deeper into the token economy:
- Full price-change feed — every recorded input/output move across all providers, newest first
- Marketwide price index — the overall direction of LLM API costs, with the intelligence-cost curve
- Per-model price history — input/output/cached price points over time for every model
- Compare current prices — head-to-head pricing across any two models