ModelPriceWatch.com
Last scan 2026-08-01 Models tracked 184 Providers 27 Cheapest paid Granite 4.0 Micro $0.017/Mtok in Every price links to its source

State of LLM Pricing — August 2026

What a million tokens costs, from verified provider list prices. All figures as of August 1, 2026; this edition does not change after publication — the live index does.

Key findings

data through 2026-08-01
  1. Frontier intelligence costs $4.66 per million tokens. The LLM Price Index — the equal-weight blended price across 9 current flagship models, one per major lab — stands at $4.66/Mtok, down 3.7% from $4.84 at the first reading on February 23, 2026.
  2. Frontier list prices are structurally sticky. Across 40 index readings over five-plus months, the level moved on exactly two dates — July 18 and July 25 — and both moves came from basket rebalances (new flagships entering), not from any lab repricing an existing model. As of August 1 the level had been flat for 7 days, since the last rebalance.
  3. The real deflation is at the floor, not the frontier. The cheapest model clearing a fixed GPT-4-class capability bar (GPQA Diamond ≥ 70) fell from $1.93/Mtok on March 1 (o4-mini) to $0.11/Mtok on July 25 (DeepSeek V4 Flash) — a 17x collapse in five months, leaving GPT-4-class capability 41x cheaper than the flagship ceiling.
  4. "Frontier flagship" spans a 21x price range. Basket constituents run from $0.54/Mtok blended (DeepSeek V4 Pro) to $11.25/Mtok (GPT-5.6 Sol). Closed-weights flagships average $4.83 against $4.08 for open-weights — an 18% proprietary premium.
  5. Price per unit of measured intelligence averages $0.0637 (blended $/Mtok ÷ benchmark score, over the 7 of 9 constituents with a published score). The best value is DeepSeek V4 Pro at $0.0076 per point — 18x better than the frontier average.

The index level

chain-linked · composition-neutral

The index opened its record at $4.84/Mtok on February 23, 2026 and held that level unchanged until mid-July. On July 18 it stepped down to $4.51 when xAI's flagship slot passed from Grok 4 to the cheaper Grok 4.5 (OpenAI's slot also passed from GPT-5.5 to GPT-5.6 Sol at an unchanged list price). On July 25 it stepped up to $4.66 when Alibaba's slot passed from Qwen3-Max to the pricier Qwen3.7-Max; the same day Anthropic's slot passed from Claude Opus 4.8 to Claude Opus 5 at identical list prices, and Moonshot (Kimi K3) entered as a ninth, previously unrepresented lab — a composition change that is chain-linked out of the trend, so it moves what is measured, not the measured trend.

Every move in the series is a succession event: a lab shipping a differently-priced flagship. No basket constituent changed its own list price in the period. That is the sticky-ceiling finding: frontier pricing moves by generation, not by discounting.

The basket, August 1

one current flagship per major lab
LLM Price Index constituents on August 1, 2026: each flagship model with provider, blended price per million tokens, and weights openness.
Model Provider Blended* $/1M Weights
GPT-5.6 SolOpenAI$11.25Closed
Claude Opus 5Anthropic$10.00Closed
Kimi K3Moonshot$6.00Open
Gemini 3.1 ProGoogle$4.50Closed
Qwen3.7-MaxAlibaba$3.75Closed
Grok 4.5xAI$3.00Closed
GLM-5.2Z.AI$2.15Open
Mistral Large 3Mistral$0.75Closed
DeepSeek V4 ProDeepSeek$0.544Closed
* Blended $/1M = (3×input + output) ÷ 4, list prices as of 2026-08-01. Current levels: /price-index/ · methodology. MODELPRICEWATCH.COM · 2026-08-01

What GPT-4-class capability costs

cheapest model clearing GPQA Diamond ≥ 70

Hold the capability bar fixed and the story inverts. The cheapest model with a verified GPQA Diamond score of 70 or better — roughly GPT-4-class general reasoning — stepped down five times between March and August:

Each model that lowered the price floor for GPT-4-class capability between March and August 2026, with the date it took the floor and its blended price per million tokens.
Floor taken by Date Blended $/1M
o4-miniMar 1, 2026$1.93
Grok 4.3May 15, 2026$1.56
DeepSeek V4 ProJun 23, 2026$0.544
GPT OSS 120BJun 29, 2026$0.262
DeepSeek V4 FlashJul 25, 2026$0.113
$1.93 → $0.113 is a 17x fall in under five months, against a flagship ceiling that moved 3.7% in the same period. The widening gap — 41x on August 1 — is the deflation story of 2026: intelligence at a fixed bar gets cheap fast, while the frontier holds its price. Full series: the Intelligence Cost Curve.

Cite today's level

"Frontier intelligence costs $4.66 per million tokens — -3.7% since Feb 23, 2026, flat for 7 days."

The LLM Price Index — ModelPriceWatch, 2026-08-01 · modelpricewatch.com/price-index/

The data behind this report

open, attributable, re-verifiable

Every figure above is computed from the ModelPriceWatch tracker: 181 models across 27 providers, 6,218 price snapshots spanning February 2024 to August 2026, each price traced to the provider's official pricing page. The full dataset is public:

Method, in one paragraph. The LLM Price Index is the equal-weighted average blended price ((3×input + output) ÷ 4 per million tokens) across a fixed basket of current general-purpose flagships, one per major lab. The basket changes only on a deliberate, dated rebalance; the trend is chain-linked from matched samples, so adding a lab never fabricates a price move. The capability floor uses one absolute metric (GPQA Diamond, raw %) with recorded provenance per score. Prices are list prices from official provider pages, re-verified on a rolling schedule — never estimated, never imputed. This report states only what that data shows; where history was backfilled from public archives it is labelled as reconstructed. This edition is fixed at August 1, 2026 — for current numbers use the live index.