ModelPriceWatch.com
Last scan 2026-09-15 Models tracked 259 Providers 32 Cheapest paid Granite 4.0 H Micro $0.017/Mtok in Every price links to its source

State of LLM Pricing — August 2026

What a million tokens costs, from verified provider list prices. All figures as of August 1, 2026; this edition does not change after publication — the live index does. Next edition: September 2026.

Key findings

data through 2026-08-01
  1. Frontier intelligence cost $4.39 per million tokens on August 1, 2026. The LLM Price Index — the equal-weight blended price across 10 current flagship models, one per major lab — closed that day at $4.39/Mtok, down 3.9% from $4.57 at the first reading on February 23, 2026.
  2. Frontier list prices are structurally sticky. Across 40 index readings over five-plus months, the level moved on exactly two dates — July 18 and July 25 — and both moves came from basket rebalances (new flagships entering), not from any lab repricing an existing model. As of August 1 the level had been flat for 7 days, since the last rebalance.
  3. The real deflation is at the floor, not the frontier. The cheapest model clearing a fixed GPT-4-class capability bar (GPQA Diamond ≥ 70) fell from $1.93/Mtok on March 1 (o4-mini) to $0.11/Mtok on July 25 (DeepSeek V4 Flash) — a 17x collapse in five months, leaving GPT-4-class capability 41x cheaper than the flagship ceiling.
  4. "Frontier flagship" spans a 21x price range. Basket constituents run from $0.54/Mtok blended (DeepSeek V4 Pro) to $11.25/Mtok (GPT-5.6 Sol). Closed-weights flagships averaged $4.47 against $4.08 for open-weights at publication — a 10% proprietary premium.
  5. Price per unit of measured intelligence averaged $0.0637 at publication (blended $/Mtok ÷ benchmark score, over the seven of ten constituents with a published score on the August-1 basket; on August 18 the Z.AI slot passed from GLM-5.2 to the not-yet-scored GLM-5.3, dropping coverage to six; on August 27 Meta's Muse Spark 1.2 earned independently measured scores, restoring coverage to seven of ten; and on September 5 Meta's slot passed in turn to the not-yet-scored Muse Spark 1.3, leaving the 6 of 10 constituents with a published score that the live index reports today). The best value at publication was DeepSeek V4 Pro at $0.0076 per point — 18x better than the frontier average.

The index level

chain-linked · composition-neutral

The index opened its record at $4.57/Mtok on February 23, 2026 and held that level unchanged until mid-July. On July 18 it stepped down to $4.25 when xAI's flagship slot passed from Grok 4 to the cheaper Grok 4.5 (OpenAI's slot also passed from GPT-5.5 to GPT-5.6 Sol at an unchanged list price). On July 25 it stepped up to $4.39 when Alibaba's slot passed from Qwen3-Max to the pricier Qwen3.7-Max; the same day Anthropic's slot passed from Claude Opus 4.8 to Claude Opus 5 at identical list prices, and Moonshot (Kimi K3) entered as a ninth, previously unrepresented lab — a composition change that is chain-linked out of the trend, so it moves what is measured, not the measured trend. On August 1 Meta entered as a tenth lab (Muse Spark 1.1, its first proprietary paid-API flagship), again as a composition change.

A note on the basis. Every level above is stated on the current ten-lab basket. The index is chain-linked and rescaled so its most recent point equals the published headline, so admitting a new lab restates the whole level history without changing the measured trend: the same series read on the nine-lab basket in force earlier on August 1 opened at $4.84 and stood at $4.66. The trend is the invariant — −3.9% since February 23 either way. Compare levels within one basis, not across two. Correction (August 4, 2026): the trend's rescale factor previously anchored on the cent-rounded headline, which let display rounding drift historical points; fixing it to anchor on the exact mean restated the February 23 first reading from $4.56 to $4.57 and the February-to-August trend from −3.7% to −3.9%. The as-published headline levels archived at each deploy are unchanged.

Every move in the series is a succession event: a lab shipping a differently-priced flagship. No basket constituent changed its own list price in the period. That is the sticky-ceiling finding: frontier pricing moves by generation, not by discounting.

The basket, August 1

one current flagship per major lab
LLM Price Index constituents on August 1, 2026: each flagship model with provider, blended price per million tokens, and weights openness.
Model Provider Blended* $/1M Weights
GPT-5.6 SolOpenAI$11.25Closed
Claude Opus 5Anthropic$10.00Closed
Kimi K3Moonshot$6.00Open
Gemini 3.1 ProGoogle$4.50Closed
Qwen3.7-MaxAlibaba$3.75Closed
Grok 4.5xAI$3.00Closed
GLM-5.2Z.AI$2.15Open
Mistral Large 3Mistral$0.75Closed
DeepSeek V4 ProDeepSeek$0.544Closed
* Blended $/1M = (3×input + output) ÷ 4, list prices as of 2026-08-01. Current levels: /price-index/ · methodology. MODELPRICEWATCH.COM · 2026-08-01

What GPT-4-class capability costs

cheapest model clearing GPQA Diamond ≥ 70

Hold the capability bar fixed and the story inverts. The cheapest model with a verified GPQA Diamond score of 70 or better — roughly GPT-4-class general reasoning — stepped down five times between March and August:

Each model that lowered the price floor for GPT-4-class capability between March and August 2026, with the date it took the floor and its blended price per million tokens.
Floor taken by Date Blended $/1M
o4-miniMar 1, 2026$1.93
Grok 4.3May 15, 2026$1.56
DeepSeek V4 ProJun 23, 2026$0.544
GPT OSS 120BJun 29, 2026$0.262
DeepSeek V4 FlashJul 25, 2026$0.113
$1.93 → $0.113 is a 17x fall in under five months, against a flagship ceiling that moved 3.9% in the same period. The widening gap — 41x on August 1 — is the deflation story of 2026: intelligence at a fixed bar gets cheap fast, while the frontier holds its price. Full series: the Intelligence Cost Curve.

Cite today's level

"Frontier intelligence costs $5.34 per million tokens — 16.8% since Feb 23, 2026, flat for 11 days."

The LLM Price Index — ModelPriceWatch, 2026-09-15 · modelpricewatch.com/price-index/

The data behind this report

open, attributable, re-verifiable

Every figure above is computed from the ModelPriceWatch tracker: 181 models across 27 providers, 6,218 price snapshots spanning February 2024 to August 2026, each price traced to the provider's official pricing page. The full dataset is public:

Method, in one paragraph. The LLM Price Index is the equal-weighted average blended price ((3×input + output) ÷ 4 per million tokens) across a fixed basket of current general-purpose flagships, one per major lab. The basket changes only on a deliberate, dated rebalance; the trend is chain-linked from matched samples, so adding a lab never fabricates a price move. The capability floor uses one absolute metric (GPQA Diamond, raw %) with recorded provenance per score. Prices are list prices from official provider pages, re-verified on a rolling schedule — never estimated, never imputed. This report states only what that data shows; where history was backfilled from public archives it is labelled as reconstructed. This edition is fixed at August 1, 2026 — for current numbers use the live index.