State of LLM Pricing — August 2026
What a million tokens costs, from verified provider list prices. All figures as of August 1, 2026; this edition does not change after publication — the live index does.
Key findings
data through 2026-08-01- Frontier intelligence costs $4.66 per million tokens. The LLM Price Index — the equal-weight blended price across 9 current flagship models, one per major lab — stands at $4.66/Mtok, down 3.7% from $4.84 at the first reading on February 23, 2026.
- Frontier list prices are structurally sticky. Across 40 index readings over five-plus months, the level moved on exactly two dates — July 18 and July 25 — and both moves came from basket rebalances (new flagships entering), not from any lab repricing an existing model. As of August 1 the level had been flat for 7 days, since the last rebalance.
- The real deflation is at the floor, not the frontier. The cheapest model clearing a fixed GPT-4-class capability bar (GPQA Diamond ≥ 70) fell from $1.93/Mtok on March 1 (o4-mini) to $0.11/Mtok on July 25 (DeepSeek V4 Flash) — a 17x collapse in five months, leaving GPT-4-class capability 41x cheaper than the flagship ceiling.
- "Frontier flagship" spans a 21x price range. Basket constituents run from $0.54/Mtok blended (DeepSeek V4 Pro) to $11.25/Mtok (GPT-5.6 Sol). Closed-weights flagships average $4.83 against $4.08 for open-weights — an 18% proprietary premium.
- Price per unit of measured intelligence averages $0.0637 (blended $/Mtok ÷ benchmark score, over the 7 of 9 constituents with a published score). The best value is DeepSeek V4 Pro at $0.0076 per point — 18x better than the frontier average.
The index level
chain-linked · composition-neutralThe index opened its record at $4.84/Mtok on February 23, 2026 and held that level unchanged until mid-July. On July 18 it stepped down to $4.51 when xAI's flagship slot passed from Grok 4 to the cheaper Grok 4.5 (OpenAI's slot also passed from GPT-5.5 to GPT-5.6 Sol at an unchanged list price). On July 25 it stepped up to $4.66 when Alibaba's slot passed from Qwen3-Max to the pricier Qwen3.7-Max; the same day Anthropic's slot passed from Claude Opus 4.8 to Claude Opus 5 at identical list prices, and Moonshot (Kimi K3) entered as a ninth, previously unrepresented lab — a composition change that is chain-linked out of the trend, so it moves what is measured, not the measured trend.
Every move in the series is a succession event: a lab shipping a differently-priced flagship. No basket constituent changed its own list price in the period. That is the sticky-ceiling finding: frontier pricing moves by generation, not by discounting.
The basket, August 1
one current flagship per major lab| Model | Provider | Blended* $/1M | Weights |
|---|---|---|---|
| GPT-5.6 Sol | OpenAI | $11.25 | Closed |
| Claude Opus 5 | Anthropic | $10.00 | Closed |
| Kimi K3 | Moonshot | $6.00 | Open |
| Gemini 3.1 Pro | $4.50 | Closed | |
| Qwen3.7-Max | Alibaba | $3.75 | Closed |
| Grok 4.5 | xAI | $3.00 | Closed |
| GLM-5.2 | Z.AI | $2.15 | Open |
| Mistral Large 3 | Mistral | $0.75 | Closed |
| DeepSeek V4 Pro | DeepSeek | $0.544 | Closed |
What GPT-4-class capability costs
cheapest model clearing GPQA Diamond ≥ 70Hold the capability bar fixed and the story inverts. The cheapest model with a verified GPQA Diamond score of 70 or better — roughly GPT-4-class general reasoning — stepped down five times between March and August:
| Floor taken by | Date | Blended $/1M |
|---|---|---|
| o4-mini | Mar 1, 2026 | $1.93 |
| Grok 4.3 | May 15, 2026 | $1.56 |
| DeepSeek V4 Pro | Jun 23, 2026 | $0.544 |
| GPT OSS 120B | Jun 29, 2026 | $0.262 |
| DeepSeek V4 Flash | Jul 25, 2026 | $0.113 |
Cite today's level
"Frontier intelligence costs $4.66 per million tokens — -3.7% since Feb 23, 2026, flat for 7 days."
The LLM Price Index — ModelPriceWatch, 2026-08-01 · modelpricewatch.com/price-index/
The data behind this report
open, attributable, re-verifiableEvery figure above is computed from the ModelPriceWatch tracker: 181 models across 27 providers, 6,218 price snapshots spanning February 2024 to August 2026, each price traced to the provider's official pricing page. The full dataset is public:
Hugging Face dataset →
The complete token price history plus the current models table and the index series, CC-BY-4.0.
Price Index API →
The live index level, constituents, series, and the dated citation string. Free, CORS-open.
Methodology →
Basket rules, the 3:1 in:out blend, chain-linking, and rebalance policy.
Method, in one paragraph. The LLM Price Index is the equal-weighted average blended price ((3×input + output) ÷ 4 per million tokens) across a fixed basket of current general-purpose flagships, one per major lab. The basket changes only on a deliberate, dated rebalance; the trend is chain-linked from matched samples, so adding a lab never fabricates a price move. The capability floor uses one absolute metric (GPQA Diamond, raw %) with recorded provenance per score. Prices are list prices from official provider pages, re-verified on a rolling schedule — never estimated, never imputed. This report states only what that data shows; where history was backfilled from public archives it is labelled as reconstructed. This edition is fixed at August 1, 2026 — for current numbers use the live index.