State of LLM Pricing — October 2026
What a million tokens costs, from verified provider list prices. All figures as of October 1, 2026. This edition does not change after publication; the live index does. Previous edition: September 2026.
Key findings
data through 2026-10-01- Frontier intelligence cost $5.14 per million tokens on October 1, 2026. The LLM Price Index is the equal-weight blended price across 10 current flagship models, one per major lab. It rose 24.2% from $4.14 on September 1, its first monthly rise on record, because OpenAI's seat moved up to a pricier tier, not because any constituent's price rose. On September 1 it stood 9.4% below its first reading of February 23; on October 1 it stood 12.5% above it.
- No flagship changed its own price. The whole rise sits in one seat. Both of September's moves were seat changes. On September 4 GPT-6 Astra ($20.00 blended) took OpenAI's seat from GPT-5.6 Sol ($8.00), the largest one-day move on record at +29.0%. On September 30 Claude Opus 5.5 ($8.00) took Anthropic's seat from Opus 5 ($10.00), a move of −3.7%. Price the ten September-1 constituents at their October-1 list prices and they still average $4.14.
- Every successor launched at the same price or cheaper. Ten September releases from the basket's labs succeeded a model already on sale in the same line. Four launched cheaper: Claude Opus 5.5 −20%, GPT-6 Sol −50% against GPT-5.6 Sol's promotional price, GPT-6 Luna −56%, DeepSeek V4.1 Flash −20%. Six held their predecessor's price, and none launched higher. The month's one step up was Astra, a new tier above Sol that OpenAI names as its default. Had the GPT-6 Sol line filled OpenAI's seat instead, the index would have read $3.54.
- GPT-4-class capability cost $0.090/Mtok, down from $0.113, and its gap to the flagship ceiling widened from 36.6x to 57.1x. Most of that widening is a change in which DeepInfra listing the tracker reads, not a price cut; without it the gap is 45.5x. DeepSeek retired the floor model from its own API on September 10, so the GPT-4-class floor is now a host's price for a model its maker no longer sells.
- Anthropic and OpenAI cut cache prices. Until September 1, no Anthropic or OpenAI model on the tracker priced cache reads below a tenth of its input rate; among index constituents only DeepSeek V4 Pro did (3.3%, $0.044 against $1.32). By October 1 Claude Fable 5.1 charged 2.5% of its input rate, and Claude Opus 5.5 and GPT-6.1 Sol charged 5%.
The index level
chain-linked · composition-neutralThe index entered September at $4.14/Mtok on the ten-lab basket and left it at $5.14. It moved on two dates:
| Date | Event | What changed | Level $/1M |
|---|---|---|---|
| Sep 1, 2026 | — | Opening level, ten-lab basket | $4.14 |
| Sep 4, 2026 | Rebalance ↑ | OpenAI's seat: GPT-5.6 Sol ($4/$20 promotional, $8.00 blended) → GPT-6 Astra ($10/$50, $20.00) | $5.34 |
| Sep 30, 2026 | Rebalance ↓ | Anthropic's seat: Claude Opus 5 ($5/$25, $10.00) → Claude Opus 5.5 ($4/$20, $8.00) | $5.14 |
| Oct 1, 2026 | — | Closing level, unchanged since September 30 | $5.14 |
On September 4 the level stepped up to $5.34 when GPT-6 Astra, released the day before at $10/$50 and named by OpenAI as its default flagship, replaced GPT-5.6 Sol in OpenAI's seat. That is +29.0% in one day, still the largest single move in the index's history and its highest level. The index stood there for 26 days. On September 30 it stepped down to $5.14 when Claude Opus 5.5 replaced Opus 5. Opus 5.5 is the first change to the Opus list price since Opus 4.5 cut it from $15/$75 to $5/$25. Its blended price is 20% lower and its cached input 60% lower ($0.50 → $0.20). Grok 4.7 took xAI's seat the same day at Grok 4.6's price, so the whole step came from Anthropic's seat. Anthropic's other September flagship, Claude Fable 5.1 at $10/$50, sits outside the basket because Anthropic's seat follows the Opus line.
A note on the basis. Every level above is on the ten-lab basket; no lab entered or left in September, so nothing was rescaled, and September 1's $4.14 is the same number the September edition published. One thing did change: how seats are filled. Since September 30 each seat follows its lab's flagship line and seats the newest current, priced model in that line automatically. Before that the basket was maintained by hand, and it lagged. Claude Opus 5.5 was on the tracker from September 23 and Grok 4.7 from September 22, but neither was seated until September 30. By the index's own rule the level should have read $5.14 from September 23. The seven published readings for September 23–29 remain at $5.34, because a late succession is never back-dated into published levels, and this note records the lag.
The index has now moved on seven dates since February 23, two of them in September. Both September moves were seat changes; neither was a reprice. The September edition left one question for this one: whether the ceiling has turned. The level rose, but not because any list price turned: a new, pricier tier took OpenAI's seat, and every constituent kept its own list price.
One seat, not the market
same basket, two datesTake the ten models that made up the index on September 1 and price them at their October 1 list prices: they average $4.14, exactly what they averaged on September 1. None of them changed its own price during the month; the last time a constituent repriced while in the basket was GPT-5.6 Sol's promotional cut on August 21. Sort the September 1 and October 1 baskets by price and nine of the ten numbers are identical: the median was $3.00 on both dates, and the nine cheapest averaged $3.49 on both. The only difference is the top price: $10.00 (Claude Opus 5) on September 1 and $20.00 (GPT-6 Astra) on October 1. Astra alone contributes $2.00 of the $5.14.
Apart from Z.AI's GLM-5.3-Flash promotion ending (see Beyond the basket), the labs changed prices in September through new models rather than by repricing old ones, and no successor launched above the model it followed:
| New model | Follows | Blended $/1M | Change |
|---|---|---|---|
| GPT-6 Luna | GPT-5.6 Luna | $0.45 → $0.20 | −56% |
| GPT-6 Sol | GPT-5.6 Sol † | $8.00 → $4.00 | −50% |
| DeepSeek V4.1 Flash ‡ | DeepSeek V4 Flash | $0.66 → $0.525 | −20% |
| Claude Opus 5.5 | Claude Opus 5 | $10.00 → $8.00 | −20% |
| Claude Fable 5.1 | Claude Fable 5 | $20.00 → $20.00 | 0% |
| Claude Sonnet 5.5 | Claude Sonnet 5 | $4.00 → $4.00 | 0% |
| GPT-6.1 Sol | GPT-6 Sol | $4.00 → $4.00 | 0% |
| Grok 4.7 | Grok 4.6 | $3.00 → $3.00 | 0% |
| Muse Spark 1.3 | Muse Spark 1.2 | $2.00 → $2.00 | 0% |
| Gemini 3.8 Flash # | Gemini 3.7 Flash | $1.50 → $1.50 | 0% |
The exception is the model that matters most to the index. GPT-6 Astra ($10/$50) is not a successor to GPT-5.6 Sol. It is a new tier above the Sol line, and OpenAI also shipped two cheaper Sol models in September: GPT-6 Sol on September 22 and GPT-6.1 Sol on September 29, both at $2/$10. The index seats Astra because OpenAI names it as its starting point. That one judgment carries $1.60 of the level: with GPT-6 Sol in OpenAI's seat instead, the October 1 index would have read $3.54. This is arithmetic only, never a published level. It does show what September's +24.2% measures. Every same-line successor launched at or below its predecessor's price; the index rose because its OpenAI seat moved up a tier.
The basket, October 1
one current flagship per major lab| Model | Provider | Blended* $/1M |
|---|---|---|
| GPT-6 Astra | OpenAI | $20.00 |
| Claude Opus 5.5 | Anthropic | $8.00 |
| Kimi K3 ‖ | Moonshot | $6.00 |
| Gemini 3.1 Pro | $4.50 | |
| Qwen3.8-Max | Alibaba | $3.00 |
| Grok 4.7 | xAI | $3.00 |
| GLM-5.3 ‖ | Z.AI | $2.15 |
| Muse Spark 1.3 | Meta | $2.00 |
| DeepSeek V4 Pro § | DeepSeek | $1.98 |
| Mistral Large 3 | Mistral | $0.75 |
The spread from the priciest constituent to the cheapest doubled, from 13x on September 1 to 27x, all of it at the top: Astra's $20.00 against Mistral Large 3's unchanged $0.75. The second-highest price was $8.00 on both dates (GPT-5.6 Sol on September 1, Claude Opus 5.5 on October 1). The tracker began counting GLM-5.3 as open weights on September 5, alongside Kimi K3 (Z.AI had published the weights by August 28). This edition does not run the open-versus-closed price comparison. Two other constituents, DeepSeek V4 Pro and Mistral Large 3, are described as open-weight on their own pages but counted closed in the data the comparison reads. A premium computed on an unsettled flag is not worth citing.
What GPT-4-class capability costs
cheapest model clearing GPQA Diamond ≥ 70Hold the capability bar fixed and the price kept falling, though September's step needs reading carefully. The cheapest model with an externally measured GPQA Diamond score of 70 or better, roughly GPT-4-class general reasoning, cost $0.113/Mtok until September 30 and $0.090 from then on:
| Floor set by | Date | Blended $/1M |
|---|---|---|
| o4-mini | Mar 1, 2026 | $1.93 |
| Grok 4.3 | May 15, 2026 | $1.56 |
| DeepSeek V4 Pro | Jun 23, 2026 | $0.544 |
| GPT-OSS 120B (via Fireworks) | Jun 29, 2026 | $0.262 |
| DeepSeek V4 Flash (via DeepInfra) | Jul 25, 2026 | $0.113 |
| DeepSeek V4 Flash (via DeepInfra, 0731 entry) ¶ | Sep 30, 2026 | $0.090 |
DeepInfra lists two entries for DeepSeek V4 Flash. The undated one, which the curve read from July 25, was $0.09/$0.18 on our receipts from both August 15 and September 30, so we have no record of it moving in September. The dated entry, DeepSeek-V4-Flash-0731, which DeepInfra describes as the official release, was $0.08/$0.18 in mid-August and $0.06/$0.18 by September 30. On September 30 the tracker switched to the 0731 entry, and the floor stepped to $0.090. DeepInfra did cut that entry's price, but our receipts can date the cut only to somewhere between August 15 and September 30. The 0731 entry was already the cheaper of the two on our August 14 and 15 reads, so by the curve's cheapest-listing rule the published $0.113 floor, and the September edition's 36.6x, overstated the cheapest DeepInfra price from mid-August until September 30. As with the index, published readings are not restated, and the 0731 price on September 1 itself is not on record. The record ceiling-to-floor gap is mostly this switch: the ceiling fell that same day, and without the switch the gap would have narrowed from 47.3x to 45.5x.
The September edition warned that a floor set by a host is a price one vendor can withdraw. That warning has sharpened. DeepSeek retired V4 Flash from its own API on September 10; its replacement, V4.1 Flash, costs $0.525 at peak, and the only GPQA score on record for it is DeepSeek's own, so it does not count. GPT-4-class capability at $0.090 now exists only as a third party's copy of a model its maker no longer sells. If DeepInfra withdrew it, the floor would move to GPT-OSS 20B on Fireworks at $0.128. The cheapest qualifying model a lab sells itself is GPT-5.6 Luna at $0.45. September added nine models that cleared the bar for the first time, but none came close to the floor; the cheapest, Gemini 3.5 Flash-Lite, was $0.85. Two September launches are priced below $0.090, Inference.net's Schematron V2 Turbo ($0.06, a 3B extraction model) and Upstage's Solar Mini 4 ($0.0875 on a promotion to October 22). Neither has an externally measured GPQA Diamond score.
Price per point of measured intelligence
blended $/Mtok ÷ benchmark scoreOn October 1 the ten constituents averaged $0.0676 per point of measured intelligence, and every constituent had a published score, as it had since September 24; no earlier edition had full coverage. The figure does not compare with September's $0.0600, which covered seven of ten. During September the tracker added independent scores for Qwen3.8-Max, Grok 4.6 and GLM-5.3 and re-scored others. Re-scored on October 1 benchmark data, the September 1 basket comes to $0.0517. The difference between $0.0517 and $0.0676 is the new seats: replacing GPT-5.6 Sol with Astra by itself adds $0.0140 to the average, and without Astra the other nine average $0.0489.
On the index's measure the best value is unchanged: Mistral Large 3, at $0.0140 per point, is 4.8x cheaper than the average. The costliest is GPT-6 Astra at $0.2356, 16.8x the cheapest. Treat these as a ranking with wide error bars, not a measurement. Each score averages whichever benchmarks a model has been measured on, and the sets differ. Grok 4.7's score rests on a single benchmark, which is why it shows almost double Grok 4.6's cost per point at the same price, even though the two score the same on the one benchmark they share. The site's value ranking, which normalizes scores differently, put Muse Spark 1.3 just ahead of Mistral Large 3 among the index constituents on October 1.
Beyond the basket
the whole tracker in SeptemberAnthropic and OpenAI cut cache prices. Until September 1, no Anthropic or OpenAI model on the tracker priced cache reads below a tenth of its input rate (DeepSeek V4 Pro, the basket's DeepSeek seat, already charged 3.3%: $0.044 against $1.32). Then Claude Fable 5.1 launched with cached input at 2.5% of its input rate ($0.25 against $10), Claude Opus 5.5 at 5% ($0.20 against $4) and GPT-6.1 Sol at 5% ($0.10 against $2). By October 7 Anthropic had halved Claude Sonnet 5.5's cache rate to the same 5%. None of this moves the index, which blends input and output list prices only. For an agent or chat product that resends a long prompt on every turn, though, the cache rate can matter more than the headline input price.
OpenAI's lineup now overlaps itself. GPT-5.6 Sol is still on sale at its promotional $4/$20, which OpenAI says is "available at least through November 21, 2026". That is twice the $2/$10 of the newer GPT-6 Sol and GPT-6.1 Sol. GPT-6 Sol is also cheaper than GPT-5.6 Terra ($2/$12), the tier below the old Sol. Anyone still calling GPT-5.6 Sol is paying twice the list price of its successor.
Every price change we can date inside September went up. No basket lab repriced an existing model on a date we can place in September except Z.AI, whose GLM-5.3-Flash promotion ended (+100%). The other dated moves came from promotions and hosts: Upstage's Solar Pro 4 launch discount shrank from 90% to 70% (+200%), and Together raised Qwen3.7-Max by 25% at some point between September 11 and 21. Three other entries in our records are not price moves we can date to September: an earlier 60% Together increase on Qwen3.7-Max (between August 25 and September 11), DeepInfra's cut on its V4 Flash 0731 entry (between August 15 and September 30), and a September 6 entry for DeepSeek V4 Pro on Together, which switched the row to the 0813 listing Together had sold at $1.32/$3.96 since at least August 17 (9% lower blended) after the undated listing left its table on August 31. Every cut we can date to September came through a new model; the increases came through promotions ending and hosts repricing.
Corrections to the September edition
the September page stands as published, with a pointer here- GPT-6 Astra's cost per point. The September edition said Astra cost $0.208 per point "at a published score of 96". That score came from a single benchmark and stood for about twelve hours. From September 5 Astra's score averaged four benchmarks, at 84.9, so at the same $20.00 price it cost $0.2356 per point, 3.9x the September 1 average of $0.0600 rather than the 3.5x printed.
- Open weights on September 1. The September edition said GLM-5.3 counted as closed until its weights landed, which left Kimi K3 as the only open-weights constituent. Z.AI had already published the weights by August 28, so GLM-5.3 should have counted as open on September 1 alongside Kimi K3. DeepSeek V4 Pro and Mistral Large 3 are also described as open-weight on their own pages but counted closed, so the September edition's nine-closed-versus-one-open price comparison should not be cited.
- The first floor set by a host. The September edition called DeepInfra's V4 Flash floor "a different kind of floor from the four before it". It was not the first: the June 29 floor, GPT-OSS 120B at $0.262, was also a host's price. Fireworks, Groq and Together all listed it at $0.15/$0.60, and there was no first-party listing.
- DeepSeek's peak hours. The September edition gave V4 Pro's peak window as 01:00–04:00 and 06:00–10:00 UTC. That window applies Monday to Friday only; weekends are entirely off-peak.
- Model count. The September edition's "244 model rows (199 currently sold)" counted two models released after its September 1 cutoff, GPT-6 Astra and Gemini 3.8 Flash. The September 1 figures were 242 rows and 197 on sale.
After the cutoff
October 2–10, 2026 · stated so this page and the live index agreeThis edition is fixed at October 1. Through October 10 the index did not move: it held $5.14 with the same ten constituents. Outside the basket:
- October 5: the tracker removed Claude Mythos 5, Claude Mythos 5.1 and GPT-5.6 Cyber. All three are sold by approval only, so an ordinary API account cannot pay the price we printed. GPT-6 Astra was checked under the same rule and stays, because it is generally available.
- October 6: Mistral released Mistral Large 4 in preview at $0.68/$2.09, which is what Mistral's API price list bills. Its launch post quotes $1.36/$4.18, double that, and its model page shows that figure struck through beside the lower one; Mistral gives no end date for the lower rate. A preview does not take a seat in the index. If it goes generally available, it would take Mistral's seat from Large 3.
- October 7–8: Anthropic released Claude Haiku 5.5 at $0.10/$0.50 for prompts up to 100K tokens.
- Recorded October 7: Anthropic's pricing page listed Claude Sonnet 5.5's cached input at $0.10, half the $0.20 it showed on September 29; our saved copies cannot say which day in between it changed.
- October 10: Upstage's Solar Pro 4 promotion ended, and its blended price rose from $0.158 to $0.525.
Cite today's level
"Frontier intelligence costs $5.14 per million tokens — up 12.5% since Feb 23, 2026, flat for 10 days."
The LLM Price Index — ModelPriceWatch, 2026-10-10 · modelpricewatch.com/price-index/
The data behind this report
open, attributable, re-verifiableEvery figure above is computed from the ModelPriceWatch tracker as it stood on October 1, 2026: 273 model rows (220 then listed as on sale) across 35 providers, and 20,162 price snapshots from February 2024 to October 1, 2026. Each price is traced to the provider's official pricing page. The full dataset is public:
Hugging Face dataset →
The complete token price history plus the current models table and the index series, CC-BY-4.0.
Price Index API →
The live index level, constituents, series, and the dated citation string. Free, CORS-open.
Methodology →
Basket rules, the 3:1 in:out blend, chain-linking, and rebalance policy.
Method, in one paragraph. The LLM Price Index is the equal-weighted average blended price ((3×input + output) ÷ 4 per million tokens) across one current general-purpose flagship from each major lab. Each seat follows its lab's flagship line and seats the newest current, priced model in that line. The line a seat follows changes only on a deliberate, dated rebalance. The trend is chain-linked from matched samples, so adding a lab never fabricates a price move. Where a vendor publishes a time-of-day schedule, the peak rate is the list price. Where a vendor labels a rate promotional, it is taken as printed and flagged. The capability floor uses one absolute metric (GPQA Diamond, raw %) with recorded provenance for each score, read across every host that serves a base model. Prices are list prices from official provider pages, re-verified on a rolling schedule, never estimated or imputed. This report states only what that data shows. This edition is fixed at October 1, 2026; for current numbers use the live index.