ModelPriceWatch.com
Last scan 2026-09-23 Models tracked 270 Providers 35 Cheapest paid Granite 4.0 H Micro $0.017/Mtok in Every price links to its source

Fireworks

Hosting provider · 14 models tracked · Founded 2022

Inference platform serving open-weight models (Llama, Qwen, DeepSeek, Kimi, GLM, MiniMax, gpt-oss) on its own infrastructure at first-party per-token pricing — distinct rates worth comparing per host.

Fireworks pricing at a glanceSeptember 2026 · $ per 1M tokens

Fireworks API pricing (September 2026): 14 current models range from $0.070 to $1.40 per 1M input tokens and $0.280 to $4.40 per 1M output tokens. The cheapest paid model is GPT-OSS 20B at $0.070/1M input; the priciest is GLM-5.3 at $1.40/1M input / $4.40 output. Every price links to Fireworks's official pricing page and refreshes twice daily.

Models14
Input range$0.070–$1.40
Output range$0.280–$4.40
Cached-input tiers14

Today's Fireworks prices

14 models · cheapest blended first

Sorted by blended cost (cheapest first). Prices per 1M tokens, September 2026 — every price links to Fireworks's official pricing page.

Current Fireworks model prices per 1M tokens, sorted by blended cost
Model Blended* Input Output Cached in Relative cost Context Status
GPT-OSS 20B
$0.128
$0.070 $0.300 $0.035
128K Current
DeepSeek V4 Flash
$0.175
$0.140 $0.280 $0.028
1M Current
GPT-OSS 120B
$0.262
$0.150 $0.600 $0.015
128K Current
MiniMax-M2.7
$0.525
$0.300 $1.20 $0.060
128K Current
MiniMax-M3
$0.525
$0.300 $1.20 $0.060
1M Current
Muse Glimmer 30B
$0.637
$0.350 $1.50 $0.040
131K Current
Qwen3.7-Plus
$0.700
$0.400 $1.60 $0.080
128K Current
NVIDIA Nemotron 3 Ultra
$1.05
$0.600 $2.40 $0.120
128K Current
Kimi K2.6
$1.71
$0.950 $4.00 $0.160
256K Current
Kimi K2.7 Code
$1.71
$0.950 $4.00 $0.190
256K Current
DeepSeek V4 Pro
$1.98
$1.32 $3.96 $0.044
1M Current
GLM-5.1
$2.15
$1.40 $4.40 $0.260
128K Current
GLM-5.2
$2.15
$1.40 $4.40 $0.140
1M Current
GLM-5.3
$2.15
$1.40 $4.40 $0.260
1M Current
* Blended = (3×input + 1×output) ÷ 4 $/1M tokens · cheap · mid · expensive MODELPRICEWATCH.COM · 2026-09-23

Quick stats

Models tracked
14
Type
hosting
Founded
2022
Cheapest model
GPT-OSS 20B
Cheapest blended
$0.128/M

Open-weight models

12 open-weight models available from this provider.

GLM-5.1
$1.40/M in · 128K ctx
GLM-5.2
$1.40/M in · 1M ctx
GLM-5.3
$1.40/M in · 1M ctx
Muse Glimmer 30B
$0.350/M in · 131K ctx
GPT-OSS 120B
$0.150/M in · 128K ctx
GPT-OSS 20B
$0.070/M in · 128K ctx

Try Fireworks

Sign up and start building with Fireworks models.

Get started
Run it yourself

Self-host Fireworks's open-weight models

These models ship with open weights, so you can serve them yourself on rented GPUs instead of paying per-token API prices — often cheaper at scale.

Partner links — we may earn a commission if you sign up, at no cost to you. We only list platforms we'd recommend regardless.