ModelPriceWatch.com
Last scan 2026-09-22 Models tracked 267 Providers 35 Cheapest paid Granite 4.0 H Micro $0.017/Mtok in Every price links to its source

DeepSeek V4 Pro

by Fireworks

Current mid tier mid tier

Today's price · per 1M tokens

Input

$1.32

per 1M tokens

Output

$3.96

per 1M tokens

Blended

$1.98

blended $/1M — 3:1 weighted input:output

Cached input

$0.044

3% of input — prompt caching

How this price is scoped: Fireworks Standard serving path; the Priority column ($1.65/$0.055/$4.95) is a different path and is not what we track. Fireworks lists exactly one DeepSeek V4 Pro row, and between our Aug 24 and Aug 31 2026 readings of the page it changed from 'DeepSeek V4 Pro' at $1.74/$0.145/$3.48 to 'DeepSeek V4 Pro (0813)' at $1.32/$0.044/$3.96 — so $1.32/$0.044/$3.96 is the only rate Fireworks now sells this model at. Those figures match DeepSeek's own peak list price exactly, including the distinctive $0.044 cached rate, which is the maker's 0813 snapshot repricing propagating to its hosts rather than a Fireworks-specific discount.

Source: official Fireworks pricing · read MODELPRICEWATCH.COM · 2026-09-22

Price receipt

We read Fireworks's own pricing page on and found DeepSeek V4 Pro (0813) listed with a price on the same row — that read is where the number above comes from, recorded as a price change. We keep the page text we read; its content hash is 21e43393b6.

DeepSeek V4 Pro is available on 4 hosts. See the full cheapest-first comparison on the consolidated DeepSeek V4 Pro page →

Overview

DeepSeek V4 Pro on Fireworks. $1.32/$3.96 per 1M; cached $0.044.

Capabilities

struck through = not supported
Input 1/4
Text Image Audio Video
Output
Text
Features 3/9
Prompt caching Reasoning Coding Fast inference Long context Open weights Multimodal Web search Realtime

Benchmark performance

accuracy % · higher is better
Avg benchmark score73.5%
Perf per $/Mtok37.1
GPQA Diamond
90.1%
SWE-Bench Verified vendor
80.6%
LiveCodeBench
93.5%
MCP Atlas
73.6%
BrowseComp
83.4%
Humanity's Last Exam
48.2%
ARC-AGI 2
30%
AIME 2025
75%
MMMLU
82%
BFCL
72%
HumanEval
80.6%
MATH 500
80%
Independent composite scores
  • AA Intelligence Index: 36.3 (Artificial Analysis composite across reasoning, knowledge and coding evals)
  • AA-Omniscience Index: 0.8 (−100–100)
  • GDPval-AA v2: 1493.3 ELO (Real-world work tasks, human baseline = 1000)
How it stacks up
  • Ranks #30 of 84 benchmarked models by average score
  • Ranks #53 of 181 comparably-measured models by percentile score, across 14 independent measurements
  • Ranks #97 of 181 comparably-tested models by normalized performance per dollar
  • Strongest at LiveCodeBench — 93.5%, #2 of 26

80 tokens/sec output

Source: Artificial Analysis, DeepSeek model card, Artificial Analysis, Vellum LLM Leaderboard · updated · See full rankings →

Specifications

Provider
Fireworks
Context window
1M tokens
Modality
text
Parameters
Proprietary
Open source
No — proprietary
Released
Status
Current
Last updated
Tags
mid-tierreasoningcaching

Availability verified: listed on Fireworks's own page