Today's price · per 1M tokens
Input
$0.600
per 1M tokens
Output
$2.40
per 1M tokens
Blended
$1.05
blended $/1M — 3:1 weighted input:output
Cached input
$0.120
20% of input — prompt caching
Price receipt
We read Fireworks's own pricing page on Aug 9, 2026 and found NVIDIA Nemotron 3 Ultra (Preview) listed with a price on the same row — that read is where the number above comes from. We keep the page text we read; its content hash is 4f5353269c.
Available on 2 hosts
cheapest blended first · $ per 1M tokensNVIDIA Nemotron 3 Ultra is sold by 2 providers. Prices are per 1M tokens (blended = (3×input + 1×output) ÷ 4). “vs cheapest” compares each host to the lowest blended price.
| Host | Input | Output | Blended | vs cheapest |
|---|---|---|---|---|
| Fireworks cheapest | $0.600 | $2.40 | $1.05 | — |
| Together details → | $0.600 | $3.60 | $1.35 | +29% |
Overview
NVIDIA Nemotron 3 Ultra (preview) on Fireworks. $0.60/$2.40 per 1M.
Deploy this open model on rented GPUs
Open weights mean you can self-host instead of paying per-token API prices. These platforms let you serve it on demand — often cheaper at scale.
Partner links — we may earn a commission if you sign up, at no cost to you. We only list platforms we'd recommend regardless.
Capabilities
struck through = not supportedBenchmark performance
accuracy % · higher is betterWe hold no per-benchmark accuracy scores for this model yet, so it has no accuracy average. That is a gap in our coverage, not a sign the model is untested — independent evaluators often publish a composite index for a new model long before releasing its per-benchmark numbers. The percentile above is its standing across the independent composites below.
- AA Intelligence Index: 38.3 (Artificial Analysis composite across reasoning, knowledge and coding evals)
- AA Agentic Index: 27.5 (Tool use, planning, autonomy)
- AA-Omniscience Index: -0.4 (−100–100)
- GDPval-AA v2: 1162 ELO (Real-world work tasks, human baseline = 1000)
- LMSYS Chatbot Arena: 1426.2 ELO (Human preference)
- Ranks #94 of 170 comparably-measured models by percentile score, across 5 independent measurements
- Ranks #68 of 170 comparably-tested models by normalized performance per dollar
Source: Artificial Analysis · updated Jun 29, 2026 · See full rankings →
Specifications
- Provider
- Fireworks
- Context window
- 128K tokens
- Modality
- text
- Parameters
- Proprietary
- Open source
- Yes — open weights available
- Released
- Jan 1, 2026
- Status
- Current
- Last updated
- Jun 25, 2026
- Tags
Availability verified: Aug 9, 2026 — listed on Fireworks's own page