Today's price · per 1M tokens
Input
$0.300
per 1M tokens
Output
$1.20
per 1M tokens
Blended
$0.525
blended $/1M — 3:1 weighted input:output
Cached input
$0.030
10% of input — prompt caching
Price receipt
Historical price.This model is retired — withdrawn on — so the provider no longer sells it and its page can no longer confirm this number: a price for a model no longer on sale is unconfirmable by definition, not missing. The figures above are the last prices we recorded for this model, kept as a historical record rather than a live quote.
Overview
MiniMax-M2.5 on Fireworks. Withdrawn from serverless; last published rate $0.30/$1.20 per 1M, cached $0.03.
Deploy this open model on rented GPUs
Open weights mean you can self-host instead of paying per-token API prices. These platforms let you serve it on demand — often cheaper at scale.
Partner links — we may earn a commission if you sign up, at no cost to you. We only list platforms we'd recommend regardless.
Capabilities
struck through = not supportedBenchmark performance
accuracy % · higher is betterWe hold no per-benchmark accuracy scores for this model yet, so it has no accuracy average. That is a gap in our coverage, not a sign the model is untested — independent evaluators often publish a composite index for a new model long before releasing its per-benchmark numbers. The percentile above is its standing across the independent composites below.
- AA-Omniscience Index: -38.9 (−100–100)
- Ranks #145 of 193 comparably-measured models by percentile score, across 1 independent measurement
- Ranks #64 of 193 comparably-tested models by normalized performance per dollar
Source: Artificial Analysis, LMSYS Chatbot Arena (UC Berkeley) · updated · See full rankings →
Specifications
- Provider
- Fireworks
- Context window
- 128K tokens
- Modality
- text
- Parameters
- Proprietary
- Open source
- Yes — open weights available
- Released
- Status
- Retired ended
- Last updated
- Tags
Availability verified: — per Fireworks's own deprecation notice