ModelPriceWatch.com
Last scan 2026-10-08 Models tracked 278 Providers 37 Cheapest paid Granite 4.0 H Micro $0.017/Mtok in Every price links to its source

MiniMax-M2.5

by Fireworks

Retired budget open weights cheap tier
Retired on . No longer served on the endpoint we price — the figures below are kept for historical reference only. Replaced by MiniMax-M2.7.
Availability: Retired as a per-token row: Fireworks no longer sells this model serverless. Its model page (app.fireworks.ai/models/fireworks/minimax-m2p5, read 2026-08-14) shows 'Serverless: Not supported' and offers only a dedicated GPU deployment, billed per GPU-hour, so no per-token rate applies; the same day, the sibling MiniMax M2.7 page read 'Serverless: Supported'. Fireworks' rate card has not listed this model since at least 2026-08-09: on 08-09 and 08-14 it listed only MiniMax M3 and MiniMax M2.7. Fireworks' size-tier fallback ('Other base models', >16B $0.90 / MoE 56.1-176B $1.20) prices serverless calls only, so it does not apply. Fireworks publishes no delisting date, so the retirement date is the day we confirmed it, not a vendor notice. The $0.30/$1.20 shown is the last rate we published and matches Fireworks' MiniMax tier; the cached $0.03 never matched either sibling (both $0.06), and we never found it on a Fireworks page. MiniMax itself lists M2.5 under Legacy Models, and MiniMax's and Together's own M2.5 listings are retired here too (Together's on 2026-05-01).

Today's price · per 1M tokens

Input

$0.300

per 1M tokens

Output

$1.20

per 1M tokens

Blended

$0.525

blended $/1M — 3:1 weighted input:output

Cached input

$0.030

10% of input — prompt caching

Source: official Fireworks pricing · read MODELPRICEWATCH.COM · 2026-10-08

Price receipt

Historical price.This model is retired — withdrawn on — so the provider no longer sells it and its page can no longer confirm this number: a price for a model no longer on sale is unconfirmable by definition, not missing. The figures above are the last prices we recorded for this model, kept as a historical record rather than a live quote.

Overview

MiniMax-M2.5 on Fireworks. Withdrawn from serverless; last published rate $0.30/$1.20 per 1M, cached $0.03.

Run it yourself

Deploy this open model on rented GPUs

Open weights mean you can self-host instead of paying per-token API prices. These platforms let you serve it on demand — often cheaper at scale.

Partner links — we may earn a commission if you sign up, at no cost to you. We only list platforms we'd recommend regardless.

Capabilities

struck through = not supported
Input 1/4
Text ✓ Image Audio Video
Output
Text ✓
Features 3/9
Prompt caching ✓ Reasoning Coding Fast inference Long context ✓ Open weights ✓ Multimodal Web search Realtime

Benchmark performance

accuracy % · higher is better
Percentile vs all tracked models 22.9th1 independent measurement
Percentile per $/Mtok 44

We hold no per-benchmark accuracy scores for this model yet, so it has no accuracy average. That is a gap in our coverage, not a sign the model is untested — independent evaluators often publish a composite index for a new model long before releasing its per-benchmark numbers. The percentile above is its standing across the independent composites below.

Independent composite scores
  • AA-Omniscience Index: -38.9 (−100–100)
How it stacks up
  • Ranks #145 of 193 comparably-measured models by percentile score, across 1 independent measurement
  • Ranks #64 of 193 comparably-tested models by normalized performance per dollar

Source: Artificial Analysis, LMSYS Chatbot Arena (UC Berkeley) · updated · See full rankings →

Specifications

Provider
Fireworks
Context window
128K tokens
Modality
text
Parameters
Proprietary
Open source
Yes — open weights available
Released
Status
Retired ended
Last updated
Tags
open-weightsbudgetcaching

Availability verified: — per Fireworks's own deprecation notice