Today's price · per 1M tokens
Input
$0.600
per 1M tokens
Output
$3.00
per 1M tokens
Blended
$1.20
blended $/1M — 3:1 weighted input:output
Cached input
$0.100
17% of input — prompt caching
Price receipt
Historical price.This model is retired — withdrawn on — so the provider no longer sells it and its page can no longer confirm this number: a price for a model no longer on sale is unconfirmable by definition, not missing. The figures above are the last prices we recorded for this model, kept as a historical record rather than a live quote.
Overview
Kimi K2.5 on Fireworks. Withdrawn from serverless; last published rate $0.60/$3.00 per 1M, cached $0.10.
Deploy this open model on rented GPUs
Open weights mean you can self-host instead of paying per-token API prices. These platforms let you serve it on demand — often cheaper at scale.
Partner links — we may earn a commission if you sign up, at no cost to you. We only list platforms we'd recommend regardless.
Capabilities
struck through = not supportedBenchmark performance
accuracy % · higher is better- AA Intelligence Index: 36 (Artificial Analysis composite across reasoning, knowledge and coding evals)
- AA-Omniscience Index: -7.3 (−100–100)
- Ranks #63 of 98 benchmarked models by average score
- Ranks #105 of 193 comparably-measured models by percentile score, across 10 independent measurements
- Ranks #86 of 193 comparably-tested models by normalized performance per dollar
- Strongest at AIME 2025 — 96.1%, #10 of 37
337.7 tokens/sec output
Source: Artificial Analysis, Vellum LLM Leaderboard, LMSYS Chatbot Arena (UC Berkeley) · updated · See full rankings →
Specifications
- Provider
- Fireworks
- Context window
- 256K tokens
- Modality
- text, image
- Parameters
- Proprietary
- Open source
- Yes — open weights available
- Released
- Status
- Retired ended
- Last updated
- Tags
Availability verified: — per Fireworks's own deprecation notice