ModelPriceWatch.com
Last scan 2026-09-15 Models tracked 259 Providers 32 Cheapest paid Granite 4.0 H Micro $0.017/Mtok in Every price links to its source

Kimi K2.6

by Moonshot

Current mid tier open weights mid tier

Today's price · per 1M tokens

Input

$0.950

per 1M tokens

Output

$4.00

per 1M tokens

Blended

$1.71

blended $/1M — 3:1 weighted input:output

Cached input

$0.160

17% of input — prompt caching

Source: official Moonshot pricing · read MODELPRICEWATCH.COM · 2026-09-15

Price receipt

No confirmed receipt. We hold dated captures of Moonshot's pricing page, but none of them tied Kimi K2.6 to the price above on that page, so we do not claim this number is sourced. Check the provider's page directly before relying on it.

Available on 2 hosts

cheapest blended first · $ per 1M tokens

Kimi K2.6 is sold by 2 providers. Prices are per 1M tokens (blended = (3×input + 1×output) ÷ 4). The first-party row is the model maker; “vs first-party” shows each host’s blended price relative to it.

Kimi K2.6 price on every host still selling it, per 1M tokens, cheapest blended first
Host Input Output Blended vs first-party
Fireworks details → $0.950 $4.00 $1.71 0%
Moonshot first-party $0.950 $4.00 $1.71
Each host links to its provider page and official pricing. Prices refresh twice daily. MODELPRICEWATCH.COM · 2026-09-15

Overview

Kimi's latest multimodal model. $0.95/$4.00 per 1M; cached $0.16.

Run it yourself

Deploy this open model on rented GPUs

Open weights mean you can self-host instead of paying per-token API prices. These platforms let you serve it on demand — often cheaper at scale.

Partner links — we may earn a commission if you sign up, at no cost to you. We only list platforms we'd recommend regardless.

Capabilities

struck through = not supported
Input 3/5
Text Image Audio Video PDF
Output 1/5
Text Image Audio Video Embedding
Features 4/9
Prompt caching Reasoning Coding Fast inference Long context Open weights Multimodal Web search Realtime

Benchmark performance

accuracy % · higher is better
Avg benchmark score76.2%
Perf per $/Mtok44.5
GPQA Diamond
90.5%
SWE-Bench Verified vendor
40%
BrowseComp
83.2%
OSWorld-Verified
73.1%
Humanity's Last Exam
54%
ARC-AGI 2 vendor
22%
AIME 2025 vendor
62%
MMMLU vendor
76%
BFCL vendor
65%
HumanEval
80.2%
MATH 500 vendor
72%
AIME 2026 *
96.4%
Independent composite scores
  • AA-Omniscience Index: 5.3 (−100–100)
  • LMSYS Chatbot Arena: 1460.4 ELO (Human preference)
How it stacks up
  • Ranks #20 of 84 benchmarked models by average score
  • Ranks #61 of 181 comparably-measured models by percentile score, across 7 independent measurements
  • Ranks #90 of 181 comparably-tested models by normalized performance per dollar
  • Strongest at Humanity's Last Exam — 54%, #12 of 65

342.6 tokens/sec output 0.68s latency to first token (TTFT)

Source: Vellum LLM Leaderboard · updated · See full rankings →

Specifications

Provider
Moonshot
Context window
262K tokens
Modality
text, image, video
Parameters
Proprietary
Open source
Yes — open weights available
Released
Status
Current
Last updated
Tags
open-weightsmid-tiercachingmultimodal

Availability verified: listed on Moonshot's own page