ModelPriceWatch.com
Last scan 2026-08-10 Models tracked 201 Providers 31 Cheapest paid Granite 4.0 Micro $0.017/Mtok in Every price links to its source

KAT-Coder-Pro V2.5 vs Relace Search

Side-by-side comparison of API pricing, specs, benchmarks, and capabilities

KAT-Coder-Pro V2.5 is 13.6% cheaper on blended cost ($1.30 vs $1.50/Mtok)
Specification
Overview
StatusCurrent coding Current coding
Released Jul 10, 2026 Dec 8, 2025
Pricing per million tokens
Input $0.741/Mtok $1.00/Mtok
Output $2.96/Mtok $3.00/Mtok
Blended avg $1.30/Mtok $1.50/Mtok
Cached input $0.148/Mtok
Specifications
Context window 256K tokens 256K tokens
Parameters Proprietary Proprietary
Speed (TPS)
Modalities
Input
text
text
Providers
Available from
Kwaipilot — $0.741/$2.96/Mtok
Relace — $1.00/$3.00/Mtok

Cost at scale

1M tokens · 50/50 input/output
Projected cost of KAT-Coder-Pro V2.5 vs Relace Search at increasing token volumes
VolumeKAT-Coder-Pro V2.5Relace SearchSavings
1M tokens $1.3 $1.5 $0.2 (13.3%)
10M tokens $12.97 $15 $2.03 (13.5%)
100M tokens $129.67 $150 $20.33 (13.6%)
1000M tokens $1296.75 $1500 $203.25 (13.6%)

When to pick which

distilled from the pricing and spec data above
Pick KAT-Coder-Pro V2.5 if…
  • cost dominates: $1.30/Mtok blended vs $1.50 — 13.6% less on the same 50/50 token mix
  • your workload re-reads context (agents, RAG, long chats): cached input costs $0.148/Mtok — 80% off its list input price, a discount Relace Search doesn't offer
Pick Relace Search if…
  • this pairing gives Relace Search no edge on price, context, cache discount, or measured performance — pick it only for qualitative fit (ecosystem, compliance, model behavior on your prompts)

Summary

KAT-Coder-Pro V2.5 by Kwaipilot costs $0.741/Mtok input and $2.96/Mtok output, with a 256K-token context window. It supports text input.

Relace Search by Relace costs $1.00/Mtok input and $3.00/Mtok output, with a 256K-token context window. It supports text input.

On a blended cost basis, KAT-Coder-Pro V2.5 is 13.6% cheaper than Relace Search.

Note: Pricing is per million tokens. Actual costs vary with usage patterns, prompt caching, and batch discounts. Always verify against official provider pricing pages.

More comparisons

pairs sharing a model with this page