ModelPriceWatch.com
Last scan 2026-07-29 Models tracked 182 Providers 27 Cheapest paid Granite 4.0 Micro $0.017/Mtok in Every price links to its source

Reka Flash vs GLM-4.5

Side-by-side comparison of API pricing, specs, benchmarks, and capabilities

Specification
by Reka
by Z.AI
Overview
StatusCurrent mid tier Current mid tier Open weights
Released Jun 1, 2025 Jan 1, 2025
Pricing per million tokens
Input $0.800/Mtok $0.600/Mtok
Output $2.00/Mtok $2.20/Mtok
Blended avg $1.40/Mtok $1.40/Mtok
Cached input $0.110/Mtok
Specifications
Context window 128K tokens 128K tokens
Parameters Proprietary Proprietary
Speed (TPS)
Modalities
Input
textimagevideoaudio
text
Benchmarks sources: LMSYS Chatbot Arena (2025-07-15 archive), Reka tech report / LMSYS Chatbot Arena (UC Berkeley)
Independent composite scores measured for both models — each on its own scale
LMSYS Chatbot Arena 1220.6 ELO 1410.7 ELO
Per-benchmark results published for Reka Flash only — no independent per-benchmark scores exist for GLM-4.5 yet
HumanEval 72 vendor
Providers
Available from
Reka — $0.800/$2.00/Mtok
Z.AI — $0.600/$2.20/Mtok

Cost at scale

1M tokens · 50/50 input/output
Projected cost of Reka Flash vs GLM-4.5 at increasing token volumes
VolumeReka FlashGLM-4.5Savings
1M tokens $1.4 $1.4 $0 (0%)
10M tokens $14 $14 $0 (0%)
100M tokens $140 $140 $0 (0%)
1000M tokens $1400 $1400 $0 (0%)
Try Reka Flash on Reka

The cheaper option here — Reka Flash costs $1.40/Mtok blended on Reka.

Get API key →

Summary

Reka Flash by Reka costs $0.800/Mtok input and $2.00/Mtok output, with a 128K-token context window. It supports text, image, video, audio input.

GLM-4.5 by Z.AI costs $0.600/Mtok input and $2.20/Mtok output, with a 128K-token context window. It supports text input.

Note: Pricing is per million tokens. Actual costs vary with usage patterns, prompt caching, and batch discounts. Always verify against official provider pricing pages.

More comparisons

pairs sharing a model with this page