Reka Flash vs GLM-4.5
Side-by-side comparison of API pricing, specs, benchmarks, and capabilities
GLM-4.5 is 9.1% cheaper on blended cost ($1.00 vs $1.10/Mtok)
| Specification |
by Reka
|
by Z.AI
|
|---|---|---|
| Overview | ||
| Status | Current mid tier | Current mid tier Open weights |
| Released | ||
| Pricing per million tokens | ||
| Input | $0.800/Mtok | $0.600/Mtok |
| Output | $2.00/Mtok | $2.20/Mtok |
| Blended avg | $1.10/Mtok | $1.00/Mtok |
| Cached input | — | $0.110/Mtok |
| Specifications | ||
| Context window | 128K tokens | 128K tokens |
| Parameters | Proprietary | Proprietary |
| Speed (TPS) | — | — |
| Modalities | ||
| Input |
textimagevideoaudio
|
text
|
| Benchmarks sources: LMSYS Chatbot Arena (2025-07-15 archive), Reka tech report / LMSYS Chatbot Arena (UC Berkeley) | ||
| Independent composite scores measured for both models — each on its own scale | ||
| LMSYS Chatbot Arena | 1220.6 ELO | 1411.3 ELO |
| AA-Omniscience Index (−100–100) | — | -27.4 |
| Per-benchmark results published for Reka Flash only — no independent per-benchmark scores exist for GLM-4.5 yet | ||
| HumanEval | 72 vendor | — |
| Providers | ||
| Available from |
Reka — $0.800/$2.00/Mtok
|
Z.AI — $0.600/$2.20/Mtok
|
Cost at scale
1M tokens · 50/50 input/output| Volume | Reka Flash | GLM-4.5 | Savings |
|---|---|---|---|
| 1M tokens | $1.1 | $1 | $0.1 (9.1%) |
| 10M tokens | $11 | $10 | $1 (9.1%) |
| 100M tokens | $110 | $100 | $10 (9.1%) |
| 1000M tokens | $1100 | $1000 | $100 (9.1%) |
When to pick which
distilled from the pricing and spec data above
Pick Reka Flash if…
- this pairing gives Reka Flash no edge on price, context, cache discount, or measured performance — pick it only for qualitative fit (ecosystem, compliance, model behavior on your prompts)
Pick GLM-4.5 if…
- cost dominates: $1.00/Mtok blended vs $1.10 — 9.1% less on the same 50/50 token mix
- your workload re-reads context (agents, RAG, long chats): cached input costs $0.110/Mtok — 82% off its list input price, a discount Reka Flash doesn't offer
- you want open weights — self-host it, fine-tune it, or exit the API entirely; Reka Flash is closed
Try GLM-4.5 on Z.AI
The cheaper option here — GLM-4.5 costs $1.00/Mtok blended on Z.AI.
Summary
Reka Flash by Reka costs $0.800/Mtok input and $2.00/Mtok output, with a 128K-token context window. It supports text, image, video, audio input.
GLM-4.5 by Z.AI costs $0.600/Mtok input and $2.20/Mtok output, with a 128K-token context window. It supports text input.
On a blended cost basis, GLM-4.5 is 9.1% cheaper than Reka Flash.
Note: Pricing is per million tokens. Actual costs vary with usage patterns, prompt caching, and batch discounts. Always verify against official provider pricing pages.
More comparisons
pairs sharing a model with this page
Qwen-Plus vs Reka Flash
46% price gap
DeepSeek V4 Pro vs Reka Flash
44% price gap
Mistral Large 3 vs Reka Flash
32% price gap
Sonar vs Reka Flash
9% price gap
GLM-4.7-Flash vs Ministral 3 3B
100% price gap
Claude Fable 5 vs DeepSeek V4 Pro
90% price gap
Claude Fable 5.1 vs DeepSeek V4 Pro
90% price gap
GPT-5.6 Terra vs GPT-5.6 Luna
90% price gap