Z.AI
Commercial provider · 14 models tracked · Founded 2019
Z.AI (formerly Zhipu AI) — makers of the GLM model family. Frontier and open-weight models with competitive pricing.
Z.AI pricing at a glanceAugust 2026 · $ per 1M tokens
Z.AI API pricing (August 2026): 14 current models range from $0.030 to $1.40 per 1M input tokens and $0.030 to $4.40 per 1M output tokens. The cheapest paid model is GLM-OCR at $0.030/1M input; the priciest is GLM-5.2 at $1.40/1M input / $4.40 output. Every price links to Z.AI's official pricing page and refreshes twice daily.
Today's Z.AI prices
14 models · cheapest blended firstSorted by blended cost (cheapest first). Prices per 1M tokens, August 2026 — every price links to Z.AI's official pricing page.
| Model | Blended* | Input | Output | Cached in | Relative cost | Context | Status |
|---|---|---|---|---|---|---|---|
| GLM-4.7-Flash | $0 |
$0 | $0 | — | 128K | Current | |
| GLM-OCR | $0.030 |
$0.030 | $0.030 | — | 128K | Current | |
| GLM-4-32B-0414 | $0.100 |
$0.100 | $0.100 | — | 128K | Current | |
| GLM-4.7-FlashX | $0.153 |
$0.070 | $0.400 | $0.010 | 128K | Current | |
| GLM-4.5-Air | $0.425 |
$0.200 | $1.10 | $0.030 | 128K | Current | |
| GLM-4.6V | $0.450 |
$0.300 | $0.900 | $0.050 | 128K | Current | |
| GLM-4.5 | $1.00 |
$0.600 | $2.20 | $0.110 | 128K | Current | |
| GLM-4.6 | $1.00 |
$0.600 | $2.20 | $0.110 | 200K | Current | |
| GLM-4.7 | $1.00 |
$0.600 | $2.20 | $0.110 | 128K | Current | |
| GLM-5 | $1.55 |
$1.00 | $3.20 | $0.200 | 128K | Current | |
| GLM-5-Turbo | $1.90 |
$1.20 | $4.00 | $0.240 | 128K | Current | |
| GLM-5V-Turbo | $1.90 |
$1.20 | $4.00 | $0.240 | 128K | Current | |
| GLM-5.1 | $2.15 |
$1.40 | $4.40 | $0.260 | 128K | Current | |
| GLM-5.2 | $2.15 |
$1.40 | $4.40 | $0.260 | 1M | Current |
Quick stats
- Models tracked
- 14
- Type
- commercial
- Founded
- 2019
- Cheapest model
- GLM-4.7-Flash
- Cheapest blended
- $0/M
Open-weight models
14 open-weight models available from this provider.
Self-host Z.AI's open-weight models
These models ship with open weights, so you can serve them yourself on rented GPUs instead of paying per-token API prices — often cheaper at scale.
Partner links — we may earn a commission if you sign up, at no cost to you. We only list platforms we'd recommend regardless.