ModelPriceWatch.com
Last scan 2026-09-15 Models tracked 259 Providers 32 Cheapest paid Granite 4.0 H Micro $0.017/Mtok in Every price links to its source

GLM-5.1

by Z.AI

Current mid tier open weights mid tier

Today's price · per 1M tokens

Input

$1.40

per 1M tokens

Output

$4.40

per 1M tokens

Blended

$2.15

blended $/1M — 3:1 weighted input:output

Cached input

$0.260

19% of input — prompt caching

Source: official Z.AI pricing · read MODELPRICEWATCH.COM · 2026-09-15

Price receipt

We read Z.AI's own pricing page on and found GLM-5.1 listed with a price on the same row — that read is where the number above comes from. We keep the page text we read; its content hash is aac79925f8.

Available on 2 hosts

cheapest blended first · $ per 1M tokens

GLM-5.1 is sold by 2 providers. Prices are per 1M tokens (blended = (3×input + 1×output) ÷ 4). The first-party row is the model maker; “vs first-party” shows each host’s blended price relative to it.

GLM-5.1 price on every host still selling it, per 1M tokens, cheapest blended first
Host Input Output Blended vs first-party
Fireworks details → $1.40 $4.40 $2.15 0%
Z.AI first-party $1.40 $4.40 $2.15
Each host links to its provider page and official pricing. Prices refresh twice daily. MODELPRICEWATCH.COM · 2026-09-15

Overview

GLM-5.1 model. $1.40/$4.40 per 1M; cached input $0.26.

Run it yourself

Deploy this open model on rented GPUs

Open weights mean you can self-host instead of paying per-token API prices. These platforms let you serve it on demand — often cheaper at scale.

Partner links — we may earn a commission if you sign up, at no cost to you. We only list platforms we'd recommend regardless.

Capabilities

struck through = not supported
Input 1/5
Text Image Audio Video PDF
Output 1/5
Text Image Audio Video Embedding
Features 3/9
Prompt caching Reasoning Coding Fast inference Long context Open weights Multimodal Web search Realtime

Benchmark performance

accuracy % · higher is better
Percentile vs all tracked models 59.6th3 independent measurements
Percentile per $/Mtok 27.7
GPQA Diamond vendor
65%
SWE-Bench Verified vendor
42%
Humanity's Last Exam vendor
25%
ARC-AGI 2 vendor
28%
AIME 2025 vendor
68%
MMMLU vendor
78%
BFCL vendor
68%
HumanEval vendor
82%
MATH 500 vendor
74%

Every per-benchmark score we hold for this model is reported by its own vendor, so none is counted toward an accuracy average or any ranking. The percentile above is its standing across the independent composites below.

Independent composite scores
  • AA Intelligence Index: 26.4 (Artificial Analysis composite across reasoning, knowledge and coding evals)
  • AA-Omniscience Index: 0.8 (−100–100)
  • LMSYS Chatbot Arena: 1465.5 ELO (Human preference)
How it stacks up
  • Ranks #49 of 181 comparably-measured models by percentile score, across 3 independent measurements
  • Ranks #106 of 181 comparably-tested models by normalized performance per dollar

160 tokens/sec output

Source: Artificial Analysis, LMSYS Chatbot Arena (UC Berkeley), Z.AI model card · updated · See full rankings →

Specifications

Provider
Z.AI
Context window
128K tokens
Modality
text
Parameters
Proprietary
Open source
Yes — open weights available
Released
Status
Current
Last updated
Tags
open-weightsmid-tiercaching

Availability verified: listed on Z.AI's own page