LIVE Cheapest paid: Granite 4.0 Micro $0.017/Mtok in 176 models tracked Updated Jul 22, 2026
Jul 22, 2026
ModelPriceWatch$/Mtok
Pricing / Compare / GLM-5.2 vs Gemini 3.6 Flash

GLM-5.2 vs Gemini 3.6 Flash

Side-by-side comparison of API pricing, specs, benchmarks, and capabilities

GLM-5.2 is 35.6% cheaper on blended cost ($2.90 vs $4.50/Mtok)
 
Zby Z.AI
2 providers: Together $1.40/$4.40 Z.AI $1.40/$4.40
Gby Google
Overview
StatusCurrent flagship Open weights Current flagship
Released Jun 13, 2026 Jul 21, 2026
Pricing per million tokens
Input $1.40/Mtok $1.50/Mtok
Output $4.40/Mtok $7.50/Mtok
Blended avg $2.90/Mtok $4.50/Mtok
Cached input $0.260/Mtok
Specifications
Context window 1M tokens 1M tokens
Parameters Proprietary Proprietary
Speed (TPS)
Modalities
Input
text
textimageaudiovideo
Providers
Available from
Together — $1.40/$4.40/Mtok
Z.AI — $1.40/$4.40/Mtok
Google — $1.50/$7.50/Mtok

Cost at scale — 1M tokens (50/50 input/output)

VolumeGLM-5.2Gemini 3.6 FlashSavings
1M tokens $2.9 $4.5 $1.6 (35.6%)
10M tokens $29 $45 $16 (35.6%)
100M tokens $290 $450 $160 (35.6%)
1000M tokens $2900 $4500 $1600 (35.6%)

Summary

GLM-5.2 by Z.AI costs $1.40/Mtok input and $4.40/Mtok output, with a 1M-token context window. It supports text input and is available from 2 providers.

Gemini 3.6 Flash by Google costs $1.50/Mtok input and $7.50/Mtok output, with a 1M-token context window. It supports text, image, audio, video input.

On a blended cost basis, GLM-5.2 is 35.6% cheaper than Gemini 3.6 Flash.

Note: Pricing is per million tokens. Actual costs vary with usage patterns, prompt caching, and batch discounts. Always verify against official provider pricing pages.