ModelPriceWatch.com
Last scan 2026-08-16 Models tracked 205 Providers 31 Cheapest paid Granite 4.0 H Micro $0.017/Mtok in Every price links to its source

Qwen3.7-Plus vs GLM-4.5

Side-by-side comparison of API pricing, specs, benchmarks, and capabilities

Qwen3.7-Plus is 30% cheaper on blended cost ($0.700 vs $1.00/Mtok)
Specification
Qwen3.7-PlusIntro price
3 providers: Alibaba $0.400/$1.60 Fireworks $0.400/$1.60 Together $0.320/$1.28
by Z.AI
Overview
StatusCurrent mid tier Current mid tier Open weights
Released May 26, 2026 Jan 1, 2025
Pricing per million tokens
Input $0.400/Mtok $0.600/Mtok
Output $1.60/Mtok $2.20/Mtok
Blended avg $0.700/Mtok $1.00/Mtok
Cached input $0.110/Mtok
Price basis Alibaba Model Studio International (Singapore) list price, 0<Token≤256K tier. The 256K<Token≤1M tier is billed higher ($1.20/$4.80). On a limited-time 20% off to $0.32 in / $1.28 out per 1M (list price tracked, as on alibaba-qwen3-7-max). Context-cache hits discount input only; no cache rate is printed for this SKU, so cached input is left unset rather than assumed.
Specifications
Context window 1M tokens 128K tokens
Parameters Proprietary Proprietary
Speed (TPS)
Modalities
Input
text
text
Benchmarks sources: Artificial Analysis / LMSYS Chatbot Arena (UC Berkeley)
Independent composite scores measured for both models — each on its own scale
AA-Omniscience Index (−100–100) 1.1 -27.4
LMSYS Chatbot Arena 1458.5 ELO 1411.3 ELO
AA Intelligence Index 39.4
AA Agentic Index 20.7
Providers
Available from
Alibaba — $0.400/$1.60/Mtok
Fireworks — $0.400/$1.60/Mtok
Together — $0.320/$1.28/Mtok
Z.AI — $0.600/$2.20/Mtok

Cost at scale

1M tokens · 50/50 input/output
Projected cost of Qwen3.7-Plus vs GLM-4.5 at increasing token volumes
VolumeQwen3.7-PlusGLM-4.5Savings
1M tokens $0.7 $1 $0.3 (30%)
10M tokens $7 $10 $3 (30%)
100M tokens $70 $100 $30 (30%)
1000M tokens $700 $1000 $300 (30%)

When to pick which

distilled from the pricing and spec data above
Pick Qwen3.7-Plus if…
  • cost dominates: $0.700/Mtok blended vs $1.00 — 30% less on the same 50/50 token mix
  • you need the longer context: 1M tokens vs 128K (7.8×)
  • provider flexibility: available from 3 providers, so you can price-shop hosts or fail over
Pick GLM-4.5 if…
  • your workload re-reads context (agents, RAG, long chats): cached input costs $0.110/Mtok — 82% off its list input price, a discount Qwen3.7-Plus doesn't offer
  • you want open weights — self-host it, fine-tune it, or exit the API entirely; Qwen3.7-Plus is closed
Try Qwen3.7-Plus on Alibaba

The cheaper option here — Qwen3.7-Plus costs $0.700/Mtok blended on Alibaba.

Get API key →

Summary

Qwen3.7-Plus by Alibaba costs $0.400/Mtok input and $1.60/Mtok output, with a 1M-token context window. It supports text input and is available from 3 providers.

GLM-4.5 by Z.AI costs $0.600/Mtok input and $2.20/Mtok output, with a 128K-token context window. It supports text input.

On a blended cost basis, Qwen3.7-Plus is 30% cheaper than GLM-4.5.It also has a larger context window.

Note: Pricing is per million tokens. Actual costs vary with usage patterns, prompt caching, and batch discounts. Always verify against official provider pricing pages.

More comparisons

pairs sharing a model with this page