ModelPriceWatch.com
Last scan 2026-07-31 Models tracked 181 Providers 27 Cheapest paid Granite 4.0 Micro $0.017/Mtok in Every price links to its source

Ministral 3 3B vs GLM-4-32B-0414

Side-by-side comparison of API pricing, specs, benchmarks, and capabilities

Specification
by Z.AI
Overview
StatusCurrent budget Open weights Current budget Open weights
Released Sep 1, 2025 Apr 1, 2025
Pricing per million tokens
Input $0.100/Mtok $0.100/Mtok
Output $0.100/Mtok $0.100/Mtok
Blended avg $0.100/Mtok $0.100/Mtok
Specifications
Context window 128K tokens 128K tokens
Parameters 3B 32B
Speed (TPS)
Modalities
Input
text
text
Benchmarks sources: Artificial Analysis, Mistral model card (reasoning variant) / LMSYS Chatbot Arena (UC Berkeley)
GPQA Diamond 53.4 vendor
AIME 2025 72.1 vendor
Independent composite scores each on its own scale — not part of the average above
AA Intelligence Index 7
LMSYS Chatbot Arena 1342.9 ELO
Providers
Available from
Mistral — $0.100/$0.100/Mtok
Z.AI — $0.100/$0.100/Mtok

Cost at scale

1M tokens · 50/50 input/output
Projected cost of Ministral 3 3B vs GLM-4-32B-0414 at increasing token volumes
VolumeMinistral 3 3BGLM-4-32B-0414Savings
1M tokens $0.1 $0.1 $0 (0%)
10M tokens $1 $1 $0 (0%)
100M tokens $10 $10 $0 (0%)
1000M tokens $100 $100 $0 (0%)
Try Ministral 3 3B on Mistral

The cheaper option here — Ministral 3 3B costs $0.100/Mtok blended on Mistral.

Get API key →

Summary

Ministral 3 3B by Mistral costs $0.100/Mtok input and $0.100/Mtok output, with a 128K-token context window. It supports text input.

GLM-4-32B-0414 by Z.AI costs $0.100/Mtok input and $0.100/Mtok output, with a 128K-token context window. It supports text input.

Note: Pricing is per million tokens. Actual costs vary with usage patterns, prompt caching, and batch discounts. Always verify against official provider pricing pages.

More comparisons

pairs sharing a model with this page