ModelPriceWatch.com
Last scan 2026-08-01 Models tracked 184 Providers 27 Cheapest paid Granite 4.0 Micro $0.017/Mtok in Every price links to its source

Granite 4.0 Micro vs Qwen-Turbo

Side-by-side comparison of API pricing, specs, benchmarks, and capabilities

Granite 4.0 Micro is 53.4% cheaper on blended cost ($0.041 vs $0.088/Mtok)
Specification
by IBM
Overview
StatusCurrent budget Open weights Current budget
Released Oct 19, 2025 Oct 1, 2025
Pricing per million tokens
Input $0.017/Mtok $0.050/Mtok
Output $0.112/Mtok $0.200/Mtok
Blended avg $0.041/Mtok $0.088/Mtok
Specifications
Context window 128K tokens 1M tokens
Parameters Proprietary Proprietary
Speed (TPS)
Modalities
Input
text
text
Benchmarks sources: Artificial Analysis, IBM model card, IBM model card (5-shot) / Artificial Analysis
Independent composite scores measured for both models — each on its own scale
AA Intelligence Index 2 6
Per-benchmark results published for Granite 4.0 Micro only — no independent per-benchmark scores exist for Qwen-Turbo yet
MMMLU 55.14 vendor
BFCL 59.98 vendor
HumanEval 80 vendor
Providers
Available from
IBM — $0.017/$0.112/Mtok
Alibaba — $0.050/$0.200/Mtok

Cost at scale

1M tokens · 50/50 input/output
Projected cost of Granite 4.0 Micro vs Qwen-Turbo at increasing token volumes
VolumeGranite 4.0 MicroQwen-TurboSavings
1M tokens $0.04 $0.09 $0.05 (57.1%)
10M tokens $0.41 $0.88 $0.47 (53.7%)
100M tokens $4.08 $8.75 $4.68 (53.5%)
1000M tokens $40.75 $87.5 $46.75 (53.4%)

When to pick which

distilled from the pricing and spec data above
Pick Granite 4.0 Micro if…
  • cost dominates: $0.041/Mtok blended vs $0.088 — 53.4% less on the same 50/50 token mix
  • you want open weights — self-host it, fine-tune it, or exit the API entirely; Qwen-Turbo is closed
Pick Qwen-Turbo if…
  • you need the longer context: 1M tokens vs 128K (7.8×)
Try Granite 4.0 Micro on IBM

The cheaper option here — Granite 4.0 Micro costs $0.041/Mtok blended on IBM.

Get API key →

Summary

Granite 4.0 Micro by IBM costs $0.017/Mtok input and $0.112/Mtok output, with a 128K-token context window. It supports text input.

Qwen-Turbo by Alibaba costs $0.050/Mtok input and $0.200/Mtok output, with a 1M-token context window. It supports text input.

On a blended cost basis, Granite 4.0 Micro is 53.4% cheaper than Qwen-Turbo.

Note: Pricing is per million tokens. Actual costs vary with usage patterns, prompt caching, and batch discounts. Always verify against official provider pricing pages.

More comparisons

pairs sharing a model with this page