Today's price · per 1M tokens
Input
$0.064
per 1M tokens
Output
$0.265
per 1M tokens
Blended
$0.114
blended $/1M — 3:1 weighted input:output
Price receipt
We read IBM's own pricing page on and found granite-4h-small listed with a price on the same row — that read is where the number above comes from. We keep the page text we read; its content hash is 2a14968924.
Overview
IBM Granite 4 H Small model. $0.0636/$0.265 per 1M tokens.
Deploy this open model on rented GPUs
Open weights mean you can self-host instead of paying per-token API prices. These platforms let you serve it on demand — often cheaper at scale.
Partner links — we may earn a commission if you sign up, at no cost to you. We only list platforms we'd recommend regardless.
Capabilities
struck through = not supportedBenchmark performance
accuracy % · higher is betterEvery per-benchmark score we hold for this model is reported by its own vendor, so none is counted toward an accuracy average or any ranking. The percentile above is its standing across the independent composites below.
- AA Intelligence Index: 5 (Artificial Analysis composite across reasoning, knowledge and coding evals)
- Ranks #181 of 187 comparably-measured models by percentile score, across 1 independent measurement
- Ranks #104 of 187 comparably-tested models by normalized performance per dollar
414.5 tokens/sec output
Source: Artificial Analysis, IBM model card, IBM model card (5-shot) · updated · See full rankings →
Specifications
- Provider
- IBM
- Context window
- 128K tokens
- Modality
- text
- Parameters
- Proprietary
- Open source
- Yes — open weights available
- Released
- Status
- Current
- Last updated
- Tags
Availability verified: — listed on IBM's own page