ModelPriceWatch.com
Last scan 2026-09-22 Models tracked 266 Providers 35 Cheapest paid Granite 4.0 H Micro $0.017/Mtok in Every price links to its source

Qwen3-32B

by Alibaba · 32B parameters

Current budget open weights cheap tier

Today's price · per 1M tokens

Input

$0.160

per 1M tokens

Output

$0.640

per 1M tokens

Blended

$0.280

blended $/1M — 3:1 weighted input:output

How this price is scoped: Alibaba Model Studio International list price, one flat rate covering both Non-Thinking and Thinking modes, from the vendor's open-source Qwen table. No context-cache or batch rate is printed for this SKU, so cached input is left unset rather than assumed. The pricing page prints no context tier for this SKU; the 128K context is what Alibaba's own served endpoint reports. Alibaba's own gateway endpoint quotes a lower routed rate of $0.104/$0.416. Alibaba's pricing page states that it lists standard prices only and directs buyers to the Model Studio console for current offers, and this SKU's row carries no discount label, so the standard list rate is the one tracked here — the same basis as every other Alibaba row on this site.

Source: official Alibaba pricing · read MODELPRICEWATCH.COM · 2026-09-22

Price receipt

We read Alibaba's own pricing page on and found qwen3-32b listed with a price on the same row — that read is where the number above comes from. We keep the page text we read; its content hash is 4603ffb7a8.

Qwen3-32B is available on 3 hosts. See the full cheapest-first comparison on the consolidated Qwen3-32B page →

Overview

Qwen3-32B served by Alibaba, the model's maker, on Model Studio. $0.16/$0.64 per 1M tokens, one flat rate covering both thinking and non-thinking modes. Open weights, 128K-token context, text-only.

Run it yourself

Deploy this open model on rented GPUs

Open weights mean you can self-host instead of paying per-token API prices. These platforms let you serve it on demand — often cheaper at scale.

Partner links — we may earn a commission if you sign up, at no cost to you. We only list platforms we'd recommend regardless.

Capabilities

struck through = not supported
Input 1/4
Text Image Audio Video
Output
Text
Features 2/9
Prompt caching Reasoning Coding Fast inference Long context Open weights Multimodal Web search Realtime

Benchmark performance

accuracy % · higher is better
Percentile vs all tracked models 9.8th3 independent measurements
Percentile per $/Mtok 35

We hold no per-benchmark accuracy scores for this model yet, so it has no accuracy average. That is a gap in our coverage, not a sign the model is untested — independent evaluators often publish a composite index for a new model long before releasing its per-benchmark numbers. The percentile above is its standing across the independent composites below.

Independent composite scores
  • AA Intelligence Index: 7.2 (Artificial Analysis composite across reasoning, knowledge and coding evals)
  • AA-Omniscience Index: -50.4 (−100–100)
  • LMSYS Chatbot Arena: 1347.1 ELO (Human preference)
How it stacks up
  • Ranks #159 of 181 comparably-measured models by percentile score, across 3 independent measurements
  • Ranks #82 of 181 comparably-tested models by normalized performance per dollar

Source: Artificial Analysis, LMSYS Chatbot Arena (UC Berkeley) · updated · See full rankings →

Specifications

Provider
Alibaba
Context window
131K tokens
Modality
text
Parameters
32B
Open source
Yes — open weights available
Released
Status
Current
Last updated
Tags
budgetopen-weights

Availability verified: listed on Alibaba's own page