Today's price · per 1M tokens
Input
$0.160
per 1M tokens
Output
$0.640
per 1M tokens
Blended
$0.280
blended $/1M — 3:1 weighted input:output
Overview
Qwen3-32B served by Alibaba, the model's maker, on Model Studio. $0.16/$0.64 per 1M tokens, one flat rate covering both thinking and non-thinking modes. Open weights, 128K-token context, text-only.
Deploy this open model on rented GPUs
Open weights mean you can self-host instead of paying per-token API prices. These platforms let you serve it on demand — often cheaper at scale.
Partner links — we may earn a commission if you sign up, at no cost to you. We only list platforms we'd recommend regardless.
Capabilities
struck through = not supportedBenchmark performance
accuracy % · higher is betterWe hold no per-benchmark accuracy scores for this model yet, so it has no accuracy average. That is a gap in our coverage, not a sign the model is untested — independent evaluators often publish a composite index for a new model long before releasing its per-benchmark numbers. The percentile above is its standing across the independent composites below.
- AA Intelligence Index: 11.4 (Artificial Analysis composite across reasoning, knowledge and coding evals)
- AA Agentic Index: 1.8 (Tool use, planning, autonomy)
- AA-Omniscience Index: -50.4 (−100–100)
- LMSYS Chatbot Arena: 1347.2 ELO (Human preference)
- Ranks #128 of 150 comparably-measured models by percentile score, across 4 independent measurements
- Ranks #58 of 150 comparably-tested models by normalized performance per dollar
Source: Artificial Analysis · updated Aug 7, 2026 · See full rankings →
Specifications
- Provider
- Alibaba
- Context window
- 131K tokens
- Modality
- text
- Parameters
- 32B
- Open source
- Yes — open weights available
- Released
- Apr 28, 2025
- Status
- Current
- Last updated
- Aug 7, 2026
- Tags
Availability verified: Aug 7, 2026 — listed on Alibaba's own page