ModelPriceWatch.com
Last scan 2026-08-22 Models tracked 220 Providers 31 Cheapest paid Granite 4.0 H Micro $0.017/Mtok in Every price links to its source

Qwen3-32B vs Qwen3.6-27B

Side-by-side comparison of API pricing, specs, benchmarks, and capabilities

Qwen3-32B is 69.6% cheaper on blended cost ($0.365 vs $1.20/Mtok)
Specification
by Groq
3 providers: Alibaba $0.160/$0.640 Groq $0.290/$0.590 DeepInfra $0.080/$0.280
by Groq
2 providers: Alibaba $0.600/$3.60 Groq $0.600/$3.00
Overview
StatusCurrent fast Open weights Current fast Open weights
Released Jan 1, 2026 Jan 1, 2026
Pricing per million tokens
Input $0.290/Mtok $0.600/Mtok
Output $0.590/Mtok $3.00/Mtok
Blended avg $0.365/Mtok $1.20/Mtok
Specifications
Context window 128K tokens 131K tokens
Parameters 32B 27B
Speed (TPS) 662 tok/s 500 tok/s
Modalities
Input
text
text
Benchmarks sources: LMSYS Chatbot Arena (UC Berkeley) / Artificial Analysis, Qwen model card
Independent composite scores measured for both models — each on its own scale
AA Intelligence Index 11.4 37.7
AA Agentic Index 1.8 27.5
AA-Omniscience Index (−100–100) -50.4 -20
LMSYS Chatbot Arena 1347.2 ELO
Per-benchmark results published for Qwen3.6-27B only — no independent per-benchmark scores exist for Qwen3-32B yet
GPQA Diamond 87.8 vendor
SWE-Bench Verified 77.2 vendor
Providers
Available from
Alibaba — $0.160/$0.640/Mtok
Groq — $0.290/$0.590/Mtok
DeepInfra — $0.080/$0.280/Mtok
Alibaba — $0.600/$3.60/Mtok
Groq — $0.600/$3.00/Mtok

Cost at scale

1M tokens · 50/50 input/output
Projected cost of Qwen3-32B vs Qwen3.6-27B at increasing token volumes
VolumeQwen3-32BQwen3.6-27BSavings
1M tokens $0.37 $1.2 $0.84 (70%)
10M tokens $3.65 $12 $8.35 (69.6%)
100M tokens $36.5 $120 $83.5 (69.6%)
1000M tokens $365 $1200 $835 (69.6%)

When to pick which

distilled from the pricing and spec data above
Pick Qwen3-32B if…
  • cost dominates: $0.365/Mtok blended vs $1.20 — 69.6% less on the same 50/50 token mix
  • latency-sensitive traffic: 662 tok/s measured throughput vs 500
  • provider flexibility: available from 3 providers, so you can price-shop hosts or fail over
Pick Qwen3.6-27B if…
  • this pairing gives Qwen3.6-27B no edge on price, context, cache discount, or measured performance — pick it only for qualitative fit (ecosystem, compliance, model behavior on your prompts)
Try Qwen3-32B on Groq

The cheaper option here — Qwen3-32B costs $0.365/Mtok blended on Groq.

Get API key →

Summary

Qwen3-32B by Groq costs $0.290/Mtok input and $0.590/Mtok output, with a 128K-token context window. It supports text input and is available from 3 providers.

Qwen3.6-27B by Groq costs $0.600/Mtok input and $3.00/Mtok output, with a 131K-token context window. It supports text input and is available from 2 providers.

On a blended cost basis, Qwen3-32B is 69.6% cheaper than Qwen3.6-27B.

Note: Pricing is per million tokens. Actual costs vary with usage patterns, prompt caching, and batch discounts. Always verify against official provider pricing pages.

More comparisons

pairs sharing a model with this page