ModelPriceWatch.com
Last scan 2026-08-01 Models tracked 181 Providers 27 Cheapest paid Granite 4.0 Micro $0.017/Mtok in Every price links to its source

Qwen3 32B vs Qwen 3.6 27B

Side-by-side comparison of API pricing, specs, benchmarks, and capabilities

Qwen3 32B is 75.6% cheaper on blended cost ($0.440 vs $1.80/Mtok)
Specification
by Groq
2 providers: Groq $0.290/$0.590 DeepInfra $0.080/$0.280
by Groq
Overview
StatusCurrent fast Open weights Current fast Open weights
Released Jan 1, 2026 Jan 1, 2026
Pricing per million tokens
Input $0.290/Mtok $0.600/Mtok
Output $0.590/Mtok $3.00/Mtok
Blended avg $0.440/Mtok $1.80/Mtok
Specifications
Context window 128K tokens 128K tokens
Parameters 32B 27B
Speed (TPS) 662 tok/s 500 tok/s
Modalities
Input
text
text
Benchmarks sources: LMSYS Chatbot Arena (UC Berkeley) / Artificial Analysis, Qwen model card
Independent composite scores measured for both models — each on its own scale
AA Intelligence Index 11.5 37
AA Agentic Index 1.8
AA-Omniscience Index (−100–100) -50.5
LMSYS Chatbot Arena 1347.2 ELO
Per-benchmark results published for Qwen 3.6 27B only — no independent per-benchmark scores exist for Qwen3 32B yet
GPQA Diamond 87.8 vendor
SWE-Bench Verified 77.2 vendor
Providers
Available from
Groq — $0.290/$0.590/Mtok
DeepInfra — $0.080/$0.280/Mtok
Groq — $0.600/$3.00/Mtok

Cost at scale

1M tokens · 50/50 input/output
Projected cost of Qwen3 32B vs Qwen 3.6 27B at increasing token volumes
VolumeQwen3 32BQwen 3.6 27BSavings
1M tokens $0.44 $1.8 $1.36 (75.6%)
10M tokens $4.4 $18 $13.6 (75.6%)
100M tokens $44 $180 $136 (75.6%)
1000M tokens $440 $1800 $1360 (75.6%)

When to pick which

distilled from the pricing and spec data above
Pick Qwen3 32B if…
  • cost dominates: $0.440/Mtok blended vs $1.80 — 75.6% less on the same 50/50 token mix
  • latency-sensitive traffic: 662 tok/s measured throughput vs 500
  • provider flexibility: available from 2 providers, so you can price-shop hosts or fail over
Pick Qwen 3.6 27B if…
  • this pairing gives Qwen 3.6 27B no edge on price, context, cache discount, or measured performance — pick it only for qualitative fit (ecosystem, compliance, model behavior on your prompts)
Try Qwen3 32B on Groq

The cheaper option here — Qwen3 32B costs $0.440/Mtok blended on Groq.

Get API key →

Summary

Qwen3 32B by Groq costs $0.290/Mtok input and $0.590/Mtok output, with a 128K-token context window. It supports text input and is available from 2 providers.

Qwen 3.6 27B by Groq costs $0.600/Mtok input and $3.00/Mtok output, with a 128K-token context window. It supports text input.

On a blended cost basis, Qwen3 32B is 75.6% cheaper than Qwen 3.6 27B.

Note: Pricing is per million tokens. Actual costs vary with usage patterns, prompt caching, and batch discounts. Always verify against official provider pricing pages.

More comparisons

pairs sharing a model with this page