ModelPriceWatch.com
Last scan 2026-08-09 Models tracked 198 Providers 30 Cheapest paid Granite 4.0 Micro $0.017/Mtok in Every price links to its source

Qwen-Flash vs Qwen-Plus

Side-by-side comparison of API pricing, specs, benchmarks, and capabilities

Qwen-Flash is 77.1% cheaper on blended cost ($0.138 vs $0.600/Mtok) — note: different model classes (budget vs mid tier)
Specification
Overview
StatusCurrent budget Current mid tier
Released Oct 1, 2025 Oct 1, 2025
Pricing per million tokens
Input $0.050/Mtok $0.400/Mtok
Output $0.400/Mtok $1.20/Mtok
Blended avg $0.138/Mtok $0.600/Mtok
Specifications
Context window 1M tokens 131K tokens
Parameters Proprietary Proprietary
Speed (TPS)
Modalities
Input
text
text
Benchmarks source: LMSYS Chatbot Arena (UC Berkeley)
Independent composite scores measured for both models — each on its own scale
LMSYS Chatbot Arena 1396.8 ELO 1346.1 ELO
Providers
Available from
Alibaba — $0.050/$0.400/Mtok
Alibaba — $0.400/$1.20/Mtok

Cost at scale

1M tokens · 50/50 input/output
Projected cost of Qwen-Flash vs Qwen-Plus at increasing token volumes
VolumeQwen-FlashQwen-PlusSavings
1M tokens $0.14 $0.6 $0.46 (76.7%)
10M tokens $1.38 $6 $4.63 (77.2%)
100M tokens $13.75 $60 $46.25 (77.1%)
1000M tokens $137.5 $600 $462.5 (77.1%)

When to pick which

distilled from the pricing and spec data above
Pick Qwen-Flash if…
  • cost dominates: $0.138/Mtok blended vs $0.600 — 77.1% less on the same 50/50 token mix
  • you need the longer context: 1M tokens vs 131K (7.6×)
Pick Qwen-Plus if…
  • this pairing gives Qwen-Plus no edge on price, context, cache discount, or measured performance — pick it only for qualitative fit (ecosystem, compliance, model behavior on your prompts)
Try Qwen-Flash on Alibaba

The cheaper option here — Qwen-Flash costs $0.138/Mtok blended on Alibaba.

Get API key →

Summary

Qwen-Flash by Alibaba costs $0.050/Mtok input and $0.400/Mtok output, with a 1M-token context window. It supports text input.

Qwen-Plus by Alibaba costs $0.400/Mtok input and $1.20/Mtok output, with a 131K-token context window. It supports text input.

On a blended cost basis, Qwen-Flash is 77.1% cheaper than Qwen-Plus.It also has a larger context window.

Note: Pricing is per million tokens. Actual costs vary with usage patterns, prompt caching, and batch discounts. Always verify against official provider pricing pages.

More comparisons

pairs sharing a model with this page