GPT-5 mini vs Qwen3.7-Plus
Side-by-side comparison of API pricing, specs, benchmarks, and capabilities
GPT-5 mini is 1.8% cheaper on blended cost ($0.688 vs $0.700/Mtok)
| Specification |
by OpenAI
|
Qwen3.7-PlusIntro price
by Alibaba
3 providers:
Alibaba $0.400/$1.60
Fireworks $0.400/$1.60
Together $0.320/$1.28
|
|---|---|---|
| Overview | ||
| Status | Current mid tier | Current mid tier |
| Released | Aug 7, 2025 | May 26, 2026 |
| Pricing per million tokens | ||
| Input | $0.250/Mtok | $0.400/Mtok |
| Output | $2.00/Mtok | $1.60/Mtok |
| Blended avg | $0.688/Mtok | $0.700/Mtok |
| Cached input | $0.025/Mtok | — |
| Price basis | — | Alibaba Model Studio International (Singapore) list price, 0<Token≤256K tier. The 256K<Token≤1M tier is billed higher ($1.20/$4.80). On a limited-time 20% off to $0.32 in / $1.28 out per 1M (list price tracked, as on alibaba-qwen3-7-max). Context-cache hits discount input only; no cache rate is printed for this SKU, so cached input is left unset rather than assumed. |
| Specifications | ||
| Context window | 400K tokens | 1M tokens |
| Parameters | Proprietary | Proprietary |
| Speed (TPS) | — | — |
| Modalities | ||
| Input |
textimage
|
text
|
| Benchmarks source: Artificial Analysis | ||
| Independent composite scores measured for both models — each on its own scale | ||
| AA Intelligence Index | 25.8 | 39.4 |
| AA Agentic Index | 19.6 | 20.7 |
| AA-Omniscience Index (−100–100) | -17.3 | 1.1 |
| LMSYS Chatbot Arena | — | 1456.3 ELO |
| Providers | ||
| Available from |
OpenAI — $0.250/$2.00/Mtok
|
Alibaba — $0.400/$1.60/Mtok
Fireworks — $0.400/$1.60/Mtok
Together — $0.320/$1.28/Mtok
|
Cost at scale
1M tokens · 50/50 input/output| Volume | GPT-5 mini | Qwen3.7-Plus | Savings |
|---|---|---|---|
| 1M tokens | $0.69 | $0.7 | $0.01 (1.4%) |
| 10M tokens | $6.88 | $7 | $0.13 (1.9%) |
| 100M tokens | $68.75 | $70 | $1.25 (1.8%) |
| 1000M tokens | $687.5 | $700 | $12.5 (1.8%) |
When to pick which
distilled from the pricing and spec data above
Pick GPT-5 mini if…
- cost dominates: $0.688/Mtok blended vs $0.700 — 1.8% less on the same 50/50 token mix
- your workload re-reads context (agents, RAG, long chats): cached input costs $0.025/Mtok — 90% off its list input price, a discount Qwen3.7-Plus doesn't offer
Pick Qwen3.7-Plus if…
- you need the longer context: 1M tokens vs 400K (2.5×)
- provider flexibility: available from 3 providers, so you can price-shop hosts or fail over
Try GPT-5 mini on OpenAI
The cheaper option here — GPT-5 mini costs $0.688/Mtok blended on OpenAI.
Summary
GPT-5 mini by OpenAI costs $0.250/Mtok input and $2.00/Mtok output, with a 400K-token context window. It supports text, image input.
Qwen3.7-Plus by Alibaba costs $0.400/Mtok input and $1.60/Mtok output, with a 1M-token context window. It supports text input and is available from 3 providers.
On a blended cost basis, GPT-5 mini is 1.8% cheaper than Qwen3.7-Plus.
Note: Pricing is per million tokens. Actual costs vary with usage patterns, prompt caching, and batch discounts. Always verify against official provider pricing pages.
More comparisons
pairs sharing a model with this page
Qwen3.7-Plus vs Mistral Large 3
7% price gap
Solar Pro 4 vs Qwen3.7-Plus
93% price gap
Solar Pro 4 vs GPT-5 mini
92% price gap
Solar Pro 3 vs Qwen3.7-Plus
63% price gap
Solar Pro 3 vs GPT-5 mini
62% price gap
GPT-5.6 Luna vs Qwen3.7-Plus
36% price gap
Qwen-Plus vs GPT-5 mini
13% price gap
GPT-5 mini vs Mistral Large 3
8% price gap