Qwen3.7-Max vs Claude Sonnet 5
Side-by-side comparison of API pricing, specs, benchmarks, and capabilities
| Specification |
by Alibaba
2 providers:
Alibaba $2.50/$7.50
Together $2.00/$6.00
|
by Anthropic
|
|---|---|---|
| Overview | ||
| Status | Current flagship | Current mid tier |
| Released | ||
| Pricing per million tokens | ||
| Input | $2.50/Mtok | $2.00/Mtok |
| Output | $7.50/Mtok | $10.00/Mtok |
| Blended avg | $3.75/Mtok | $4.00/Mtok |
| Cached input | $0.250/Mtok | $0.200/Mtok |
| Price basis | $2.50 in / $7.50 out per 1M is Alibaba's standing rate. The Model Studio price table prints no time-limited reduction against this SKU — its qwen3.7-plus rows on the same table do carry one, so the absence is that table's own signal, not a gap in our reading. | Standard pricing, per 1M tokens — not a promotional rate. $2/$10 was announced at launch as introductory pricing through Aug 31 2026, but Anthropic's pricing page now states that it "is now the standard price" and that "the previously scheduled increase to $3/$15 per million input/output tokens on September 1, 2026 will not occur". There is therefore no reversion date to watch. Verified 2026-08-13 against platform.claude.com/docs/en/about-claude/pricing. |
| Specifications | ||
| Context window | 1M tokens | 1M tokens |
| Parameters | Proprietary | Proprietary |
| Speed (TPS) | — | — |
| Modalities | ||
| Input |
text
|
textimage
|
| Benchmarks sources: Artificial Analysis, Alibaba, Qwen model card / Vellum LLM Leaderboard | ||
| GPQA Diamond | 68 vendor | 96.2 |
| SWE-Bench Verified | 45 vendor | 85.2 |
| Humanity's Last Exam | 25 vendor | 57.4 |
| HumanEval | 84 vendor | 85.2 |
| Avg benchmark score | — | 81.5 |
| Perf / dollar | — | 20.4 |
| Terminal-Bench 2.1 | — | 80.4 |
| BrowseComp | — | 84.7 |
| OSWorld-Verified | — | 81.2 |
| ARC-AGI 2 | 28 vendor | — |
| AIME 2025 | 70 vendor | — |
| MMMLU | 80 vendor | — |
| BFCL | 68 vendor | — |
| MATH 500 | 76 vendor | — |
| AutoBench | — | 13.5 |
| Independent composite scores each on its own scale — not part of the average above | ||
| AA Intelligence Index | 29.9 | 38.4 |
| AA Agentic Index | 23.9 | 44.3 |
| AA-Omniscience Index (−100–100) | 13.5 | 16.4 |
| GDPval-AA v2 | 1271.8 ELO | 1598.2 ELO |
| Providers | ||
| Available from |
Alibaba — $2.50/$7.50/Mtok
Together — $2.00/$6.00/Mtok
|
Anthropic — $2.00/$10.00/Mtok
|
Cost at scale
1M tokens · 50/50 input/output| Volume | Qwen3.7-Max | Claude Sonnet 5 | Savings |
|---|---|---|---|
| 1M tokens | $3.75 | $4 | $0.25 (6.3%) |
| 10M tokens | $37.5 | $40 | $2.5 (6.3%) |
| 100M tokens | $375 | $400 | $25 (6.3%) |
| 1000M tokens | $3750 | $4000 | $250 (6.3%) |
When to pick which
distilled from the pricing and spec data above- cost dominates: $3.75/Mtok blended vs $4.00 — 6.3% less on the same 50/50 token mix
- provider flexibility: available from 2 providers, so you can price-shop hosts or fail over
- this pairing gives Claude Sonnet 5 no edge on price, context, cache discount, or measured performance — pick it only for qualitative fit (ecosystem, compliance, model behavior on your prompts)
The cheaper option here — Qwen3.7-Max costs $3.75/Mtok blended on Alibaba.
Summary
Qwen3.7-Max by Alibaba costs $2.50/Mtok input and $7.50/Mtok output, with a 1M-token context window. It supports text input and is available from 2 providers.
Claude Sonnet 5 by Anthropic costs $2.00/Mtok input and $10.00/Mtok output, with a 1M-token context window. It supports text, image input.
On a blended cost basis, Qwen3.7-Max is 6.3% cheaper than Claude Sonnet 5.
The two aren't directly comparable on average benchmark score: Claude Sonnet 5 has published per-benchmark results, while Qwen3.7-Max does not yet — it is measured today on independent composites (see the table above).
Note: Pricing is per million tokens. Actual costs vary with usage patterns, prompt caching, and batch discounts. Always verify against official provider pricing pages.