Qwen3-Max vs Command A
Side-by-side comparison of API pricing, specs, benchmarks, and capabilities
Qwen3-Max is 45.1% cheaper on blended cost ($2.40 vs $4.38/Mtok)
| Specification |
by Alibaba
|
by Cohere
|
|---|---|---|
| Overview | ||
| Status | Current flagship | Current flagship |
| Released | Jan 1, 2026 | Mar 13, 2025 |
| Pricing per million tokens | ||
| Input | $1.20/Mtok | $2.50/Mtok |
| Output | $6.00/Mtok | $10.00/Mtok |
| Blended avg | $2.40/Mtok | $4.38/Mtok |
| Price basis | Alibaba Model Studio International list price, 0<Token≤32K tier. Tiered above that: 32K<Token≤128K is $2.40/$12.00 and 128K<Token≤256K is $3.00/$15.00 per 1M tokens. | — |
| Specifications | ||
| Context window | 262K tokens | 256K tokens |
| Parameters | Proprietary | 111B |
| Speed (TPS) | — | — |
| Modalities | ||
| Input |
text
|
text
|
| Benchmarks sources: Alibaba model card / Cohere model card | ||
| GPQA Diamond | 70 vendor | 50 vendor |
| SWE-Bench Verified | 48 vendor | 25 vendor |
| Humanity's Last Exam | 25 vendor | 15 vendor |
| ARC-AGI 2 | 28 vendor | 12 vendor |
| AIME 2025 | 72 vendor | 45 vendor |
| MMMLU | 80 vendor | 74 vendor |
| BFCL | 70 vendor | 60 vendor |
| HumanEval | 86 vendor | 78 vendor |
| MATH 500 | 78 vendor | 62 vendor |
| Independent composite scores each on its own scale — not part of the average above | ||
| LMSYS Chatbot Arena | 1434.3 ELO | 1353.9 ELO |
| Providers | ||
| Available from |
Alibaba — $1.20/$6.00/Mtok
|
Cohere — $2.50/$10.00/Mtok
|
Cost at scale
1M tokens · 50/50 input/output| Volume | Qwen3-Max | Command A | Savings |
|---|---|---|---|
| 1M tokens | $2.4 | $4.38 | $1.98 (45.3%) |
| 10M tokens | $24 | $43.75 | $19.75 (45.1%) |
| 100M tokens | $240 | $437.5 | $197.5 (45.1%) |
| 1000M tokens | $2400 | $4375 | $1975 (45.1%) |
When to pick which
distilled from the pricing and spec data above
Pick Qwen3-Max if…
- cost dominates: $2.40/Mtok blended vs $4.38 — 45.1% less on the same 50/50 token mix
Pick Command A if…
- this pairing gives Command A no edge on price, context, cache discount, or measured performance — pick it only for qualitative fit (ecosystem, compliance, model behavior on your prompts)
Try Qwen3-Max on Alibaba
The cheaper option here — Qwen3-Max costs $2.40/Mtok blended on Alibaba.
Summary
Qwen3-Max by Alibaba costs $1.20/Mtok input and $6.00/Mtok output, with a 262K-token context window. It supports text input.
Command A by Cohere costs $2.50/Mtok input and $10.00/Mtok output, with a 256K-token context window. It supports text input.
On a blended cost basis, Qwen3-Max is 45.1% cheaper than Command A.It also has a larger context window.
Note: Pricing is per million tokens. Actual costs vary with usage patterns, prompt caching, and batch discounts. Always verify against official provider pricing pages.
More comparisons
pairs sharing a model with this page
Gemini 3 Flash Preview vs Command A
74% price gap
Grok 4.3 vs Command A
64% price gap
Command A vs GLM-5.2
51% price gap
Reka Core vs Command A
31% price gap
Grok 4.5 vs Command A
31% price gap
Command A vs Gemini 3.5 Flash
23% price gap
GLM-5.2 vs Qwen3-Max
10% price gap
Command A vs Gemini 3.1 Pro
3% price gap