ModelPriceWatch.com
Last scan 2026-09-14 Models tracked 260 Providers 32 Cheapest paid Granite 4.0 H Micro $0.017/Mtok in Every price links to its source

Qwen3.7-Max vs Gemini 3.1 Pro

Side-by-side comparison of API pricing, specs, benchmarks, and capabilities

Qwen3.7-Max is 16.7% cheaper on blended cost ($3.75 vs $4.50/Mtok)
Specification
2 providers: Alibaba $2.50/$7.50 Together $2.00/$6.00
Overview
StatusCurrent flagship Current flagship
Released
Pricing per million tokens
Input$2.50/Mtok $2.00/Mtok
Output $7.50/Mtok $12.00/Mtok
Blended avg $3.75/Mtok $4.50/Mtok
Cached input $0.250/Mtok $0.200/Mtok
Price basis $2.50 in / $7.50 out per 1M is Alibaba's standing rate. The Model Studio price table prints no time-limited reduction against this SKU — its qwen3.7-plus rows on the same table do carry one, so the absence is that table's own signal, not a gap in our reading.
Specifications
Context window 1M tokens 2M tokens
Parameters Proprietary Proprietary
Speed (TPS)
Modalities
Input
text
textimageaudiovideo
Benchmarks sources: Artificial Analysis, Alibaba, Qwen model card / Vellum LLM Leaderboard
GPQA Diamond 68 vendor 94.3
Humanity's Last Exam 25 vendor 44.4
ARC-AGI 2 28 vendor 77.1
HumanEval 84 vendor 80.6
Avg benchmark score 74.8
Perf / dollar 16.6
SWE-Bench Verified 45 vendor
Terminal-Bench 2.1 70.3
MCP Atlas 69.2
BrowseComp 85.9
OSWorld-Verified 76.2
AIME 2025 70 vendor
MMMLU 80 vendor
BFCL 68 vendor
MATH 500 76 vendor
Independent composite scores each on its own scale — not part of the average above
AA Intelligence Index 29.9 30.4
AA Agentic Index 23.9 10.3
AA-Omniscience Index (−100–100) 13.5 31.9
GDPval-AA v2 1271.8 ELO
LMSYS Chatbot Arena 1486.8 ELO
Providers
Available from
Alibaba — $2.50/$7.50/Mtok
Together — $2.00/$6.00/Mtok
Google — $2.00/$12.00/Mtok

Cost at scale

1M tokens · 50/50 input/output
Projected cost of Qwen3.7-Max vs Gemini 3.1 Pro at increasing token volumes
VolumeQwen3.7-MaxGemini 3.1 ProSavings
1M tokens $3.75 $4.5 $0.75 (16.7%)
10M tokens $37.5 $45 $7.5 (16.7%)
100M tokens $375 $450 $75 (16.7%)
1000M tokens $3750 $4500 $750 (16.7%)

When to pick which

distilled from the pricing and spec data above
Pick Qwen3.7-Max if…
  • cost dominates: $3.75/Mtok blended vs $4.50 — 16.7% less on the same 50/50 token mix
  • provider flexibility: available from 2 providers, so you can price-shop hosts or fail over
Pick Gemini 3.1 Pro if…
  • you need the longer context: 2M tokens vs 1M (2×)
Try Qwen3.7-Max on Alibaba

The cheaper option here — Qwen3.7-Max costs $3.75/Mtok blended on Alibaba.

Get API key →

Summary

Qwen3.7-Max by Alibaba costs $2.50/Mtok input and $7.50/Mtok output, with a 1M-token context window. It supports text input and is available from 2 providers.

Gemini 3.1 Pro by Google costs $2.00/Mtok input and $12.00/Mtok output, with a 2M-token context window. It supports text, image, audio, video input.

On a blended cost basis, Qwen3.7-Max is 16.7% cheaper than Gemini 3.1 Pro.

The two aren't directly comparable on average benchmark score: Gemini 3.1 Pro has published per-benchmark results, while Qwen3.7-Max does not yet — it is measured today on independent composites (see the table above).

Note: Pricing is per million tokens. Actual costs vary with usage patterns, prompt caching, and batch discounts. Always verify against official provider pricing pages.

More comparisons

pairs sharing a model with this page