ModelPriceWatch.com
Last scan 2026-09-15 Models tracked 259 Providers 32 Cheapest paid Granite 4.0 H Micro $0.017/Mtok in Every price links to its source

Qwen3.7-Max

by Alibaba

Current flagship mid tier

Today's price · per 1M tokens

Input

$2.50

per 1M tokens

Output

$7.50

per 1M tokens

Blended

$3.75

blended $/1M — 3:1 weighted input:output

Cached input

$0.250

10% of input — prompt caching

How this price is scoped: $2.50 in / $7.50 out per 1M is Alibaba's standing rate. The Model Studio price table prints no time-limited reduction against this SKU — its qwen3.7-plus rows on the same table do carry one, so the absence is that table's own signal, not a gap in our reading.

Source: official Alibaba pricing · read MODELPRICEWATCH.COM · 2026-09-15

Price receipt

We read Alibaba's own pricing page on and found qwen3.7-max listed with a price on the same row — that read is where the number above comes from. We keep the page text we read; its content hash is 1c6db23291.

Available on 2 hosts

cheapest blended first · $ per 1M tokens

Qwen3.7-Max is sold by 2 providers. Prices are per 1M tokens (blended = (3×input + 1×output) ÷ 4). The first-party row is the model maker; “vs first-party” shows each host’s blended price relative to it.

Qwen3.7-Max price on every host still selling it, per 1M tokens, cheapest blended first
Host Input Output Blended vs first-party
Together details → $2.00 $6.00 $3.00 -20%
Alibaba first-party $2.50 $7.50 $3.75
Each host links to its provider page and official pricing. Prices refresh twice daily. MODELPRICEWATCH.COM · 2026-09-15

Overview

Alibaba's current proprietary flagship — a reasoning-first agent model with a 1M-token context and native extended thinking, built for long-horizon autonomy (up to ~35 hours / 1,000+ tool calls) across coding and workflow automation. API-only (no open weights). $2.50/$7.50 per 1M tokens list; cached input $0.25 (90% off). Text-only.

Capabilities

struck through = not supported
Input 1/5
Text Image Audio Video PDF
Output 1/5
Text Image Audio Video Embedding
Features 2/9
Prompt caching Reasoning Coding Fast inference Long context Open weights Multimodal Web search Realtime

Benchmark performance

accuracy % · higher is better
Percentile vs all tracked models 50.9th3 independent measurements
Percentile per $/Mtok 13.6
GPQA Diamond vendor
68%
SWE-Bench Verified vendor
45%
Humanity's Last Exam vendor
25%
ARC-AGI 2 vendor
28%
AIME 2025 vendor
70%
MMMLU vendor
80%
BFCL vendor
68%
HumanEval vendor
84%
MATH 500 vendor
76%

Every per-benchmark score we hold for this model is reported by its own vendor, so none is counted toward an accuracy average or any ranking. The percentile above is its standing across the independent composites below.

Independent composite scores
  • AA Intelligence Index: 29.9 (Artificial Analysis composite across reasoning, knowledge and coding evals)
  • AA-Omniscience Index: 13.5 (−100–100)
  • GDPval-AA v2: 1271.8 ELO (Real-world work tasks, human baseline = 1000)
How it stacks up
  • Ranks #71 of 181 comparably-measured models by percentile score, across 3 independent measurements
  • Ranks #142 of 181 comparably-tested models by normalized performance per dollar

100 tokens/sec output

Source: Artificial Analysis, Alibaba, Qwen model card · updated · See full rankings →

Specifications

Provider
Alibaba
Context window
1M tokens
Modality
text
Parameters
Proprietary
Open source
No — proprietary
Released
Status
Current
Last updated
Tags
flagshipcaching

Availability verified: listed on Alibaba's own page