LIVE Cheapest paid: Granite 4.0 Micro $0.017/Mtok in 174 models tracked Updated Jul 25, 2026
Jul 25, 2026
ModelPriceWatch$/Mtok
Pricing / Alibaba / Qwen3.7-Max

Qwen3.7-Max

by Alibaba

Current  Intro price flagship mid tier

Pricing · per 1M tokens

Input
$2.50
per million tokens
Output
$7.50
per million tokens
Cached input$0.250/Mtok (10% of input — prompt caching)
Blended avg cost*$5.00/Mtok
Qwen3.7-Max is available on 2 hosts. See the full cheapest-first comparison on the consolidated Qwen3.7-Max page →

Overview

Alibaba's current proprietary flagship — a reasoning-first agent model with a 1M-token context and native extended thinking, built for long-horizon autonomy (up to ~35 hours / 1,000+ tool calls) across coding and workflow automation. API-only (no open weights). $2.50/$7.50 per 1M tokens list; cached input $0.25 (90% off). Text-only.

Capabilities

Input 1/5
Text ✓ Image Audio Video PDF
Output 1/5
Text ✓ Image Audio Video Embedding
Features 2/9
Prompt caching ✓ Reasoning Coding Fast inference Long context ✓ Open weights Multimodal Web search Realtime

Benchmark performance

39.8th
percentile vs all tracked models · 4 independent measurements
8
percentile / $/Mtok
GPQA Diamond
68%
SWE-Bench Verified
45%
Humanity's Last Exam
25%
ARC-AGI 2
28%
AIME 2025
70%
MMMLU
80%
BFCL
68%
HumanEval
84%
MATH 500
76%
Independent composite scores
  • AA Intelligence Index: 46 (Artificial Analysis composite across reasoning, knowledge and coding evals)
  • AA Agentic Index: 30.6 (Tool use, planning, autonomy)
  • AA-Omniscience Index: 14.1 (−100–100)
  • GDPval-AA v2: 1271.3 ELO (Real-world work tasks, human baseline = 1000)
How it stacks up
  • Ranks #70 of 126 comparably-measured models by percentile score, across 4 independent measurements
  • Ranks #95 of 126 comparably-tested models by normalized performance per dollar
  • Strongest at HumanEval — 84%, #31 of 89
Inference speed: 100 tokens/sec

Source: Artificial Analysis, Alibaba, Qwen model card · Updated Apr 1, 2026 · See full rankings →

Specifications

ProviderAlibaba
Context window1M tokens
Modalitytext
ParametersProprietary
Open sourceNo — proprietary
ReleasedMay 20, 2026
Status Current
Availability verifiedJul 25, 2026 — listed on Alibaba's own page
Last updatedJul 25, 2026
Tagsflagship caching

Cost calculator

1K tokens (in)$0.003
1K tokens (out)$0.007
100K tokens (in)$0.250
100K tokens (out)$0.750
1M tokens (in)$2.50
1M tokens (out)$7.50
10M tokens (blended)$50.00
Full calculator →

At a glance

Input$2.50/M
Output$7.50/M
Blended$5.00/M
Context1M
Tiermid

Embed this price

A live, auto-updating badge for your README or docs:

Qwen3.7-Max live API price

Price history

Last updated: Jul 25, 2026 · price links to the official provider pricing page (source above)

No price changes detected — this model's pricing has been stable since we began tracking it.

Compare Qwen3.7-Max

Featured in these rankings

Related models