ModelPriceWatch.com
Last scan 2026-08-01 Models tracked 181 Providers 27 Cheapest paid Granite 4.0 Micro $0.017/Mtok in Every price links to its source

Qwen3.7-Max

by Together

Current Intro price mid tier mid tier

Today's price · per 1M tokens

Input

$1.25

per 1M tokens

Output

$3.75

per 1M tokens

Blended

$2.50

avg of input & output $/1M

Cached input

$0.130

10% of input — prompt caching

Source: official Together pricing · read Jul 25, 2026 MODELPRICEWATCH.COM · 2026-08-01
Qwen3.7-Max is available on 2 hosts. See the full cheapest-first comparison on the consolidated Qwen3.7-Max page →

Overview

Alibaba's proprietary flagship agent model — reasoning-first with a 1M-token context and native extended thinking, built for long-horizon autonomy (up to ~35 hours / 1,000+ tool calls) across coding and workflow automation. API-only (no open weights). Text-only. Together lists it at $1.25/$3.75 per 1M tokens — Alibaba's 50% launch promo off the $2.50/$7.50 list.

Capabilities

struck through = not supported
Input 1/5
Text ✓ Image Audio Video PDF
Output 1/5
Text ✓ Image Audio Video Embedding
Features 2/9
Prompt caching ✓ Reasoning Coding Fast inference Long context ✓ Open weights Multimodal Web search Realtime

Benchmark performance

accuracy % · higher is better
GPQA Diamond vendor
68%
SWE-Bench Verified vendor
45%
Humanity's Last Exam vendor
25%
ARC-AGI 2 vendor
28%
AIME 2025 vendor
70%
MMMLU vendor
80%
BFCL vendor
68%
HumanEval vendor
84%
MATH 500 vendor
76%

Every per-benchmark score we hold for this model is reported by its own vendor, so none is counted toward an accuracy average or any ranking.

100 tokens/sec output

Source: Together AI, Qwen model card · updated Apr 1, 2026 · See full rankings →

Specifications

Provider
Together
Context window
1M tokens
Modality
text
Parameters
Proprietary
Open source
No — proprietary
Released
May 20, 2026
Status
Current
Last updated
Jul 25, 2026
Tags
cachingmid-tier

Availability verified: Not re-verified. We have no recent evidence that Together still lists this model — treat availability as unconfirmed.