ModelPriceWatch.com
Last scan 2026-09-08 Models tracked 248 Providers 32 Cheapest paid Granite 4.0 H Micro $0.017/Mtok in Every price links to its source

Gemini 3.6 Flash vs Qwen3.7-Max

Side-by-side comparison of API pricing, specs, benchmarks, and capabilities

Gemini 3.6 Flash is 60% cheaper on blended cost ($1.50 vs $3.75/Mtok)
Specification
Gemini 3.6 FlashIntro price
Qwen3.7-MaxIntro price
2 providers: Alibaba $2.50/$7.50 Together $1.25/$3.75
Overview
StatusCurrent flagship Current flagship
Released
Pricing per million tokens
Input $0.750/Mtok $2.50/Mtok
Output $3.75/Mtok $7.50/Mtok
Blended avg $1.50/Mtok $3.75/Mtok
Cached input $0.075/Mtok $0.250/Mtok
Price basis Google's Gemini API pricing page prints a dated, two-phase schedule for this model: $0.75 per 1M input tokens through December 31, 2026, rising to $1.50 on January 1, 2027, and $3.75 per 1M output tokens through December 31, 2026, rising to $7.50. Context caching follows the same dates ($0.075, then $0.15). We publish the rate you would pay today and re-verify it before the changeover date. Checked 2026-08-13 against ai.google.dev/gemini-api/docs/pricing; the same page listed a flat $1.50/$7.50 on 2026-08-09, so this is a genuine 50% price cut rather than a correction to an earlier mistake. On a 50% launch promo to $1.25 in / $3.75 out per 1M (list price tracked).
Specifications
Context window 1M tokens 1M tokens
Parameters Proprietary Proprietary
Speed (TPS)
Modalities
Input
textimageaudiovideo
text
Benchmarks sources: Artificial Analysis, Google DeepMind Gemini 3.6 Flash model card, LMSYS Chatbot Arena (UC Berkeley) / Artificial Analysis, Alibaba, Qwen model card
Independent composite scores measured for both models — each on its own scale
AA Intelligence Index 40.3 36.6
AA Agentic Index 30.3 23.9
AA-Omniscience Index (−100–100) 22.1 13.5
GDPval-AA v2 1422.3 ELO 1271.8 ELO
LMSYS Chatbot Arena 1480.2 ELO
Per-benchmark results no per-benchmark score is published for both models — the composites above are what compares them
GPQA Diamond 68 vendor
SWE-Bench Verified 45 vendor
Terminal-Bench 2.1 78 vendor
OSWorld-Verified 83 vendor
Humanity's Last Exam 25 vendor
ARC-AGI 2 28 vendor
AIME 2025 70 vendor
MMMLU 80 vendor
BFCL 68 vendor
HumanEval 84 vendor
MATH 500 76 vendor
Providers
Available from
Google — $0.750/$3.75/Mtok
Alibaba — $2.50/$7.50/Mtok
Together — $1.25/$3.75/Mtok

Cost at scale

1M tokens · 50/50 input/output
Projected cost of Gemini 3.6 Flash vs Qwen3.7-Max at increasing token volumes
VolumeGemini 3.6 FlashQwen3.7-MaxSavings
1M tokens $1.5 $3.75 $2.25 (60%)
10M tokens $15 $37.5 $22.5 (60%)
100M tokens $150 $375 $225 (60%)
1000M tokens $1500 $3750 $2250 (60%)

When to pick which

distilled from the pricing and spec data above
Pick Gemini 3.6 Flash if…
  • cost dominates: $1.50/Mtok blended vs $3.75 — 60% less on the same 50/50 token mix
Pick Qwen3.7-Max if…
  • provider flexibility: available from 2 providers, so you can price-shop hosts or fail over

Price movement

recorded changes since we began tracking each model · last checked
Recorded price changes for Gemini 3.6 Flash and Qwen3.7-Max — how many moves, the direction and date of the latest one, and its old → new prices per million tokens
Model Changes Latest move Old → new $/M
Gemini 3.6 Flash 1 price cut on $1.5/$7.5 → $0.75/$3.75/Mtok
Qwen3.7-Max 0 — no move Stable since we began tracking — no change recorded
Only Gemini 3.6 Flash has changed price since we began tracking; Qwen3.7-Max has held steady. See all recent price moves → MODELPRICEWATCH.COM · 2026-09-08
Try Gemini 3.6 Flash on Google

The cheaper option here — Gemini 3.6 Flash costs $1.50/Mtok blended on Google.

Get API key →

Summary

Gemini 3.6 Flash by Google costs $0.750/Mtok input and $3.75/Mtok output, with a 1M-token context window. It supports text, image, audio, video input.

Qwen3.7-Max by Alibaba costs $2.50/Mtok input and $7.50/Mtok output, with a 1M-token context window. It supports text input and is available from 2 providers.

On a blended cost basis, Gemini 3.6 Flash is 60% cheaper than Qwen3.7-Max.

Note: Pricing is per million tokens. Actual costs vary with usage patterns, prompt caching, and batch discounts. Always verify against official provider pricing pages.

More comparisons

pairs sharing a model with this page