LIVE Cheapest paid: Granite 4.0 Micro $0.017/Mtok in 174 models tracked Updated Jul 25, 2026
Jul 25, 2026
ModelPriceWatch$/Mtok
Pricing / Compare / Gemini 3.6 Flash vs Qwen3.7-Max

Gemini 3.6 Flash vs Qwen3.7-Max

Side-by-side comparison of API pricing, specs, benchmarks, and capabilities

Gemini 3.6 Flash is 10% cheaper on blended cost ($4.50 vs $5.00/Mtok)
 
Gby Google
Qwen3.7-Max Intro price
2 providers: Alibaba $2.50/$7.50 Together $1.25/$3.75
Overview
StatusCurrent flagship Current flagship
Released Jul 21, 2026 May 20, 2026
Pricing per million tokens
Input $1.50/Mtok $2.50/Mtok
Output $7.50/Mtok $7.50/Mtok
Blended avg $4.50/Mtok $5.00/Mtok
Cached input $0.250/Mtok
Specifications
Context window 1M tokens 1M tokens
Parameters Proprietary Proprietary
Speed (TPS)
Modalities
Input
textimageaudiovideo
text
Benchmarks sources: Artificial Analysis / Artificial Analysis, Alibaba, Qwen model card
Independent composite scores measured for both models — each on its own scale
AA Intelligence Index 50.1 46
AA Agentic Index 38.7 30.6
AA-Omniscience Index (−100–100) 23.5 14.1
GDPval-AA v2 1423 ELO 1271.3 ELO
LMSYS Chatbot Arena 1485 ELO
Per-benchmark results published for Qwen3.7-Max only — no independent per-benchmark scores exist for Gemini 3.6 Flash yet
GPQA Diamond 68
SWE-Bench Verified 45
Humanity's Last Exam 25
ARC-AGI 2 28
AIME 2025 70
MMMLU 80
BFCL 68
HumanEval 84
MATH 500 76
Providers
Available from
Google — $1.50/$7.50/Mtok
Alibaba — $2.50/$7.50/Mtok
Together — $1.25/$3.75/Mtok

Cost at scale — 1M tokens (50/50 input/output)

VolumeGemini 3.6 FlashQwen3.7-MaxSavings
1M tokens $4.5 $5 $0.5 (10%)
10M tokens $45 $50 $5 (10%)
100M tokens $450 $500 $50 (10%)
1000M tokens $4500 $5000 $500 (10%)

Summary

Gemini 3.6 Flash by Google costs $1.50/Mtok input and $7.50/Mtok output, with a 1M-token context window. It supports text, image, audio, video input.

Qwen3.7-Max by Alibaba costs $2.50/Mtok input and $7.50/Mtok output, with a 1M-token context window. It supports text input and is available from 2 providers.

On a blended cost basis, Gemini 3.6 Flash is 10% cheaper than Qwen3.7-Max.

Note: Pricing is per million tokens. Actual costs vary with usage patterns, prompt caching, and batch discounts. Always verify against official provider pricing pages.