LIVE Cheapest paid: Granite 4.0 Micro $0.017/Mtok in 176 models tracked Updated Jul 22, 2026
Jul 22, 2026
ModelPriceWatch$/Mtok
Pricing / Compare / Qwen3-Max vs Gemini 3.6 Flash

Qwen3-Max vs Gemini 3.6 Flash

Side-by-side comparison of API pricing, specs, benchmarks, and capabilities

Qwen3-Max is 20% cheaper on blended cost ($3.60 vs $4.50/Mtok)
 
Gby Google
Overview
StatusCurrent flagship Current flagship
Released Jan 1, 2026 Jul 21, 2026
Pricing per million tokens
Input $1.20/Mtok $1.50/Mtok
Output $6.00/Mtok $7.50/Mtok
Blended avg $3.60/Mtok $4.50/Mtok
Specifications
Context window 262K tokens 1M tokens
Parameters Proprietary Proprietary
Speed (TPS)
Modalities
Input
text
textimageaudiovideo
Providers
Available from
Alibaba — $1.20/$6.00/Mtok
Google — $1.50/$7.50/Mtok

Cost at scale — 1M tokens (50/50 input/output)

VolumeQwen3-MaxGemini 3.6 FlashSavings
1M tokens $3.6 $4.5 $0.9 (20%)
10M tokens $36 $45 $9 (20%)
100M tokens $360 $450 $90 (20%)
1000M tokens $3600 $4500 $900 (20%)

Summary

Qwen3-Max by Alibaba costs $1.20/Mtok input and $6.00/Mtok output, with a 262K-token context window. It supports text input.

Gemini 3.6 Flash by Google costs $1.50/Mtok input and $7.50/Mtok output, with a 1M-token context window. It supports text, image, audio, video input.

On a blended cost basis, Qwen3-Max is 20% cheaper than Gemini 3.6 Flash.

Note: Pricing is per million tokens. Actual costs vary with usage patterns, prompt caching, and batch discounts. Always verify against official provider pricing pages.