LIVE Cheapest paid: Granite 4.0 Micro $0.017/Mtok in 174 models tracked Updated Jul 25, 2026
Jul 25, 2026
ModelPriceWatch$/Mtok
Pricing / Compare / Gemini 3 Flash Preview vs Qwen3.7-Max

Gemini 3 Flash Preview vs Qwen3.7-Max

Side-by-side comparison of API pricing, specs, benchmarks, and capabilities

Gemini 3 Flash Preview is 65% cheaper on blended cost ($1.75 vs $5.00/Mtok)
 
Gby Google
Qwen3.7-Max Intro price
2 providers: Alibaba $2.50/$7.50 Together $1.25/$3.75
Overview
StatusPreview flagship Current flagship
Released Dec 17, 2025 May 20, 2026
Pricing per million tokens
Input $0.500/Mtok $2.50/Mtok
Output $3.00/Mtok $7.50/Mtok
Blended avg $1.75/Mtok $5.00/Mtok
Cached input $0.050/Mtok $0.250/Mtok
Specifications
Context window 1M tokens 1M tokens
Parameters Proprietary Proprietary
Speed (TPS)
Modalities
Input
textimageaudiovideo
text
Benchmarks sources: Artificial Analysis, Google model card, Google model card (no tools) / Artificial Analysis, Alibaba, Qwen model card
GPQA Diamond 90.4 68
SWE-Bench Verified 78 45
Humanity's Last Exam 33.7 25
ARC-AGI 2 28
AIME 2025 70
MMMLU 80
BFCL 68
HumanEval 84
MATH 500 76
Independent composite scores each on its own scale — not part of the average above
AA Intelligence Index 27 46
AA Agentic Index 30.6
AA-Omniscience Index (−100–100) 14.1
GDPval-AA v2 1271.3 ELO
LMSYS Chatbot Arena 1473.1 ELO
Providers
Available from
Google — $0.500/$3.00/Mtok
Alibaba — $2.50/$7.50/Mtok
Together — $1.25/$3.75/Mtok

Cost at scale — 1M tokens (50/50 input/output)

VolumeGemini 3 Flash PreviewQwen3.7-MaxSavings
1M tokens $1.75 $5 $3.25 (65%)
10M tokens $17.5 $50 $32.5 (65%)
100M tokens $175 $500 $325 (65%)
1000M tokens $1750 $5000 $3250 (65%)

Summary

Gemini 3 Flash Preview by Google costs $0.500/Mtok input and $3.00/Mtok output, with a 1M-token context window. It supports text, image, audio, video input.

Qwen3.7-Max by Alibaba costs $2.50/Mtok input and $7.50/Mtok output, with a 1M-token context window. It supports text input and is available from 2 providers.

On a blended cost basis, Gemini 3 Flash Preview is 65% cheaper than Qwen3.7-Max. It also has a larger context window.

Note: Pricing is per million tokens. Actual costs vary with usage patterns, prompt caching, and batch discounts. Always verify against official provider pricing pages.