ModelPriceWatch.com
Last scan 2026-08-03 Models tracked 188 Providers 28 Cheapest paid Granite 4.0 Micro $0.017/Mtok in Every price links to its source

Qwen3.8-Max vs Gemini 3.6 Flash

Side-by-side comparison of API pricing, specs, benchmarks, and capabilities

Specification
Qwen3.8-MaxIntro price
Overview
StatusCurrent flagship Current flagship
Released Aug 3, 2026 Jul 21, 2026
Pricing per million tokens
Input $2.00/Mtok $1.50/Mtok
Output $6.00/Mtok $7.50/Mtok
Blended avg $3.00/Mtok $3.00/Mtok
Cached input $0.200/Mtok
Specifications
Context window 1M tokens 1M tokens
Parameters 2.4T Proprietary
Speed (TPS)
Modalities
Input
textimagevideo
textimageaudiovideo
Benchmarks sources: Qwen official launch blog (Full Benchmark Table) / Artificial Analysis
GPQA Diamond 92.6 vendor
Humanity's Last Exam 43.6 vendor
Independent composite scores each on its own scale — not part of the average above
AA Intelligence Index 50.1
AA Agentic Index 38.7
AA-Omniscience Index (−100–100) 23.5
GDPval-AA v2 1423 ELO
LMSYS Chatbot Arena 1482.7 ELO
Providers
Available from
Alibaba — $2.00/$6.00/Mtok
Google — $1.50/$7.50/Mtok

Cost at scale

1M tokens · 50/50 input/output
Projected cost of Qwen3.8-Max vs Gemini 3.6 Flash at increasing token volumes
VolumeQwen3.8-MaxGemini 3.6 FlashSavings
1M tokens $3 $3 $0 (0%)
10M tokens $30 $30 $0 (0%)
100M tokens $300 $300 $0 (0%)
1000M tokens $3000 $3000 $0 (0%)

When to pick which

distilled from the pricing and spec data above
Pick Qwen3.8-Max if…
  • your workload re-reads context (agents, RAG, long chats): cached input costs $0.200/Mtok — 90% off its list input price, a discount Gemini 3.6 Flash doesn't offer
Pick Gemini 3.6 Flash if…
  • this pairing gives Gemini 3.6 Flash no edge on price, context, cache discount, or measured performance — pick it only for qualitative fit (ecosystem, compliance, model behavior on your prompts)
Try Qwen3.8-Max on Alibaba

The cheaper option here — Qwen3.8-Max costs $3.00/Mtok blended on Alibaba.

Get API key →

Summary

Qwen3.8-Max by Alibaba costs $2.00/Mtok input and $6.00/Mtok output, with a 1M-token context window. It supports text, image, video input.

Gemini 3.6 Flash by Google costs $1.50/Mtok input and $7.50/Mtok output, with a 1M-token context window. It supports text, image, audio, video input.

Note: Pricing is per million tokens. Actual costs vary with usage patterns, prompt caching, and batch discounts. Always verify against official provider pricing pages.

More comparisons

pairs sharing a model with this page