ModelPriceWatch.com
Last scan 2026-08-03 Models tracked 188 Providers 28 Cheapest paid Granite 4.0 Micro $0.017/Mtok in Every price links to its source

Qwen3.8-Max

by Alibaba · 2.4T parameters

Current new · 0d Intro price flagship mid tier

Today's price · per 1M tokens

Input

$2.00

per 1M tokens

Output

$6.00

per 1M tokens

Blended

$3.00

blended $/1M — 3:1 weighted input:output

Cached input

$0.200

10% of input — prompt caching

Source: official Alibaba pricing · read Aug 3, 2026 MODELPRICEWATCH.COM · 2026-08-03

Overview

Alibaba's current proprietary flagship — a 2.4T-parameter mixture-of-experts model with a 1M-token context that takes text, image and video input and returns text. $2.00/$6.00 per 1M tokens, a single flat tier across the whole 1M context, with thinking tokens billed at the output rate. Open weights are announced but not yet released.

Capabilities

struck through = not supported
Input 3/5
Text ✓ Image ✓ Audio Video ✓ PDF
Output 1/5
Text ✓ Image Audio Video Embedding
Features 3/9
Prompt caching ✓ Reasoning Coding Fast inference Long context ✓ Open weights Multimodal ✓ Web search Realtime

Benchmark performance

accuracy % · higher is better
GPQA Diamond vendor
92.6%
Humanity's Last Exam vendor
43.6%

Every per-benchmark score we hold for this model is reported by its own vendor, so none is counted toward an accuracy average or any ranking.

Source: Qwen official launch blog (Full Benchmark Table) · updated Aug 3, 2026 · See full rankings →

Specifications

Provider
Alibaba
Context window
1M tokens
Modality
text, image, video
Parameters
2.4T
Open source
No — proprietary
Released
Aug 3, 2026
Status
Current
Last updated
Aug 3, 2026
Tags
flagshipcaching

Availability verified: Aug 3, 2026 — listed on Alibaba's own page