ModelPriceWatch.com
Last scan 2026-09-17 Models tracked 261 Providers 33 Cheapest paid Granite 4.0 H Micro $0.017/Mtok in Every price links to its source

Qwen3.8-Max

by Alibaba · 2.4T parameters

Current new · 45d flagship mid tier

Today's price · per 1M tokens

Input

$2.00

per 1M tokens

Output

$6.00

per 1M tokens

Blended

$3.00

blended $/1M — 3:1 weighted input:output

Cached input

$0.200

10% of input — prompt caching

How this price is scoped: Alibaba Model Studio International list price, single 0<Token≤1M tier — unlike qwen3-max this SKU is not tiered by request size. Explicit context-cache hits bill at 10% of the input rate (the cached_input_per_mtok we track); implicit-cache hits bill at 20%.

Source: official Alibaba pricing · read MODELPRICEWATCH.COM · 2026-09-17

Price receipt

We read Alibaba's own pricing page on and found qwen3.8-max listed with a price on the same row — that read is where the number above comes from. We keep the page text we read; its content hash is 167fe0b7ad.

Overview

Alibaba's current proprietary flagship — a 2.4T-parameter mixture-of-experts model with a 1M-token context that takes text, image and video input and returns text. $2.00/$6.00 per 1M tokens, a single flat tier across the whole 1M context, with thinking tokens billed at the output rate. The open weights shipped on 2026-08-08 as Qwen3.8-2.4T-A95B, which Alibaba now also sells as its own Model Studio SKU at the same $2.00/$6.00 rate; the price here is for this proprietary multimodal SKU.

Capabilities

struck through = not supported
Input 3/5
Text Image Audio Video PDF
Output 1/5
Text Image Audio Video Embedding
Features 3/9
Prompt caching Reasoning Coding Fast inference Long context Open weights Multimodal Web search Realtime

Benchmark performance

accuracy % · higher is better
Percentile vs all tracked models 83.3th4 independent measurements
Percentile per $/Mtok 27.8
GPQA Diamond vendor
92.6%
Humanity's Last Exam vendor
43.6%

Every per-benchmark score we hold for this model is reported by its own vendor, so none is counted toward an accuracy average or any ranking. The percentile above is its standing across the independent composites below.

Independent composite scores
  • AA Intelligence Index: 40.3 (Artificial Analysis composite across reasoning, knowledge and coding evals)
  • AA-Omniscience Index: 3.4 (−100–100)
  • GDPval-AA v2: 1735.2 ELO (Real-world work tasks, human baseline = 1000)
  • LMSYS Chatbot Arena: 1480.6 ELO (Human preference)
How it stacks up
  • Ranks #8 of 181 comparably-measured models by percentile score, across 4 independent measurements
  • Ranks #105 of 181 comparably-tested models by normalized performance per dollar

Source: Artificial Analysis, LMSYS Chatbot Arena (UC Berkeley), Qwen official launch blog (Full Benchmark Table) · updated · See full rankings →

Specifications

Provider
Alibaba
Context window
1M tokens
Modality
text, image, video
Parameters
2.4T
Open source
No — proprietary
Released
Status
Current
Last updated
Tags
flagshipcaching

Availability verified: listed on Alibaba's own page