Qwen3.8-Flash
by Alibaba
Today's price · per 1M tokens
Input
$0.150
per 1M tokens
Output
$0.470
per 1M tokens
Blended
$0.230
blended $/1M — 3:1 weighted input:output
How this price is scoped: Alibaba Model Studio International list price, the SKU's single 0<Token≤1M tier — unlike Qwen3.6-Flash and Qwen3.7-Flash this row is not tiered by request size. The row is labelled context-cache eligible but the table prints no cache rate for it, so cached input is left unset rather than assumed; unlike its Qwen3.7 predecessor it carries no 50% batch-inference label. A free quota of 1M tokens (90 days) applies in Singapore only and is a trial credit, not a list price.
Price receipt
We read Alibaba's own pricing page on Aug 27, 2026 and found qwen3.8-flash listed with a price on the same row — that read is where the number above comes from, recorded as a new model. We keep the page text we read; its content hash is 43447f123f.
Overview
Alibaba's current-generation cheap tier on Model Studio, and the first Qwen flash SKU priced as one flat rate across its whole context: $0.15/$0.47 per 1M tokens on any request up to 1M tokens, with no request-size tiers. That makes it more expensive than Qwen3.7-Flash on short prompts (which starts at $0.03/$0.13 below 32K) but cheaper on long ones (Qwen3.7-Flash bills $0.20/$0.80 in its 256K-1M tier). 1M-token context, text-only.
Capabilities
struck through = not supportedSpecifications
- Provider
- Alibaba
- Context window
- 1M tokens
- Modality
- text
- Parameters
- Proprietary
- Open source
- No — proprietary
- Released
- Aug 26, 2026
- Status
- Current
- Last updated
- Aug 27, 2026
- Tags
Availability verified: Aug 27, 2026 — listed on Alibaba's own page