ModelPriceWatch.com
Last scan 2026-08-01 Models tracked 181 Providers 27 Cheapest paid Granite 4.0 Micro $0.017/Mtok in Every price links to its source

Qwen3.5-397B-A17B

by Together · 397B (A17B) parameters

Retired mid tier open weights mid tier
Retired on Jun 29, 2026. No longer served on the endpoint we price — the figures below are kept for historical reference only. Replaced by MiniMax M3.
Availability: Together's changelog (June 15, 2026) scheduled Qwen/Qwen3.5-397B-A17B for removal from serverless after June 29, 2026, naming MiniMaxAI/MiniMax-M3 as the replacement; the June 29 entry confirms the removal. The model is still offered on on-demand DEDICATED endpoints, which are billed per minute — so the $0.60/$3.60 per-1M serverless price shown here is no longer obtainable.

Today's price · per 1M tokens

Input

$0.600

per 1M tokens

Output

$3.60

per 1M tokens

Blended

$2.10

avg of input & output $/1M

Cached input

$0.350

58% of input — prompt caching

Source: official Together pricing · read Jul 31, 2026 MODELPRICEWATCH.COM · 2026-08-01

Overview

Large MoE Qwen model (397B total, 17B active) on Together. $0.60/$3.60 per 1M tokens. Retired from Together serverless 2026-06-29.

Run it yourself

Deploy this open model on rented GPUs

Open weights mean you can self-host instead of paying per-token API prices. These platforms let you serve it on demand — often cheaper at scale.

Partner links — we may earn a commission if you sign up, at no cost to you. We only list platforms we'd recommend regardless.

Capabilities

struck through = not supported
Input 1/5
Text ✓ Image Audio Video PDF
Output 1/5
Text ✓ Image Audio Video Embedding
Features 3/9
Prompt caching ✓ Reasoning Coding Fast inference Long context ✓ Open weights ✓ Multimodal Web search Realtime

Benchmark performance

accuracy % · higher is better
Percentile vs all tracked models 27.8th3 independent measurements
Percentile per $/Mtok 13.2

We hold no per-benchmark accuracy scores for this model yet, so it has no accuracy average. That is a gap in our coverage, not a sign the model is untested — independent evaluators often publish a composite index for a new model long before releasing its per-benchmark numbers. The percentile above is its standing across the independent composites below.

Independent composite scores
  • AA Intelligence Index: 33.7 (Artificial Analysis composite across reasoning, knowledge and coding evals)
  • AA Agentic Index: 19.8 (Tool use, planning, autonomy)
  • AA-Omniscience Index: -29.8 (−100–100)
How it stacks up
  • Ranks #89 of 133 comparably-measured models by percentile score, across 3 independent measurements
  • Ranks #92 of 133 comparably-tested models by normalized performance per dollar

Source: Artificial Analysis · updated Jun 29, 2026 · See full rankings →

Specifications

Provider
Together
Context window
128K tokens
Modality
text
Parameters
397B (A17B)
Open source
Yes — open weights available
Released
Jan 1, 2026
Status
Retired ended Jun 29, 2026
Last updated
Jul 31, 2026
Tags
cachingmid-tieropen-weightsmoe

Availability verified: Not re-verified. We have no recent evidence that Together still lists this model — treat availability as unconfirmed.