Qwen3.8-27B vs Muse Glimmer 30B
Side-by-side comparison of API pricing, specs, benchmarks, and capabilities
| Specification |
by Alibaba
|
by Fireworks
2 providers:
Fireworks $0.350/$1.50
Together $0.350/$1.50
|
|---|---|---|
| Overview | ||
| Status | Current mid tier Open weights | Current budget Open weights |
| Released | Aug 14, 2026 | Aug 9, 2026 |
| Pricing per million tokens | ||
| Input | $0.500/Mtok | $0.350/Mtok |
| Output | $3.00/Mtok | $1.50/Mtok |
| Blended avg | $1.13/Mtok | $0.637/Mtok |
| Cached input | — | $0.040/Mtok |
| Price basis | Alibaba Model Studio International list price, single 0<Token≤1M tier, read off the vendor's open-source Qwen3.8 table directly beneath the Qwen3.8-2.4T-A95B row. Alibaba marks the SKU as eligible for the context-caching discount but prints no cache rate in that table, so cached input is left unset rather than inferred — the same treatment as its Qwen3.8-2.4T-A95B sibling. Alibaba's own endpoint on OpenRouter quotes $0.425/$2.55; the number here is the maker's published rate card, not the gateway's. | — |
| Specifications | ||
| Context window | 1M tokens | 131K tokens |
| Parameters | 27B | 30B |
| Speed (TPS) | — | — |
| Modalities | ||
| Input |
text
|
text
|
| Benchmarks source: Artificial Analysis | ||
| Independent composite scores measured for both models — each on its own scale | ||
| AA Intelligence Index | 52 | 35.1 |
| AA Agentic Index | 50.9 | 22.9 |
| AA-Omniscience Index (−100–100) | -10 | -32.9 |
| GDPval-AA v2 | 1543 ELO | — |
| Providers | ||
| Available from |
Alibaba — $0.500/$3.00/Mtok
|
Fireworks — $0.350/$1.50/Mtok
Together — $0.350/$1.50/Mtok
|
Cost at scale
1M tokens · 50/50 input/output| Volume | Qwen3.8-27B | Muse Glimmer 30B | Savings |
|---|---|---|---|
| 1M tokens | $1.13 | $0.64 | $0.49 (43.6%) |
| 10M tokens | $11.25 | $6.38 | $4.88 (43.4%) |
| 100M tokens | $112.5 | $63.75 | $48.75 (43.3%) |
| 1000M tokens | $1125 | $637.5 | $487.5 (43.3%) |
When to pick which
distilled from the pricing and spec data above- you need the longer context: 1M tokens vs 131K (7.6×)
- cost dominates: $0.637/Mtok blended vs $1.13 — 43.3% less on the same 50/50 token mix
- your workload re-reads context (agents, RAG, long chats): cached input costs $0.040/Mtok — 89% off its list input price, a discount Qwen3.8-27B doesn't offer
- provider flexibility: available from 2 providers, so you can price-shop hosts or fail over
The cheaper option here — Muse Glimmer 30B costs $0.637/Mtok blended on Fireworks.
Summary
Qwen3.8-27B by Alibaba costs $0.500/Mtok input and $3.00/Mtok output, with a 1M-token context window. It supports text input.
Muse Glimmer 30B by Fireworks costs $0.350/Mtok input and $1.50/Mtok output, with a 131K-token context window. It supports text input and is available from 2 providers.
On a blended cost basis, Muse Glimmer 30B is 43.3% cheaper than Qwen3.8-27B.
Note: Pricing is per million tokens. Actual costs vary with usage patterns, prompt caching, and batch discounts. Always verify against official provider pricing pages.