Grok 4.3 vs Muse Spark 1.3
Side-by-side comparison of API pricing, specs, benchmarks, and capabilities
| Specification |
by xAI
|
by Meta
|
|---|---|---|
| Overview | ||
| Status | Current flagship | Current flagship |
| Released | ||
| Pricing per million tokens | ||
| Input | $1.25/Mtok | $1.25/Mtok |
| Output | $2.50/Mtok | $4.25/Mtok |
| Blended avg | $1.56/Mtok | $2.00/Mtok |
| Cached input | $0.200/Mtok | $0.150/Mtok |
| Price basis | — | Meta Model API standard-tier list price, per 1M tokens, with no long-context premium — the tier on which Meta states prompts and completions are not used for training. Meta's card prices 1.3, 1.2 and 1.1 off one shared standard-tier table. A cheaper contributor tier ($0.10 input / $0.20 output per 1M) trades training rights and a far lower rate limit (100 vs 3,000 requests/min) for the discount, so it is not the tracked list price. The "max" extended-reasoning effort level is a request parameter on this same model and Standard-tier rate, not a separately priced SKU. Search grounding is billed separately at $2.50 per 1,000 queries and is not included in the token rate. |
| Specifications | ||
| Context window | 1M tokens | 1M tokens |
| Parameters | Proprietary | Proprietary |
| Speed (TPS) | — | — |
| Modalities | ||
| Input |
textimage
|
textimagevideoaudio
|
| Benchmarks sources: xAI model card, Artificial Analysis / Artificial Analysis | ||
| Independent composite scores measured for both models — each on its own scale | ||
| AA Intelligence Index | 37.9 | 53 |
| AA Agentic Index | 17.3 | 55.6 |
| AA-Omniscience Index (−100–100) | 18 | 24.9 |
| GDPval-AA v2 | — | 1720.2 ELO |
| LMSYS Chatbot Arena | 1443 ELO | — |
| Per-benchmark results published for Grok 4.3 only — no independent per-benchmark scores exist for Muse Spark 1.3 yet | ||
| Avg benchmark score | 72.4 | — |
| Perf / dollar | 46.3 | — |
| GPQA Diamond | 80 | — |
| SWE-Bench Verified | 65 | — |
| Humanity's Last Exam | 38 | — |
| ARC-AGI 2 | 42 | — |
| AIME 2025 | 85 | — |
| MMMLU | 86 | — |
| BFCL | 78 | — |
| HumanEval | 90 | — |
| MATH 500 | 88 | — |
| Providers | ||
| Available from |
xAI — $1.25/$2.50/Mtok
|
Meta — $1.25/$4.25/Mtok
|
Cost at scale
1M tokens · 50/50 input/output| Volume | Grok 4.3 | Muse Spark 1.3 | Savings |
|---|---|---|---|
| 1M tokens | $1.56 | $2 | $0.44 (22%) |
| 10M tokens | $15.63 | $20 | $4.38 (21.9%) |
| 100M tokens | $156.25 | $200 | $43.75 (21.9%) |
| 1000M tokens | $1562.5 | $2000 | $437.5 (21.9%) |
When to pick which
distilled from the pricing and spec data above- cost dominates: $1.56/Mtok blended vs $2.00 — 21.9% less on the same 50/50 token mix
- your workload re-reads context (agents, RAG, long chats): cached input costs $0.150/Mtok — 88% off its list input price
The cheaper option here — Grok 4.3 costs $1.56/Mtok blended on xAI.
Summary
Grok 4.3 by xAI costs $1.25/Mtok input and $2.50/Mtok output, with a 1M-token context window. It supports text, image input.
Muse Spark 1.3 by Meta costs $1.25/Mtok input and $4.25/Mtok output, with a 1M-token context window. It supports text, image, video, audio input.
On a blended cost basis, Grok 4.3 is 21.9% cheaper than Muse Spark 1.3.
The two aren't directly comparable on average benchmark score: Grok 4.3 has published per-benchmark results, while Muse Spark 1.3 does not yet — it is measured today on independent composites (see the table above).
Note: Pricing is per million tokens. Actual costs vary with usage patterns, prompt caching, and batch discounts. Always verify against official provider pricing pages.