ModelPriceWatch.com
Last scan 2026-09-22 Models tracked 267 Providers 35 Cheapest paid Granite 4.0 H Micro $0.017/Mtok in Every price links to its source

Grok 4

by xAI

Retired flagship mid tier
Retired on . No longer served on the endpoint we price — the figures below are kept for historical reference only. Replaced by Grok 4.3. Full retirement status & replacements →
Availability: Retired 2026-05-15; the grok-4 slug still resolves but is served by grok-4.3, so any price or benchmark attributed to grok-4 is wrong.

Today's price · per 1M tokens

Input

$3.00

per 1M tokens

Output

$15.00

per 1M tokens

Blended

$6.00

blended $/1M — 3:1 weighted input:output

Source: official xAI pricing · read MODELPRICEWATCH.COM · 2026-09-22

Price receipt

Historical price.This model is retiredwithdrawn on so the provider no longer sells it and its page can no longer confirm this number: a price for a model no longer on sale is unconfirmable by definition, not missing. The figures above are the last prices we recorded for this model, kept as a historical record rather than a live quote.

Overview

Original Grok 4. $3/$15 per 1M tokens.

Capabilities

struck through = not supported
Input 2/4
Text Image Audio Video
Output
Text
Features 2/9
Prompt caching Reasoning Coding Fast inference Long context Open weights Multimodal Web search Realtime

Benchmark performance

accuracy % · higher is better
Avg benchmark score63%
Perf per $/Mtok10.5
GPQA Diamond
87.5%
SWE-Bench Verified
55%
Terminal-Bench 2.1
27.2%
LiveCodeBench
79%
Humanity's Last Exam
25.4%
ARC-AGI 2
16%
AIME 2025
91.7%
MMMLU
82%
BFCL
72%
HumanEval
75%
MATH 500
82%
Independent composite scores
  • LMSYS Chatbot Arena: 1411 ELO (Human preference)
How it stacks up
  • Ranks #50 of 84 benchmarked models by average score
  • Ranks #82 of 181 comparably-measured models by percentile score, across 12 independent measurements
  • Ranks #162 of 181 comparably-tested models by normalized performance per dollar
  • Strongest at MATH 500 — 82%, #4 of 24

52 tokens/sec output 13.3s latency to first token (TTFT)

Source: Vellum LLM Leaderboard · updated · See full rankings →

Specifications

Provider
xAI
Context window
256K tokens
Modality
text, image
Parameters
Proprietary
Open source
No — proprietary
Released
Status
Retired ended
Last updated
Tags
flagshipmultimodalretired

Availability verified: per xAI's own deprecation notice