Today's price · per 1M tokens
Input
$1.00
per 1M tokens
Output
$5.00
per 1M tokens
Blended
$3.00
avg of input & output $/1M
Cached input
$0.100
10% of input — prompt caching
Source: official Anthropic pricing · read Jun 25, 2026
MODELPRICEWATCH.COM · 2026-08-01
Overview
Anthropic's fast, low-cost tier — ~97 tokens/sec, built for high-volume classification, routing, and code review where heavy reasoning isn't needed. 200K context. $1/$5 per 1M tokens.
Capabilities
struck through = not supportedInput 2/5
Text ✓
Image ✓
Audio
Video
PDF
Output 1/5
Text ✓
Image
Audio
Video
Embedding
Features 4/9
Prompt caching ✓
Reasoning
Coding
Fast inference ✓
Long context ✓
Open weights
Multimodal ✓
Web search
Realtime
Benchmark performance
accuracy % · higher is betterAvg benchmark score65.5%
Perf per $/Mtok21.8
GPQA Diamond
73%
Terminal-Bench 2.1
27.5%
MCP Atlas
40.2%
AIME 2025
96.3%
MMMLU
83%
HumanEval
73.3%
MATH 500 vendor
78%
Independent composite scores
- AA Intelligence Index: 29.6 (Artificial Analysis composite across reasoning, knowledge and coding evals)
- AA Agentic Index: 16.4 (Tool use, planning, autonomy)
- AA-Omniscience Index: -4.2 (−100–100)
- LMSYS Chatbot Arena: 1412.2 ELO (Human preference)
How it stacks up
- Ranks #38 of 68 benchmarked models by average score
- Ranks #80 of 133 comparably-measured models by percentile score, across 10 independent measurements
- Ranks #98 of 133 comparably-tested models by normalized performance per dollar
- Strongest at AIME 2025 — 96.3%, #8 of 36
51 tokens/sec output 39.9s latency to first token (TTFT)
Source: Vellum LLM Leaderboard · updated Aug 1, 2026 · See full rankings →
Specifications
- Provider
- Anthropic
- Context window
- 200K tokens
- Modality
- text, image
- Parameters
- Proprietary
- Open source
- No — proprietary
- Released
- Oct 15, 2025
- Status
- Current
- Last updated
- Jun 25, 2026
- Tags
Availability verified: Jul 25, 2026 — listed on Anthropic's own page