ModelPriceWatch.com
Last scan 2026-08-15 Models tracked 205 Providers 31 Cheapest paid Granite 4.0 H Micro $0.017/Mtok in Every price links to its source

Claude Sonnet 5

by Anthropic

Current mid tier mid tier

Today's price · per 1M tokens

Input

$2.00

per 1M tokens

Output

$10.00

per 1M tokens

Blended

$4.00

blended $/1M — 3:1 weighted input:output

Cached input

$0.200

10% of input — prompt caching

Source: official Anthropic pricing · read Aug 13, 2026 MODELPRICEWATCH.COM · 2026-08-15

Price receipt

We read Anthropic's own pricing page on Aug 13, 2026 and found Claude Sonnet 5 listed with a price on the same row — that read is where the number above comes from. We keep the page text we read; its content hash is 89a12b8bd5.

Overview

Anthropic's most agentic Sonnet yet — performance approaching Opus 4.8 at a fraction of the price, strong on reasoning, tool use, coding and autonomous agent workloads. $2/$10 per 1M tokens — first announced as launch pricing, now Anthropic's standard price; the increase once scheduled for Sep 1 2026 will not happen.

Capabilities

struck through = not supported
Input 2/5
Text ✓ Image ✓ Audio Video PDF
Output 1/5
Text ✓ Image Audio Video Embedding
Features 4/9
Prompt caching ✓ Reasoning ✓ Coding Fast inference Long context ✓ Open weights Multimodal ✓ Web search Realtime

Benchmark performance

accuracy % · higher is better
Avg benchmark score81.5%
Perf per $/Mtok20.4
GPQA Diamond
96.2%
SWE-Bench Verified
85.2%
Terminal-Bench 2.1
80.4%
BrowseComp
84.7%
OSWorld-Verified
81.2%
Humanity's Last Exam
57.4%
HumanEval
85.2%
AutoBench *
13.5%
Independent composite scores
  • AA Intelligence Index: 55.3 (Artificial Analysis composite across reasoning, knowledge and coding evals)
  • AA Agentic Index: 49.7 (Tool use, planning, autonomy)
  • AA-Omniscience Index: 16.4 (−100–100)
  • GDPval-AA v2: 1598.2 ELO (Real-world work tasks, human baseline = 1000)
How it stacks up
  • Ranks #10 of 69 benchmarked models by average score
  • Ranks #13 of 155 comparably-measured models by percentile score, across 11 independent measurements
  • Ranks #111 of 155 comparably-tested models by normalized performance per dollar
  • Strongest at GPQA Diamond — 96.2%, #1 of 63

56.3 tokens/sec output 20.69s latency to first token (TTFT)

Source: Vellum LLM Leaderboard · updated Aug 15, 2026 · See full rankings →

Specifications

Provider
Anthropic
Context window
1M tokens
Modality
text, image
Parameters
Proprietary
Open source
No — proprietary
Released
Jun 30, 2026
Status
Current
Last updated
Aug 13, 2026
Tags
reasoningmid-tiermultimodalagentic

Availability verified: Aug 13, 2026 — listed on Anthropic's own page