Claude Sonnet 5
Anthropic · Released June 2026
Input Price
$3.00
per 1M tokens
Output Price
$15.00
per 1M tokens
Context Window
1.0M
tokens
Speed Score
88
out of 100
Benchmarks
Avg 92/100Pricing
Input
Prompt / context tokens
$3.00/ 1M
Output
Generated tokens
$15.00/ 1M
Batch API
Async processing discount
$1.50 / $7.50
50% off standard price
Best For
Claude Sonnet 5 in context
On a typical 70% input / 30% output mix, Claude Sonnet 5 blends to $6.60 per million tokens. That puts it above 72% of the 32 models tracked here, and 10 other models in the same mid-range tier undercut it. Tier is a capability label, not a price band — the spread inside one tier is often wider than the gap between tiers, which is why picking by tier alone tends to overspend.
Input runs $3.00 per million and output $15.00 — a 5.0x ratio. That ratio matters more than either number on its own: a summarisation or classification workload reads far more than it writes and will track the input price, while a code-generation or long-form writing workload inverts that and will track the output price. Compare models on the ratio your own traffic actually has, not on the input price alone.
The benchmark profile is unusually flat — coding leads at 95/100 but only 5 points separate the strongest category from the weakest. That makes Claude Sonnet 5 a safe default for mixed workloads where you cannot predict in advance which capability a request will lean on, and a poor choice if you need a specialist.
The context window is 1M tokens and the throughput score is 88/100. Context only earns its keep if you fill it, and filling it is also what makes a request expensive — a large window is an option, not a discount. For a mid-range-tier model, judge the speed score against what you are doing: it decides everything for interactive chat and almost nothing for overnight batch work.
Worth checking before you commit: DeepSeek R1 scores at least as well on combined coding and reasoning and blends to $1.04 per million against Claude Sonnet 5's $6.60. That does not make Claude Sonnet 5 the wrong choice — context window, latency, vendor lock-in, and specific capabilities all sit outside a benchmark score — but it does mean the price difference needs a reason behind it.
The closest comparison is GPT-5.6 Terra from OpenAI, at $5.00 per million against $6.60. 3 models in this tier score higher on combined coding and reasoning, which is the useful framing: the question is rarely whether a model is good, it is whether it is the cheapest thing that is good enough for the specific work you are sending it.
What it costs in production
At 10M + 2M tokens a month — a realistic mid-size production workload — Claude Sonnet 5 runs about $60.00. Routing the latency-tolerant share through the batch API at 50% off would bring that to roughly $30.00, which is usually a bigger saving than switching models.
Price History
No price changes since launch
Strengths
- Near-Opus coding quality at Sonnet pricing
- 1M context window
- Adaptive thinking on by default
Avoid For
- Deepest multi-step reasoning — the Opus tier still wins
- Very high-volume routing (Haiku is 3x cheaper)