HomeModelsClaude Haiku 4.5
AnthropicFast Tier Vision Batch 50% Off

Claude Haiku 4.5

Anthropic · Released October 2025

Input Price

$1.00

per 1M tokens

Output Price

$5.00

per 1M tokens

Context Window

200K

tokens

Speed Score

97

out of 100

Benchmarks

Avg 84/100
Coding83/100
Reasoning82/100
Extraction91/100
Creative82/100
Vision80/100

Pricing

Input

Prompt / context tokens

$1.00/ 1M

Output

Generated tokens

$5.00/ 1M

Batch API

Async processing discount

$0.50 / $2.50

50% off standard price

Best For

extractionchatbotanalysis

Claude Haiku 4.5 in context

On a typical 70% input / 30% output mix, Claude Haiku 4.5 blends to $2.20 per million tokens. That puts it above 38% of the 32 models tracked here, and 10 other models in the same budget tier undercut it. Tier is a capability label, not a price band — the spread inside one tier is often wider than the gap between tiers, which is why picking by tier alone tends to overspend.

Input runs $1.00 per million and output $5.00 — a 5.0x ratio. That ratio matters more than either number on its own: a summarisation or classification workload reads far more than it writes and will track the input price, while a code-generation or long-form writing workload inverts that and will track the output price. Compare models on the ratio your own traffic actually has, not on the input price alone.

The benchmark profile is uneven rather than flat: data extraction at 91/100 against vision at 80/100, a 11-point spread. An average hides that. If your workload sits on the strong end this model punches above its price; if it sits on the weak end, a cheaper model with a flatter profile will serve you better.

The context window is 200K tokens and the throughput score is 97/100. Context only earns its keep if you fill it, and filling it is also what makes a request expensive — a large window is an option, not a discount. For a budget-tier model, judge the speed score against what you are doing: it decides everything for interactive chat and almost nothing for overnight batch work.

Worth checking before you commit: DeepSeek V3.2 scores at least as well on combined coding and reasoning and blends to $0.30 per million against Claude Haiku 4.5's $2.20. That does not make Claude Haiku 4.5 the wrong choice — context window, latency, vendor lock-in, and specific capabilities all sit outside a benchmark score — but it does mean the price difference needs a reason behind it.

The closest comparison is DeepSeek V3.2 from DeepSeek, at $0.30 per million against $2.20. 3 models in this tier score higher on combined coding and reasoning, which is the useful framing: the question is rarely whether a model is good, it is whether it is the cheapest thing that is good enough for the specific work you are sending it.

What it costs in production

At 10M + 2M tokens a month — a realistic mid-size production workload — Claude Haiku 4.5 runs about $20.00. Routing the latency-tolerant share through the batch API at 50% off would bring that to roughly $10.00, which is usually a bigger saving than switching models.

Price History

No price changes since launch

Strengths

  • Very reliable structured/JSON output for its price
  • 200K context — large enough for most document tasks
  • Fast enough for interactive chat

Avoid For

  • Frontier reasoning tasks
  • Workloads needing a context window above 200K