StackPricing
All AI models

Anthropic · AI model API pricing

How much does Claude Haiku 4.5 cost?

Anthropic's fast small model — matches GPT-5.6 Luna on input and beats it slightly on output.

Rate card

Input tokens$1 / 1M
Cached input$0.10 / 1M
Output tokens$5 / 1M
Context window200K

What that means per month

WorkloadTokens / monthCost / month
Prototype 1M in · 0.3M out $2.50/mo
Small app 10M in · 3M out $25/mo
Production app 50M in · 15M out $125/mo
High volume 200M in · 60M out $500/mo
Heavy platform 1B in · 300M out $2,500/mo
Price it at your exact usage Official pricing page

Standard on-demand rates from Anthropic's official pricing docs, last checked 2026-07-16. Monthly figures exclude caching and batch discounts. Confirm live pricing before committing.

Where Claude Haiku 4.5 sits on price

At the reference month of 10M input and 3M output tokens, Claude Haiku 4.5 costs $25, which makes it the 15th cheapest of the 30 models tracked here. That is 15× the price of Qwen-Flash, the cheapest model at this mix ($1.70). Among the 9 budget-tier models it ranks 9th, behind Qwen-Flash at $1.70.

Cache economics

Cached input on Claude Haiku 4.5 bills at $0.10 per 1M — 10% of its own input rate, against a 10% median across the models here that publish a cached-read price. 2 models discount cached reads more steeply than this. On the reference month, serving 80% of that input from cache takes the input side of the bill from $10 to $2.80, leaving output untouched at $15 — which is why a high reuse rate changes the ranking above rather than just shaving the total.

Cache reads only. Writes are billed separately by some providers and are excluded here — see the rate note above for Anthropic's terms. Try it at your own hit rate on the AI cost calculator.

Cheaper alternatives

At a reference workload of 10M input / 3M output tokens a month, these cost less than Claude Haiku 4.5 ($25/mo):

Claude Haiku 4.5 head-to-head