AI API PricesAnthropic › Claude Haiku 4.5

Claude Haiku 4.5 pricing

Claude Haiku 4.5 API cost from Anthropic — $1.00 per 1M input tokens and $5.00 per 1M output tokens. Fastest, near-frontier. Estimate your real monthly bill below, then compare it against every other model.

Input price
$1.00
per 1M input tokens
Output price
$5.00
per 1M output tokens
Cached input
$0.100
per 1M · ~90% off input
Reference call
$0.0060
1K in + 1K out

Fast / cheap 200K context · all prices USD per 1M tokens, list price last verified 2026-06-27.

Claude Haiku 4.5 capabilities & limits

Representative specs for this tier — confirm exact limits in the Anthropic docs.

Context window200K
Max output16K
Knowledge cutoffSep 2025
API formatanthropic
✓ Vision✓ Tools / functions✓ JSON / structured✓ Prompt caching✓ Batch API

API acting up? Check live Anthropic status → · all AI API status

Call Claude Haiku 4.5 from your code

The exact model string and endpoint to use. Uses the Anthropic API format.

Model stringclaude-haiku-4-5
Base URLhttps://api.anthropic.com/v1
API key envANTHROPIC_API_KEY
curl https://api.anthropic.com/v1/messages \
  -H "x-api-key: $ANTHROPIC_API_KEY" \
  -H "anthropic-version: 2023-06-01" \
  -H "content-type: application/json" \
  -d '{
    "model": "claude-haiku-4-5",
    "max_tokens": 1024,
    "messages": [{"role": "user", "content": "Hello"}]
  }'

Model strings follow each provider's naming convention; confirm the current id in the Anthropic docs before relying on it.

What a Claude Haiku 4.5 API call actually costs

Blended input + output cost at Claude Haiku 4.5's list rates — useful yardsticks before you commit to volume.

$0.0060
Short call
1K in · 1K out
$0.0055
Chat turn
3K in · 500 out
$0.025
RAG / long prompt
20K in · 1K out
$6.00
1M in + 1M out
bulk job

Claude Haiku 4.5 cost calculator

Enter your monthly volume to estimate the Claude Haiku 4.5 bill. Runs entirely in your browser.

Estimated monthly cost

For the full multi-model breakdown, use the cost calculator.

Anthropic models — Claude Haiku 4.5 in context

How Claude Haiku 4.5 prices against its siblings. Cheapest reference cost first.

Model Input /1M Output /1M Cost / call* Context Type
Claude Haiku 4.5
Fastest, near-frontier
$1.00 $5.00 $0.0060 200K Fast / cheap
Claude Sonnet 4.6
Best speed/intelligence; caching cuts input ~90%
$3.00 $15.00 $0.018 1M Balanced
Claude Opus 4.8
Top Opus reasoning/agentic
$5.00 $25.00 $0.030 1M Flagship
Claude Fable 5
Most capable widely released
$10.00 $50.00 $0.060 1M Flagship

*Reference cost of one call with 1,000 input + 1,000 output tokens — a neutral yardstick. Use the calculator for your real usage. Cheapest row highlighted.

Cheaper than Claude Haiku 4.5? In the same fast / cheap tier, Gemini 1.5 Flash-8B runs $0.037 in / $0.150 out. The overall cheapest model tracked is Gemini 1.5 Flash-8B — see the cheapest LLM API guide.

How Claude Haiku 4.5 pricing works

Claude Haiku 4.5 is billed per token: a lower input price of $1.00 per million tokens for everything you send, and a higher output price of $5.00 per million for everything the model generates. To turn that into a real Claude Haiku 4.5 bill, multiply by your monthly request volume and average prompt/response length — the calculator above does exactly that.

Prompt caching is the biggest lever on your Claude Haiku 4.5 cost: cached input tokens bill at just $0.100 per million — about 90% cheaper. If you reuse a long system prompt or context across calls, caching can dramatically cut the input portion of your bill.

List prices move fast — always confirm the current numbers on the official Anthropic pricing page before relying on them for budgeting.

Claude Haiku 4.5 pricing FAQ

How much does the Claude Haiku 4.5 API cost?

Claude Haiku 4.5 costs $1.00 per 1M input tokens and $5.00 per 1M output tokens, with cached input at $0.100 per 1M (about 90% off). A typical 3K-input / 500-output call works out to about $0.0055.

What is Claude Haiku 4.5's context window?

Claude Haiku 4.5 has a 200K-token context window. Tokens you put in the prompt are billed at the $1.00 per 1M input rate.

Does Claude Haiku 4.5 support prompt caching?

Yes. Claude Haiku 4.5 bills cached input tokens at $0.100 per 1M instead of $1.00 — roughly 90% cheaper for the repeated part of your prompts.

Is there a cheaper alternative to Claude Haiku 4.5?

In the same fast / cheap tier, Gemini 1.5 Flash-8B (Google) is cheaper at $0.037 in / $0.150 out. The overall cheapest model on the market is Gemini 1.5 Flash-8B.

Compare Claude Haiku 4.5 with other models