Magic Tools
LLM API Pricing Calculator (all models)

Claude Haiku 4.5 API Pricing and Cost Estimates

Anthropic Claude Haiku 4.5 API pricing is $1 per 1M input tokens and $5 per 1M output tokens. Cached input is billed at $0.1 per 1M tokens. The context window is 200K tokens. The Batch API bills at 50% of standard. Prices verified 2026-09-04 against the official pricing page.

Input / 1M tokens

$1

Output / 1M tokens

$5

Cached input / 1M

$0.1

Context window

200K

ActiveNo deprecation announced (verified 2026-09-05). Tentative retirement not sooner than Oct 15, 2026.
Full deprecation timeline →

Monthly cost at three workloads

WorkloadTokens / monthStandard50% cache hitBatch API
Light (prototype / side project)5M in + 1M out$10.00$7.75$5.00
Medium (small production app)50M in + 10M out$100$77.50$50.00
Heavy (production at scale)500M in + 100M out$1000$775$500

Closest-priced alternatives

ModelInput / 1MOutput / 1MMedium workload / mo
Anthropic Claude Haiku 4.5$1$5$100
DeepSeek DeepSeek V4 Pro$1.3$3.91$104
xAI Grok 4.20$1.25$2.5$87.50
Moonshot Kimi K2.7 Code$0.95$4$87.50
Zhipu (z.ai) GLM-5.2$1.4$4.4$114

Medium workload = 50M input + 10M output tokens per month at standard price, no caching. Prices verified 2026-09-04.

Claude Haiku 4.5 vs the rest of the Anthropic lineup

ModelInput / 1MOutput / 1MCached / 1MContextMedium workload / mo
Claude Haiku 4.5$1$5$0.1200K$100
Claude Fable 5.1$10$50$0.251M$1000
Claude Fable 5$10$50$11M$1000
Claude Opus 5$5$25$0.51M$500
Claude Sonnet 5$2$10$0.21M$200

Calculate with your own volume

Back

LLM API Pricing Calculator

Compare current API prices for Claude, GPT, Gemini, DeepSeek, Kimi, Grok and other LLMs, and estimate your monthly bill from token volume. Includes prompt-caching prices and batch discounts that most comparison tables miss. Pick your current model to see exactly how much migrating to each alternative saves. Prices last verified 2026-09-04. Runs entirely in your browser — no upload, no signup.

M tokens
M tokens

Share of input tokens served from prompt cache

Adds a column showing what switching to each model saves (or costs) per month at this usage.

ModelInput $/MOutput $/MCached $/MContextEst. monthlyvs current

GPT-5.6 luna

OpenAI

$0.2$1.2$0.021.05M$22.00−$78.00 (−78%)

Gemini 3.1 Flash-Lite

Google

$0.25$1.5$0.0251M$27.50−$72.50 (−73%)

DeepSeek V4 Flash

DeepSeek · open-weight

$0.435$1.3$0.01451M$34.75−$65.25 (−65%)

Gemini 3.6 Flash

Google

$0.75$3.75$0.0751M$75.00−$25.00 (−25%)

GLM-5

Zhipu (z.ai) · open-weight

$1$3.2$0.2200K$82.00−$18.00 (−18%)

Grok 4.20

xAI

$1.25$2.5$0.21M$87.50−$12.50 (−13%)

Kimi K2.7 Code

Moonshot · open-weight

$0.95$4$0.19262K$87.50−$12.50 (−13%)

Claude Haiku 4.5current

Anthropic

$1$5$0.1200K$100

DeepSeek V4 Pro

DeepSeek · open-weight

$1.3$3.91$0.0431M$104+$4.10 (+4%)

GLM-5.2

Zhipu (z.ai) · open-weight

$1.4$4.4$0.26200K$114+$14.00 (+14%)

Grok 4.6

xAI

$2$6$0.5500K$160+$60.00 (+60%)

Grok 4.5

xAI

$2$6$0.3500K$160+$60.00 (+60%)

Claude Sonnet 5

Anthropic

$2$10$0.21M$200+$100 (+100%)

GPT-5.6 terra

OpenAI

$2$12$0.21.05M$220+$120 (+120%)

Gemini 3.1 Pro Preview

Google

$2$12$0.21M$220+$120 (+120%)

Kimi K3

Moonshot · open-weight

$3$15$0.31.048576M$300+$200 (+200%)

GPT-5.6 sol

OpenAI

$4$20$0.41.05M$400+$300 (+300%)

Claude Opus 5

Anthropic

$5$25$0.51M$500+$400 (+400%)

Claude Fable 5.1

Anthropic

$10$50$0.251M$1000+$900 (+900%)

Claude Fable 5

Anthropic

$10$50$11M$1000+$900 (+900%)

Prices are per 1 million tokens in USD, standard (non-batch) tier, verified 2026-09-04. Claude Fable 5.1: Released Sep 1, 2026. Anthropic's Mythos-class flagship tier above Opus 5 — same $10/$50 base rate as Fable 5, but cache reads drop to 2.5% of the input price ($0.25/MTok vs the 10% every other Claude model charges). Shares its underlying model with Claude Mythos 5.1, offered only to approved organizations. Claude Fable 5: Superseded by Claude Fable 5.1 (Sep 1, 2026), which keeps the same base rate but cuts cache reads to $0.25/MTok. Fable 5 remains available as a pinned snapshot with cache reads at the standard 10% ($1.00/MTok). Shares its underlying model with Claude Mythos 5. Claude Sonnet 5: The $2 / $10 launch price is now permanent; the planned Sep 2026 increase to $3 / $15 was cancelled. GPT-5.6 sol: Promotional price, available at least through Nov 21, 2026 (previously $5 / $30). Requests beyond the long-context threshold bill at 2× input / 1.5× output. GPT-5.6 terra: Requests beyond the long-context threshold bill at 2× input / 1.5× output. GPT-5.6 luna: Requests beyond the long-context threshold bill at 2× input / 1.5× output. Gemini 3.1 Pro Preview: Prompts over 200K tokens bill at $4 / $18. Gemini 3.6 Flash: Intro price through Dec 31, 2026; $1.50 / $7.50 from Jan 1, 2027. Grok 4.6: Requests over 200K tokens bill at 2×. Grok 4.5: Requests over 200K tokens bill at 2×. Kimi K3: Always-on reasoning; output includes thinking tokens. DeepSeek V4 Pro: Peak-hour price (Beijing weekdays 09:00–12:00, 14:00–18:00) effective Aug 17, 2026; all other hours bill at 50% of peak. DeepSeek V4 Flash: Peak-hour price (Beijing weekdays 09:00–12:00, 14:00–18:00) effective Aug 17, 2026; all other hours bill at 50% of peak.

Batch discount is applied only to models whose provider offers an async batch tier. Cached-input pricing models the read price; cache-write surcharges (Anthropic) are not included. Long-context surcharges apply above provider thresholds and are noted per model.

FAQ

How much does the Claude Haiku 4.5 API cost?

$1 per 1M input tokens and $5 per 1M output tokens, with cached input at $0.1 per 1M. (Verified 2026-09-04.)

What does a month of Claude Haiku 4.5 cost in practice?

A medium workload of 50M input + 10M output tokens per month runs about $100/mo, or about $77.50/mo if half the input hits the prompt cache. Offline jobs can use the Batch API at 50% of standard.

How does prompt caching change Claude Haiku 4.5 pricing?

Input tokens served from the prompt cache bill at $0.1 per 1M — about 10% of the standard input price. Repeated prefixes like system prompts and few-shot examples benefit most.

Is Claude Haiku 4.5 cheaper than DeepSeek V4 Pro?

At the same medium workload, Claude Haiku 4.5 costs about $100/mo versus about $104/mo for DeepSeek DeepSeek V4 Pro. Also weigh quality, latency, and context window (200K vs 1M tokens).

How does Claude Haiku 4.5 pricing compare with other Anthropic models?

By input price: Claude Fable 5.1 is $10/$50 per 1M tokens in/out (about 1000% of Claude Haiku 4.5's input price); Claude Fable 5 is $10/$50 per 1M tokens in/out (about 1000% of Claude Haiku 4.5's input price); Claude Opus 5 is $5/$25 per 1M tokens in/out (about 500% of Claude Haiku 4.5's input price); Claude Sonnet 5 is $2/$10 per 1M tokens in/out (about 200% of Claude Haiku 4.5's input price). See the full-lineup table above.

API pricing for other models