$1.00 input · $5.00 output per 1M tokens — Anthropic's budget tier for high-volume, low-latency work.
Claude Haiku 4.5 is Anthropic's budget tier: $1.00 input and $5.00 output per million tokens, with cached input at $0.10. It's the right call for high-volume, simpler tasks — classification, extraction, summarization, and straightforward tool calls.
Claude Haiku 4.5 · USD official list price
| Tier | Rate |
|---|---|
| Input | $1.00 |
| Output | $5.00 |
| Cache hit | $0.10 |
DeepSeek lists official prices in CNY; USD uses ≈ 6.9 exchange rate
Budget-tier list price for input and output per million tokens.
Cached input at one-tenth of the uncached rate.
Designed for high-volume, latency-sensitive workloads like classification and extraction.
High-volume classification, extraction, summarization, and straightforward tool calls where the extra reasoning headroom of Sonnet 5 isn't needed.
Published API pricing as of September 2026, standard tier, USD per million tokens. Also qualifies for Anthropic's 50% Batch API discount on offline workloads.
At $1.00/$5.00, Haiku 4.5 is more expensive than GPT-5.6 Luna ($0.20/$1.20) and DeepSeek V4 Flash off-peak (¥1.00/¥4.00) — Anthropic's budget tier still carries a premium over the cheapest models on the market.
Claude Haiku 4.5 bills $1.00 per million input tokens and $5.00 per million output tokens, with cached input at $0.10 per million — Anthropic's budget tier for high-volume, latency-sensitive work.
Haiku 4.5 is the right call for high-volume, simpler tasks — classification, extraction, summarization, and straightforward tool calls — where the extra reasoning headroom of Sonnet 5 isn't needed. It also benefits from Anthropic's 50% Batch discount.
Yes. Cached input is $0.10 per million tokens, one-tenth of the uncached input rate, making stable-prefix workloads especially cheap at this tier.
Compare this model against the rest of the lineup.
Model your own workload — input/output tokens, cache hit rate, and peak-hour share — against any model mix with the interactive pricing calculator.
Open the Pricing CalculatorWe are using these tools ourselves for Development / Deployment. Check out for more details.
You will receive $5 when subscribed, which directly offsets the 1st month's Go fee.
Could be the cheapest option for calling DeepSeek V4 Flash model in the world.
RReceive $300 to test out VVultr platform
Referred user must be active 30+ days and spend $10–$25
Deploy anything without the complexity
Connect your repo, Railway handles the rest.
Cloud servers, cloud databases, COS, CDN, SMS and other cloud products are on special offer now.
AWS high-quality alternative, we use COS as a replacement for S3
¥16 universal voucher for all platforms, valid for 180 days from the date of receipt
A comprehensive product matrix supporting the full-process implementation of AI applications.
One of the best AI terminal tools in the world
Interestingly, we can directly use various top-tier models in Warp without leaving the terminal.