The Tool Desk
Outbyte PC Repair FREERepair Windows errors before they cause bigger problemsFix Now →Outbyte Driver Updater FREEScan for outdated or missing drivers - takes under a minuteDriver Scan →Some links on this page are affiliate links: if you buy through them we may earn a commission, at no extra cost to you.
Anthropic’s Claude 3.5 Haiku became much more expensive than the Claude 3 Haiku model it replaced. The initial November 2024 announcement raised input and output prices fourfold—from $0.25 to $1 and from $1.25 to $5 per million tokens. Anthropic later revised those prices to $0.80 and $4.
The change mattered because Haiku had been positioned as Anthropic’s fast, lower-cost model. It also showed that “small” does not necessarily mean “cheap” when a model becomes more capable.
What changed in Claude Haiku pricing?
Claude 3.5 Haiku was announced on October 22, 2024, as a faster and more capable small model. Pricing initially appeared to track the older Claude 3 Haiku:
| Date | Model or event | Input | Output |
|---|---|---|---|
| March 13, 2024 | Claude 3 Haiku launched | $0.25 per million tokens | $1.25 per million tokens |
| October 22, 2024 | Claude 3.5 Haiku announced | $0.25 expected | $1.25 expected |
| November 4, 2024 | Initial Claude 3.5 Haiku pricing | $1 | $5 |
| December 3, 2024 | Revised Claude 3.5 Haiku pricing | $0.80 | $4 |
The original November announcement therefore represented a fourfold price increase over Claude 3 Haiku. In percentage terms, both input and output prices were 300% higher.
#1 Best Overall
Anthropic subsequently reduced the new model’s standard price to $0.80 per million input tokens and $4 per million output tokens. Compared with Claude 3 Haiku, that final historical price was 3.2 times higher—a 220% increase—not four times higher.
Why did Anthropic raise the price?
Anthropic said final testing showed that Claude 3.5 Haiku was more intelligent than expected. The company pointed to benchmark performance, including results it said surpassed Claude 3 Opus on several evaluations. That was Anthropic’s stated justification for charging more.
Anthropic reported a 40.6% score on SWE-bench Verified and described the result as outperforming several publicly available models, including the original Claude 3.5 Sonnet and GPT-4o. This was a vendor-reported benchmark result, so it should not be treated as proof that the model was better on every developer’s workload. Results can depend on the prompts, test harness, benchmark version and evaluation date.
Quick wins for a faster PC:
Scan for outdated or missing drivers - takes under a minuteDriver Scan →Repair Windows errors before they cause bigger problemsFix Now →Fix the driver behind crashes, sound loss and screen glitchesFind Drivers →The broader commercial message was clear: Anthropic was pricing Haiku according to capability, not simply model size or speed. A compact model that performs more difficult coding, tool-use and instruction-following tasks can occupy a higher price tier even if it remains intended for low-latency applications.
Rank #2
What did customers get for the higher price?
Anthropic positioned Claude 3.5 Haiku for:
- coding and code-related subtasks;
- more reliable instruction following;
- tool use and specialized sub-agents;
- interactive, low-latency applications; and
- high-volume processing of structured or semi-structured data.
A stronger small model could potentially reduce retries, failed tool calls, escalations to a more expensive model or human review. But those are workload-specific possibilities, not a guarantee that the higher token price produced a lower total cost.
The important limitation: it initially lacked image input
Claude 3.5 Haiku launched as a text-only model, with image input planned for a later update. Claude 3 Haiku supported vision. That meant Claude 3.5 Haiku was not a universal replacement despite its higher intelligence claims.
For developers processing images, the cheaper older model could remain the more suitable option. The comparison was therefore not simply “newer and better versus older and worse”; the models differed in modality, speed, knowledge cutoff and task performance.
Do these 3 things before closing this tab:
1Clear out junk files and repair common Windows errors2Fix the driver behind crashes, sound loss and screen glitches3Repair Windows errors before they cause bigger problemsHow much did the change affect usage costs?
Token prices apply separately to input and output. That distinction matters because applications that generate long answers can incur substantially higher output costs.
Example: 1 million input and 1 million output tokens
| Model or price | Input | Output | Total |
|---|---|---|---|
| Claude 3 Haiku | $0.25 | $1.25 | $1.50 |
| Claude 3.5 Haiku at initial price | $1 | $5 | $6 |
| Claude 3.5 Haiku after revision | $0.80 | $4 | $4.80 |
Example: 100 million input and 20 million output tokens
| Model or price | Input cost | Output cost | Total |
|---|---|---|---|
| Claude 3 Haiku | $25 | $25 | $50 |
| Claude 3.5 Haiku at initial price | $100 | $100 | $200 |
| Claude 3.5 Haiku after revision | $80 | $80 | $160 |
These are illustrative token charges, excluding taxes, cloud-provider billing differences and other application costs. Prompt length, retries, caching, batch eligibility and the ratio of input to output tokens can materially change the final bill.
Could customers keep using Claude 3 Haiku?
Yes, at the time of the price change. Anthropic representatives said Claude 3 Haiku would remain available for customers prioritizing maximum cost efficiency or image processing. That made the announcement a model-segmentation decision rather than an immediate forced migration.
However, legacy availability changes over time. Developers should check the current deprecation and retirement documentation before building a new production system around an older model.
What’s actually slowing this PC down?
Pick the symptom - the matching free tool is one click away.
Which platforms were affected?
Claude 3.5 Haiku was made available through Anthropic’s API, Amazon Bedrock and Google Cloud Vertex AI. Customers using those services could therefore encounter the model through different procurement and billing arrangements.
The headline Anthropic price should not be assumed to be every customer’s effective price. Cloud platforms may differ by region, account terms, marketplace arrangements, infrastructure options and billing rules. Check the applicable Bedrock pricing or Vertex AI pricing before estimating production costs.
When could the more expensive model make financial sense?
The higher-priced model could be reasonable when its additional capability reduces other costs. Examples include:
- coding workflows where fewer failed attempts save engineering time;
- agent systems where a stronger small model avoids escalation to Sonnet or Opus;
- tool-calling workflows where errors create expensive retries; and
- interactive products where latency and response quality matter more than the lowest token price.
Teams should measure success using the cost per completed task, not just the cost per million tokens. A more expensive model is not automatically cheaper overall, and a cheaper model may be the better choice when simple classification, extraction, moderation or labeling is accurate enough.
Ways to control costs
For asynchronous work such as bulk extraction or classification, Anthropic’s Message Batches API offered a 50% discount on standard input and output token pricing. Batch processing is not suitable for user-facing requests that require immediate results.
Best Value
Other cost variables include prompt caching, prompt length, retry rates, output limits, regional availability and whether image input is required. Current Anthropic pricing documentation separates standard token charges from cache and batch pricing, so these mechanisms should not be conflated with the model’s headline rate.
What is the current status?
This was a November 2024 pricing event, not a new 2026 price hike. As of August 18, 2026, Anthropic’s main pricing documentation lists Claude Haiku 4.5 as the current mainline Haiku model at $0.50 per million input tokens and $2.50 per million output tokens. Claude 3.5 Haiku is listed as retired except on AWS Bedrock and Google Cloud.
That current status is separate from the historical Claude 3.5 Haiku pricing story. Anyone choosing a model today should verify current model IDs, retirement dates, availability and platform-specific prices rather than relying on the 2024 figures.
What the decision signaled about AI pricing
Anthropic’s move challenged the assumption that a “small” or “fast” model must remain the budget option. The company treated improved capability as a reason to move Haiku into a higher pricing band, even though the model still targeted latency-sensitive and high-volume applications.
For buyers, the practical lesson is to compare quality-adjusted cost: how much a completed task costs after accounting for retries, escalation, latency, human review and required modalities. The model with the lowest token price is not always the least expensive choice, but neither does a benchmark advantage prove that a premium model will pay for itself.
Quick Recap
Product prices and availability are accurate as of the date/time indicated and are subject to change. Any price and availability information displayed on Amazon at the time of purchase will apply.

