Anthropic launched Claude Haiku 5.5 on October 7, 2026, introducing a small AI model designed for fast responses, high-volume workloads and lower operating costs.
The company targets applications such as summarization, text classification, database queries, request routing and live customer support. The model also works as a subagent, handling smaller tasks within a larger AI workflow.

Anthropic describes Haiku 5.5 as its fastest model to date at each model’s standard speed. The company says the model costs around 75% less to run than Claude Haiku 4.5 on average.
The published API rates are $0.10 per million input tokens and $0.50 per million output tokens for prompts up to 100,000 tokens. For prompts above 100,000 tokens, the rates rise to $0.50 per million input tokens and $2.50 per million output tokens.
Anthropic says Haiku 5.5 has 90% lower token rates than Haiku 4.5 for prompts up to 100,000 tokens and 50% lower rates for longer prompts. The company notes that changes to tokenization affect the final savings, so the reduction in token prices does not translate directly into the same reduction in every workload’s total cost.
The release gives developers another option for applications that need to process large numbers of requests. Instead of sending every task to a larger model, developers might use Haiku 5.5 for routine jobs and reserve more capable models for complex work.
Sources: Anthropic’s official Claude Haiku 5.5 announcement and Claude Haiku model page.
Key Facts About Claude Haiku 5.5
| Detail | Verified information |
|---|---|
| Release date | October 7, 2026 |
| Developer | Anthropic |
| Model name | Claude Haiku 5.5 |
| Main focus | Fast, high-volume and cost-sensitive tasks |
| Common uses | Summarization, classification, routing, database queries and live support |
| Input price for prompts up to 100K tokens | $0.10 per million tokens |
| Output price for prompts up to 100K tokens | $0.50 per million tokens |
| Input price for prompts over 100K tokens | $0.50 per million tokens |
| Output price for prompts over 100K tokens | $2.50 per million tokens |
| Estimated average cost reduction | Around 75% compared with Haiku 4.5 |
| API model identifier | claude-haiku-5-5 |
Source: Anthropic’s official announcement.
What Is Claude Haiku 5.5?
Claude Haiku 5.5 is a small AI model in Anthropic’s Claude family. It focuses on tasks where response speed, operating costs and the ability to process many requests matter.
A business might use the model to summarize customer conversations, classify incoming emails or extract details from documents. A developer might use it to route requests to the correct system or prepare information for a larger model.
These tasks often involve clear instructions and repeated operations. A fast, lower-cost model helps developers process them without sending every request to a more expensive model.
Haiku 5.5 also supports AI agent workflows. An agent is a system that performs tasks through a series of steps, sometimes using tools or other models. A subagent handles one part of the larger workflow.
For example, a larger model might prepare a business report while Haiku 5.5 extracts figures from a document or summarizes a section. The larger model then uses the result to complete the main task.
Anthropic says Haiku 5.5 works well for this kind of supporting role, particularly when the application needs fast responses across many requests.
How Much Does Claude Haiku 5.5 Cost?
Pricing is a major part of the release.
Anthropic publishes different rates depending on whether a prompt contains up to 100,000 tokens or exceeds that threshold.
| API usage | Prompts up to 100K tokens | Prompts over 100K tokens |
|---|---|---|
| Input tokens | $0.10 per million | $0.50 per million |
| Output tokens | $0.50 per million | $2.50 per million |
| Cache reads | $0.01 per million | $0.05 per million |
Source: Anthropic’s pricing table.
Input tokens represent the content sent to the model. Output tokens represent the content generated in response. Cache reads apply when an application uses eligible cached input.
These rates help developers estimate the cost of running the model through the Claude Platform. The final bill depends on token usage and the applicable pricing rules.
How Much Cheaper Is Haiku 5.5 Than Haiku 4.5?
Anthropic estimates that Haiku 5.5 costs around 75% less to run on average than Haiku 4.5.
For prompts up to 100,000 tokens, Haiku 5.5’s input rate is $0.10 per million tokens, compared with $1.00 for Haiku 4.5. The output rate falls from $5.00 to $0.50 per million tokens.
For prompts above 100,000 tokens, Haiku 5.5 costs $0.50 per million input tokens and $2.50 per million output tokens. Anthropic lists the corresponding Haiku 4.5 rates as $1.00 and $5.00.
The company says around 90% of requests to its previous Haiku model fell into the shorter-prompt category. Its average cost estimate also accounts for differences in token usage between the models.
Haiku 5.5 uses an updated tokenizer, which means the same text might use more tokens than it did with Haiku 4.5. Developers should therefore test their own workloads instead of assuming every task will cost exactly 75% less.
Main Uses for Claude Haiku 5.5
1. Summarization
Businesses often process reports, customer conversations, meeting notes and long documents. Summarization turns these materials into shorter versions containing the main points.
Haiku 5.5 targets this type of repetitive work. A support team might use the model to prepare a summary before an employee reviews a customer case. A business might also use it to produce brief notes from internal documents.
Users should check the results for missing details and factual errors, particularly when summaries inform important decisions.
2. Text Classification
Classification involves placing text into predefined categories.
A customer service system might sort messages into billing, delivery, technical support and account access. A business might classify documents by topic or urgency.
Haiku 5.5 is designed for these high-volume tasks. Its pricing makes it an option for applications that process many short requests.
Developers should test classification accuracy against real examples. Unclear messages and unusual cases still need careful handling.
3. Live Customer Support
Customer support systems need to identify requests and respond quickly.
Haiku 5.5 targets speed-sensitive applications such as live support and browser use. A business might use the model to identify a customer’s issue, retrieve relevant information and prepare a response.
For complex complaints or sensitive account issues, businesses should set rules for human review and escalation.
4. Database Queries and Request Routing
Anthropic lists database queries and routing among the model’s intended workloads.
Routing directs an incoming request to the correct system or process. For example, an application might identify whether a customer needs billing assistance, technical support or information about a product.
A model designed for fast, repeated tasks provides an option for automating these steps. Developers should test the results and define what happens when the model is unsure.
5. AI Agent Workflows
Haiku 5.5 can handle smaller tasks inside a larger AI workflow.
A larger model might plan a complex task, while Haiku 5.5 performs a short lookup, summarizes information or extracts a specific detail.
Anthropic also identifies coding subagent work as a suitable use case. The model is not positioned as the best choice for every complex coding task. Its role is to handle focused work quickly and at lower cost.
How Haiku 5.5 Compares With Larger Claude Models
Anthropic’s model family includes Haiku, Sonnet and Opus models, which serve different workloads.
Haiku 5.5 targets fast, high-volume and cost-sensitive work. Sonnet models offer a different balance of capability, speed and cost. Opus models target more demanding reasoning and coding tasks.
This allows developers to choose models based on the task instead of sending every request to the same system.
For example, an application might use Haiku 5.5 to classify incoming requests, then send complicated cases to a larger model. This approach is useful only if testing shows that the combined workflow meets the application’s accuracy and cost requirements.
Anthropic’s own announcement says Sonnet 5.5 and Opus 5.5 remain better choices for complex agentic coding tasks such as those measured by Terminal-Bench 4.0.
Read Anthropic’s model announcement and benchmark information.
Where Is Claude Haiku 5.5 Available?
Anthropic says Haiku 5.5 is available through Claude.ai, the Claude Platform and supported cloud services, including Amazon Web Services, Google Cloud and Microsoft Azure.
Developers using the Claude Platform should select the model identifier claude-haiku-5-5.
Access, billing and available features depend on the service and account. Developers should check the relevant platform documentation before adding the model to an application.
The official Claude Haiku 5.5 page contains the announcement, pricing and availability information.
What Developers Should Consider Before Switching
Lower prices do not automatically make a model suitable for every task.
Developers should test Haiku 5.5 with representative prompts and compare its accuracy, response times and total operating costs with their current model.
Token usage matters as well. A model might have lower rates per million tokens but use a different number of tokens to complete the same task.
Caching and batch processing also affect costs. Anthropic states that eligible prompt caching offers savings of up to 90% on cached tokens, while batch processing offers a 50% discount on input and output rates under its published terms.
Teams should also review data handling, access controls and human review requirements before deploying the model in customer-facing systems.
Conclusion
Anthropic released Claude Haiku 5.5 on October 7, 2026, with a focus on speed, efficiency and lower costs for high-volume AI workloads.
The model targets summarization, classification, database queries, request routing and live customer support. Anthropic estimates around 75% lower average operating costs compared with Haiku 4.5, although the final savings depend on token usage and the task.
Its pricing and supporting role in AI agent systems give developers another option for routine work. More demanding reasoning and coding tasks might still require a larger model.
Developers should check Anthropic’s official pricing and model documentation, test their own workloads and calculate actual costs before changing production systems.
Frequently Asked Questions
When did Anthropic release Claude Haiku 5.5?
Anthropic announced Claude Haiku 5.5 on October 7, 2026.
What is Claude Haiku 5.5 designed for?
The model targets summarization, classification, database queries, request routing, live customer support and subagent workflows.
