- Anthropic released Claude Haiku 5.5 on 7 October 2026.
- It is the fastest and cheapest small model in the Claude lineup.
- The model features adjustable reasoning effort settings for different tasks.
- Pricing starts at ten cents per million input tokens up to one hundred thousand tokens.
In this guide
What is Claude Haiku 5.5?
Claude Haiku 5.5 is a fast, low-cost small artificial intelligence model developed by Anthropic and released on 7 October 2026. This model is designed specifically to handle quick, repetitive, and cost-sensitive tasks at scale.
Compared to its predecessor, Claude Haiku 4.5, this updated version is significantly more capable and runs at a much lower cost. According to Anthropic, running the new model is about 75 percent cheaper on average.
- Developer
- Anthropic
- Release Date
- 7 October 2026
- Model Identifier
- claude-haiku-5-5
- Base Input Price
- $0.10 per million tokens
- Base Output Price
- $0.50 per million tokens
- Context Limit for Base Price
- 100,000 tokens
What can Claude Haiku 5.5 do?
Claude Haiku 5.5 excels at high-volume, speed-sensitive applications such as live customer support, database queries, summaries, and classification requests. It is also designed to act as an efficient subagent alongside larger models like Claude Sonnet 5.5 and Claude Opus 5.5 to perform specific tasks on coding projects.
This is the first Haiku-class model to feature an adjustable effort setting, letting users balance cost and intelligence. Users can choose between low, medium, high, xhigh, and max reasoning effort levels, though reasoning cannot be disabled entirely and defaults to medium.
In early evaluations shared by Anthropic, corporate users reported notable performance gains. HubSpot recorded a 92.8 percent average score on CRM deal reporting, while Box noted that the model completed analytical tasks with roughly half the latency of Claude Haiku 4.5.
- High-volume, repetitive workloads
- Speed-sensitive tasks like live chat support
- Subagent operations on coding tasks
- Budget-friendly processing under 100,000 tokens
- Offensive cybersecurity or penetration testing
- Workloads where reasoning must be entirely disabled
- Prompts exceeding 100,000 tokens due to price hikes
How much does Claude Haiku 5.5 cost?
Claude Haiku 5.5 matches the pricing of competing small models like OpenAI's GPT-6 Luna, costing $0.10 per million input tokens and $0.50 per million output tokens. However, this competitive rate only applies to prompts containing up to 100,000 tokens.
For tasks exceeding 100,000 tokens, the cost of Claude Haiku 5.5 increases fivefold to $0.50 per million input tokens and $2.50 per million output tokens. Additionally, the model utilizes a less generous tokenizer that requires roughly 1.25 times more tokens for the same long prompt compared to Claude Haiku 4.5.
| Token Type | Up to 100k Tokens (per M) | Over 100k Tokens (per M) |
|---|---|---|
| Input Tokens | $0.10 | $0.50 |
| Output Tokens | $0.50 | $2.50 |
| Cache Reads | $0.01 | $0.05 |
| Cache Writes | $0.125 | $0.625 |
How to use Claude Haiku 5.5
Claude Haiku 5.5 is available across multiple cloud platforms and directly through the Claude API. Developers can access the model using Amazon Web Services, Google Cloud, Microsoft Azure, or the Claude Platform.
Following these steps allows developers to run prompt queries using Claude Haiku 5.5 locally and manage monthly credits.
- Install or update the Anthropic command-line tool plugin: llm install -U llm-anthropic
- Refresh your active models to fetch the new release: llm anthropic refresh
- Run a test query with your preferred effort setting: llm -m claude-haiku-5.5 "Your prompt here" -o thinking_effort low
- To use monthly API credits as a Claude Max or Team subscriber, navigate to Settings, select Billing, and assign the credit to your chosen API organization.
Anthropic allows users to disable auto-reload on the API. This ensures requests stop when your monthly credit balance runs out, preventing unexpected billing charges.
What are the limits of Claude Haiku 5.5?
Claude Haiku 5.5 is subject to strict safety and alignment guardrails set by Anthropic. While the cybersecurity safeguards are more permissive than those on Claude Sonnet 5.5 to allow defensive operations, the model still blocks penetration testing and other offensive techniques.
Biology safeguards restrict requests that could cause physical harm while still permitting basic research biology questions. Furthermore, developers working with large volumes should monitor context limits closely, as prompts longer than 100,000 tokens trigger the 5x pricing multiplier.
Frequently asked questions
What is Claude Haiku 5.5?
How much does Claude Haiku 5.5 cost?
Can I adjust the reasoning levels in Claude Haiku 5.5?
How do I access Claude Haiku 5.5?
What are the safety restrictions on Claude Haiku 5.5?
Sources: Anthropic · Simon Willison
Want your brand to be the answer AI gives?
See how ready your website is for ChatGPT, Gemini and Perplexity, free, in about ten seconds.