New AI release ·

Claude Haiku 5.5: Features, Pricing and How to Use It

Claude Haiku 5.5 is a fast, cost-effective small AI model from Anthropic built for high-volume workloads. It features adjustable reasoning levels and is available across major cloud platforms and the Claude API.

In short
  • Anthropic released Claude Haiku 5.5 on 7 October 2026.
  • It is the fastest and cheapest small model in the Claude lineup.
  • The model features adjustable reasoning effort settings for different tasks.
  • Pricing starts at ten cents per million input tokens up to one hundred thousand tokens.
In this guide
  1. What is Claude Haiku 5.5?
  2. What can Claude Haiku 5.5 do?
  3. How much does Claude Haiku 5.5 cost?
  4. How to use Claude Haiku 5.5
  5. What are the limits of Claude Haiku 5.5?
  6. Frequently asked questions

What is Claude Haiku 5.5?

Claude Haiku 5.5 is a fast, low-cost small artificial intelligence model developed by Anthropic and released on 7 October 2026. This model is designed specifically to handle quick, repetitive, and cost-sensitive tasks at scale.

Compared to its predecessor, Claude Haiku 4.5, this updated version is significantly more capable and runs at a much lower cost. According to Anthropic, running the new model is about 75 percent cheaper on average.

Claude Haiku 5.5 At A Glance
Developer
Anthropic
Release Date
7 October 2026
Model Identifier
claude-haiku-5-5
Base Input Price
$0.10 per million tokens
Base Output Price
$0.50 per million tokens
Context Limit for Base Price
100,000 tokens

What can Claude Haiku 5.5 do?

Claude Haiku 5.5 excels at high-volume, speed-sensitive applications such as live customer support, database queries, summaries, and classification requests. It is also designed to act as an efficient subagent alongside larger models like Claude Sonnet 5.5 and Claude Opus 5.5 to perform specific tasks on coding projects.

This is the first Haiku-class model to feature an adjustable effort setting, letting users balance cost and intelligence. Users can choose between low, medium, high, xhigh, and max reasoning effort levels, though reasoning cannot be disabled entirely and defaults to medium.

In early evaluations shared by Anthropic, corporate users reported notable performance gains. HubSpot recorded a 92.8 percent average score on CRM deal reporting, while Box noted that the model completed analytical tasks with roughly half the latency of Claude Haiku 4.5.

Good for
  • High-volume, repetitive workloads
  • Speed-sensitive tasks like live chat support
  • Subagent operations on coding tasks
  • Budget-friendly processing under 100,000 tokens
Not ideal for
  • Offensive cybersecurity or penetration testing
  • Workloads where reasoning must be entirely disabled
  • Prompts exceeding 100,000 tokens due to price hikes

How much does Claude Haiku 5.5 cost?

Claude Haiku 5.5 matches the pricing of competing small models like OpenAI's GPT-6 Luna, costing $0.10 per million input tokens and $0.50 per million output tokens. However, this competitive rate only applies to prompts containing up to 100,000 tokens.

For tasks exceeding 100,000 tokens, the cost of Claude Haiku 5.5 increases fivefold to $0.50 per million input tokens and $2.50 per million output tokens. Additionally, the model utilizes a less generous tokenizer that requires roughly 1.25 times more tokens for the same long prompt compared to Claude Haiku 4.5.

Claude Haiku 5.5 Pricing Structure
Token TypeUp to 100k Tokens (per M)Over 100k Tokens (per M)
Input Tokens$0.10$0.50
Output Tokens$0.50$2.50
Cache Reads$0.01$0.05
Cache Writes$0.125$0.625

How to use Claude Haiku 5.5

Claude Haiku 5.5 is available across multiple cloud platforms and directly through the Claude API. Developers can access the model using Amazon Web Services, Google Cloud, Microsoft Azure, or the Claude Platform.

What you'll get

Following these steps allows developers to run prompt queries using Claude Haiku 5.5 locally and manage monthly credits.

  1. Install or update the Anthropic command-line tool plugin: llm install -U llm-anthropic
  2. Refresh your active models to fetch the new release: llm anthropic refresh
  3. Run a test query with your preferred effort setting: llm -m claude-haiku-5.5 "Your prompt here" -o thinking_effort low
  4. To use monthly API credits as a Claude Max or Team subscriber, navigate to Settings, select Billing, and assign the credit to your chosen API organization.
Tip

Anthropic allows users to disable auto-reload on the API. This ensures requests stop when your monthly credit balance runs out, preventing unexpected billing charges.

What are the limits of Claude Haiku 5.5?

Claude Haiku 5.5 is subject to strict safety and alignment guardrails set by Anthropic. While the cybersecurity safeguards are more permissive than those on Claude Sonnet 5.5 to allow defensive operations, the model still blocks penetration testing and other offensive techniques.

Biology safeguards restrict requests that could cause physical harm while still permitting basic research biology questions. Furthermore, developers working with large volumes should monitor context limits closely, as prompts longer than 100,000 tokens trigger the 5x pricing multiplier.

Frequently asked questions

What is Claude Haiku 5.5?
Claude Haiku 5.5 is a fast, highly capable small AI model developed by Anthropic. Released on 7 October 2026, it is engineered for high-volume, cost-sensitive tasks like data classification, short summaries, and customer support.
How much does Claude Haiku 5.5 cost?
Claude Haiku 5.5 costs $0.10 per million input tokens and $0.50 per million output tokens for prompts up to 100,000 tokens. Beyond 100,000 tokens, the price increases fivefold to $0.50 per million input and $2.50 per million output tokens.
Can I adjust the reasoning levels in Claude Haiku 5.5?
Yes, Claude Haiku 5.5 is the first model in its tier to feature adjustable reasoning effort levels. Users can select settings from low to max, though reasoning cannot be turned off completely and defaults to medium.
How do I access Claude Haiku 5.5?
Developers can access Claude Haiku 5.5 via the Claude Platform using the model identifier claude-haiku-5-5. The model is also available through major cloud providers including Amazon Web Services, Google Cloud, and Microsoft Azure.
What are the safety restrictions on Claude Haiku 5.5?
Claude Haiku 5.5 includes defensive cybersecurity safeguards that block penetration testing and offensive exploits. It also features biology safeguards that restrict access to harmful requests while permitting standard scientific research.

Sources: Anthropic · Simon Willison

Want your brand to be the answer AI gives?

See how ready your website is for ChatGPT, Gemini and Perplexity, free, in about ten seconds.

Related: How to get your brand cited by ChatGPT