HomeLearn

What Is Claude Haiku 5.5? Anthropic's Fastest Small Model Explained

What is Claude Haiku 5.5? Anthropic's fastest small model costs $0.10 per 1M input tokens. See what's new, pricing tiers, benchmarks, and where to use it.

Bhavyadeep
Written by
Bhavyadeep
Priyanka Singh
Reviewed by
Priyanka Singh
Last updated: 
October 8, 2026
0
 min read
Select Emergent as your Preferred news source
Table of Contents

TL;DR

  • What it is: Claude Haiku 5.5 is Anthropic's fastest and cheapest small model, released October 7, 2026, for high-volume work like summaries, classification, routing, and subagent tasks
  • Price: $0.10 per million input tokens and $0.50 per million output tokens for prompts up to 100,000 tokens, about 75% cheaper to run than Haiku 4.5 on average
  • What's new: an adjustable effort setting, a 1-million-token context window, 128,000-token output, and a large jump in computer use
  • Limits: Sonnet 5.5 and Opus 5.5 still lead on complex coding and long, multi-step agent work
  • Best for: well-defined jobs you need to run thousands of times a day


Claude Haiku 5.5 is the model for tasks that are simple but happen thousands of times a day. Anthropic released it on October 7, 2026, at a tenth of its predecessor's per-token rate for most requests.

If you pay for a large model to sort emails, summarize calls, or answer routine support questions, that gap matters. Most of those jobs don't need the most powerful model, and paying for one adds up fast.

Here's what Claude Haiku 5.5 is, what changed from Haiku 4.5, what it costs, and how it scores in Anthropic's tests and in independent ones. All figures are as of October 2026.

What is Claude Haiku 5.5?

Claude Haiku 5.5 is Anthropic's fastest and lowest-cost Claude model, built for high-volume, time-sensitive work such as classification, extraction, routing, and subagent tasks. Released on October 7, 2026, Haiku 5.5 costs $0.10 per million input tokens for prompts up to 100,000 tokens and reads up to 1 million tokens at once.

In its launch announcement, Anthropic names the jobs it was built for:

  • Quick, repetitive work: summaries, compressing long conversations, database lookups, and sorting requests into categories.
  • Subagent tasks: a larger model such as Opus 5.5 or Sonnet 5.5 plans the work and hands Haiku the small, well-defined pieces.
  • Speed-sensitive jobs: live customer support and browser automation, where every second of waiting counts.

A subagent is simply a helper model. Think of a lead model as the manager who decides what needs doing, and Haiku as the assistant who fetches a figure from a report or files a ticket. Developers call it with the model ID claude-haiku-5-5.

Where Haiku 5.5 fits in the Claude 5.5 family

Haiku 5.5 is the fastest and cheapest of Anthropic's four current model tiers. It costs one-twentieth of Sonnet 5.5 per token on short prompts, and it is the tier Anthropic recommends for volume rather than difficulty.

Model Price per 1M tokens (input / output) Speed Default effort Positioned for
Claude Fable 5.1 $10 / $50 Slower High Anthropic's most capable tier
Claude Opus 5.5 $4 / $20 Moderate Medium Complex coding and knowledge work
Claude Sonnet 5.5 $2 / $10 Fast High Well-scoped coding and agent tasks
Claude Haiku 5.5 From $0.10 / $0.50 Fastest Medium High-volume, speed-sensitive tasks

The Claude model lineup, pricing as of October 2026. All four models share a 1-million-token context window, 128,000-token output, and a June 2026 knowledge cutoff. Source: Claude Platform docs.

The practical takeaway is to treat these tiers as a team, not a ranking. A bigger model is worth its price on hard problems, while Haiku handles the many small jobs around them.

What's new in Claude Haiku 5.5

Haiku 5.5 brings the main features of Anthropic's larger 5.5 models down to the lowest price tier. Five changes matter most.

1. A much lower price, with a 100K-token line

Haiku 5.5 charges 90% less per token than Haiku 4.5 for prompts up to 100,000 tokens, and 50% less above that. Anthropic says about 90% of requests to Haiku 4.5 fell under the line, so most workloads land in the cheap tier.

The average saving is lower, at about 75%. Haiku 5.5 uses a newer tokenizer that counts the same text as roughly 30% more tokens than Haiku 4.5 did, which eats into the per-token discount. Text that cost $1 to read on Haiku 4.5 costs about $0.13 on Haiku 5.5 in the low tier, not $0.10. Add the higher rate on prompts over 100,000 tokens, and Anthropic puts the average saving at about 75%.

2. Adjustable effort, for the first time on a Haiku model

Effort controls how long the model thinks before answering. Haiku 5.5 offers settings from low to max and starts at medium. Low suits quick labels and routing, while higher settings help on tricky, multi-step checks.

Haiku 4.5 had no effort dial, only a fixed thinking budget that developers set by hand. More effort means better answers on hard tasks, but also more thinking tokens on your bill.

3. A 1-million-token context window and 128K-token output

The context window is how much text a model can read in one request. Haiku 5.5 takes up to 1 million tokens, five times Haiku 4.5's 200,000, and can write up to 128,000 tokens in a single reply. Through Anthropic's Batch API, a beta option raises that output limit to 300,000 tokens.

That lets Haiku read the same large inputs as the bigger model directing it, such as a full contract set or a long support history. Reading that much, though, pushes a request past the 100,000-token line and into the higher price tier.

4. Much stronger computer and browser use

Haiku 5.5 can now operate a computer reliably enough to be useful. On OSWorld 2.1, a test of completing tasks on a real desktop, Anthropic reports 72.4% against 15.7% for Haiku 4.5.

Anthropic points to repetitive jobs like form filling, data entry, and moving information between apps. It also added computer use and browser use tools to its Python and TypeScript SDKs in beta on launch day.

5. Built for subagent work

Anthropic designed Haiku 5.5 to work under a larger model. Fable or Opus plans the work, and Haiku runs the small subtasks, which makes it practical to run many agents in parallel.

Customers describe the same split. Financial AI company Rogo says a bigger model builds a presentation while Haiku 5.5 pulls the revenue figure it needs from an annual report.

Claude Haiku 5.5 vs Haiku 4.5: what changed

Haiku 5.5 is a generational jump over Haiku 4.5: five times the context, a tenth of the short-prompt price, and usable scores on tasks Haiku 4.5 could barely attempt.

Spec Claude Haiku 4.5 Claude Haiku 5.5
Released October 15, 2025 October 7, 2026
Status Legacy Latest
Context window 200,000 tokens 1,000,000 tokens
Max output 64,000 tokens 128,000 tokens
Input / output price per 1M tokens $1 / $5 $0.10 / $0.50 up to 100K tokens, $0.50 / $2.50 above
Thinking Manual thinking budget Adaptive thinking with an effort setting
Knowledge cutoff February 2025 June 2026
OSWorld 2.1 (computer use, Anthropic) 15.7% 72.4%
Terminal-Bench 4.0 (command line, Anthropic) 0.0% 39.2%
Intelligence Index (Artificial Analysis, max effort) 17 43

Claude Haiku 4.5 vs Claude Haiku 5.5 specs, pricing, and scores, as of October 2026. Sources: Anthropic and Artificial Analysis.

Haiku 4.5 is still available, but Anthropic now lists it as a legacy model and recommends moving to Haiku 5.5.

How much does Claude Haiku 5.5 cost?

Claude Haiku 5.5 costs $0.10 per million input tokens and $0.50 per million output tokens for prompts up to 100,000 tokens. Once a prompt goes over 100,000 tokens, every rate rises five times.

Rate per 1M tokens Prompts up to 100K tokens Prompts over 100K tokens
Input $0.10 $0.50
Output $0.50 $2.50
Cache read $0.01 $0.05
Cache write, 5 minutes $0.125 $0.625
Cache write, 1 hour $0.20 $1.00

Claude Haiku 5.5 API pricing as of October 2026, from Anthropic's Claude Platform docs. The Batch API takes 50% off input and output.

A token is a small chunk of text, often a whole short word or part of a longer one. Input is what you send the model, and output is what it writes back.

In everyday terms, the low tier is very cheap. A support reply that reads 3,000 tokens of context and writes 500 tokens costs about $0.00055, so 1,000 replies come to roughly $0.55. A 150,000-token contract summary is a different story: at the higher tier, 1,000 of them with 2,000-token outputs cost about $80.

Two more factors move the real bill. Hidden thinking tokens count as output, so higher effort settings cost more per request. On the other side, Anthropic says prompt caching can save up to 90% when you reuse the same instructions or documents, and batch processing saves 50% on work that can wait.

Claude Haiku 5.5 benchmarks: Anthropic's numbers vs independent tests

Anthropic's tests show Haiku 5.5 far ahead of Haiku 4.5 and ahead of OpenAI's GPT-6 Luna, but still behind Sonnet 5.5 everywhere. Independent results from Artificial Analysis confirm a big jump and add two caveats Anthropic's table leaves out.

1. What Anthropic reports

Benchmark What it tests Haiku 5.5 Haiku 4.5 GPT-6 Luna Sonnet 5.5
GDPval-AA v2.1 (Elo) Real-world professional tasks 1,620 735 1,437 1,840
AA-Briefcase v1.1 (Elo) Longer knowledge-work projects 1,578 614 1,336 1,824
OSWorld 2.1, offline subset Operating a computer 72.4% 15.7% 48.9% 83.9%
Terminal-Bench 4.0 Multi-step command-line tasks 39.2% 0.0% 16.4% 70.6%
Humanity's Last Exam, with tools Expert-level reasoning 57.4% 18.7% Not reported 64.5%
Chartography, no tools Reading charts 46.4% 6.4% 29.1% 61.6%

Vendor-reported results from Anthropic's Claude Haiku 5.5 launch announcement, October 2026. Elo scores are relative ratings, not percentages.

The gap to Sonnet 5.5 is widest on long coding jobs. Haiku 5.5 completes about four in ten Terminal-Bench tasks, against seven in ten for Sonnet 5.5, which is why Anthropic still recommends its larger models for complex coding.

2. What Artificial Analysis measured

Artificial Analysis, an independent benchmarking firm, scores Haiku 5.5 at 43 on its Intelligence Index at max effort. That is 26 points above the previous Haiku and a little ahead of other small models such as GPT-6 Luna at 38, according to its launch analysis. Sonnet 5.5 still leads at 56. Its own Terminal-Bench 4.0 run put Haiku 5.5 at 33%, below the 39.2% in Anthropic's table, because the two test setups differ.

The same analysis flags two caveats:

  • Heavy token use: at max effort, Haiku 5.5 used about 162,000 output tokens per test task, roughly three times GPT-6 Luna. More tokens mean a higher bill per finished job, even at the same per-token price.
  • Less stored knowledge, fewer made-up answers: Haiku 5.5 answered 36% of factual questions correctly, below GPT-6 Luna's 44%. Its hallucination rate was lower, though, at 40% against 77% for Luna.

Artificial Analysis also expects one weak score to rise. On AutomationBench, which tests multi-step business workflows, Haiku 5.5 scored 35%. A safety issue during pre-release testing made the model refuse too often, and Artificial Analysis plans to re-run the test once Anthropic's fix is in place.

What Claude Haiku 5.5 is good at, and where it falls short

Haiku 5.5 is strongest on short, repeated, well-defined jobs and weakest on long, open-ended ones. The table below sums up where it fits.

Task Fit Why
Sorting, tagging, and routing requests Strong Fast, cheap, and simple enough for medium or low effort
Summarizing documents and conversations Strong Cheap under 100,000 tokens, with a 1M-token window for longer inputs
Live chat, voice, and in-app assistants Strong Anthropic's fastest model at standard speed
Form filling and data entry in a browser Strong 72.4% on OSWorld 2.1, up from 15.7%
Helper tasks under a larger model Strong Designed as a subagent for Opus 5.5 or Fable 5.1
Complex coding and long agent tasks Weak 39.2% on Terminal-Bench 4.0, against 70.6% for Sonnet 5.5
Answering facts from memory Weak 36% accuracy on Artificial Analysis's knowledge test
Very long documents at high volume Costly Prompts over 100,000 tokens pay five times the base rate

Where Claude Haiku 5.5 fits, based on Anthropic's guidance and benchmark results as of October 2026.

Early customers report results that match the strong rows. AlphaSense, which runs about 8 million document questions a week, saw a statistically significant accuracy gain over Haiku 4.5 on 400 test queries. Box reported scores 11 points higher than Haiku 4.5 at about half the latency. Both are testimonials Anthropic published, not independent tests.

The weak rows have simple workarounds. Hand complex builds to a larger model, and give Haiku the source material rather than asking it to recall facts unaided.

Where you can use Claude Haiku 5.5

Claude Haiku 5.5 is available in the Claude apps for every plan, through Anthropic's API, and on the three major cloud platforms.

Where Who it's for How to access it
Claude apps (web, iOS, Android) Everyone Select Haiku 5.5 on the Free, Pro, Max, Team, or Enterprise plan
Claude API Developers Model IDclaude-haiku-5-5
Amazon Bedrock AWS customers Model IDanthropic.claude-haiku-5-5
Google Cloud and Microsoft Foundry Cloud customers Model IDclaude-haiku-5-5
Claude Code Developers Available as a model choice
GitHub Copilot Developers Generally available, billed at list price under usage-based billing

Claude Haiku 5.5 availability as of October 2026, from Anthropic's Haiku page and the GitHub changelog.

Anthropic has committed to keeping Haiku 5.5 available until at least October 7, 2027. That gives teams that build on it a full year before any retirement.

Should you switch from Haiku 4.5?

For most high-volume tasks, yes. Haiku 5.5 is cheaper per token on short prompts and far more capable, but swapping the model name is not the whole job.

Check three things before you move a live product or automation:

  • Settings that now fail: Haiku 5.5 rejects any non-default value for temperature, top_p, or top_k. A setup that pinned one of these, such as a classifier set to temperature 0, will return an error until the setting is removed.
  • Thinking is on by default: Haiku 5.5 decides when to think, and that thinking shows up in its responses and on your bill. Pick an effort level on purpose rather than relying on the default.
  • Token counts go up: the newer tokenizer counts the same text as about 30% more tokens. Recount your prompts, both for cost and to see which ones now cross the 100,000-token line.

Plan the move soon. Anthropic's docs list Haiku 4.5's retirement as no sooner than October 15, 2026, just eight days after Haiku 5.5 launched, so start testing Haiku 5.5 on a slice of real traffic now.

What Claude Haiku 5.5 means if you're building AI-powered apps

Haiku 5.5 is a reminder that the right AI model depends on the job, not the leaderboard. Most features inside an app, like tagging support tickets or summarizing a customer's history, don't need the most powerful model available.

A simple way to think about it: match each feature to the smallest model that does it well. Use a fast, low-cost model for frequent, simple jobs, and save a larger one for the work where mistakes are expensive. Splitting features this way can cut your running costs sharply as usage grows.

Use Haiku 5.5 for volume and a larger model for depth

Claude Haiku 5.5 is Anthropic's answer to the jobs that happen all day: sorting, summarizing, looking things up, and helping a bigger model finish its work. It costs a tenth of Haiku 4.5's rate on short prompts, reads up to 1 million tokens, and finally handles computer use well enough to be practical.

It is not a replacement for Sonnet 5.5 or Opus 5.5. Independent tests show a real jump in capability, but also heavy token use at max effort and weaker recall of facts. Long prompts over 100,000 tokens also cost five times as much.

The smart next step is to list the AI tasks you run most often and try the simplest ones on Haiku 5.5 at medium effort. If you're turning those tasks into an app, choose from Claude, GPT, and Gemini models for each feature and start building on Emergent.

Was this article helpful?
About the writer

Bhavyadeepsinh Rathod is SEO Content Manager at Emergent.sh, where he covers the tools, frameworks, and workflows driving the next era of vibe coding. With 8+ years in tech content marketing, he brings a sharp SEO lens to complex subjects, making Emergent's ecosystem of AI builder tools discoverable for the builders, creators, and teams that need them most. He specializes in making complex topics feel simple, relevant, and easy to act on.

Cta image

Most AI app builders stop at prototypes. Emergent creates production-ready apps you can actually launch.

  • Production-ready apps
  • Web & mobile apps
  • Deploy in minutes
Try For Free
Share this article:

Frequently Asked Questions

Your Questions, Answered

What is Claude Haiku 5.5 used for?
Claude Haiku 5.5 is built for high-volume, time-sensitive work: classification, summarization, request routing, data extraction, and live chat or support. It also works as a subagent, handling small subtasks for a larger model such as Opus 5.5. Anthropic recommends its bigger models for complex coding and long, open-ended agent tasks.
What is the price of Claude Haiku 5.5 per million tokens?
On Anthropic's API, Claude Haiku 5.5 costs $0.10 per million input tokens and $0.50 per million output tokens for prompts up to 100,000 tokens. Longer prompts cost $0.50 and $2.50. Cache reads start at $0.01 per million tokens, and the Batch API takes 50% off input and output.
Is Claude Haiku 5.5 free?
You can select Claude Haiku 5.5 in the Claude apps on the Free plan, as well as on Pro, Max, Team, and Enterprise. Using it through Anthropic's API, cloud platforms, or tools like GitHub Copilot is paid, billed by the number of tokens you send and receive.
What is the Claude Haiku 5.5 model ID?
The model ID is claude-haiku-5-5 on the Claude API, Google Cloud, Microsoft Foundry, and Claude Platform on AWS. On Amazon Bedrock, it is anthropic.claude-haiku-5-5.
Is Claude Haiku 5.5 better than Sonnet 5.5?
No. Sonnet 5.5 scores higher on every benchmark Anthropic published, and the gap is largest on long coding tasks: 70.6% vs 39.2% on Terminal-Bench 4.0. Haiku 5.5 costs one-twentieth as much per token on short prompts, so it suits simple, repeated tasks where Sonnet's extra capability isn't needed.
Why is Claude Haiku 5.5 only about 75% cheaper on average?
The per-token rate is 90% lower than Haiku 4.5 for prompts up to 100,000 tokens, but two things shrink the real saving. Haiku 5.5's newer tokenizer counts the same text as about 30% more tokens, and prompts over 100,000 tokens get only a 50% discount.
Start Building
on Emergent today
Try Emergent

https://api.linear.app/graphql