Justin McKelvey
Fractional CTO · 15 years, 50+ products shipped
Claude Haiku 5.5 (2026): API Pricing, the 100K-Token Price Line, and What It Costs on Real Work
Claude Haiku 5.5 is Anthropic's new small model, released October 7, 2026. On the API it costs $0.10 per million input tokens and $0.50 per million output tokens for prompts up to 100,000 tokens, which is 90% under Haiku 4.5's $1 / $5. Prompts over 100,000 tokens pay $0.50 / $2.50. It is built for high-volume, repetitive work: classification, routing, summaries, live support and subagent jobs. It's also selectable in the Claude app on every plan, Free included. The 100,000-token line is the number to design around.
Small models don't get headlines. They get the bill. If your business runs anything through the Claude API at volume, tagging tickets, summarizing calls, pulling fields out of invoices, this is the launch that moves the number.
Here's what Anthropic shipped, what it costs as of October 8, 2026, two workloads priced out, and the one design rule the new price sheet quietly creates.
What is Claude Haiku 5.5?
It's the third model in the Claude 5.5 family, after Opus 5.5 (September 22) and Sonnet 5.5 (September 28). Anthropic calls it "the cheapest, fastest, and most capable small model we've ever released," aimed at "quick and repetitive workloads (like summaries, compactions, database queries, and classification requests)."
- Speed: Anthropic's fastest model at standard speed, which is why it's pitched for live customer support and browser use.
- Effort control: the first Haiku with an adjustable effort setting, so you can trade cost against quality per task.
-
Where it runs: the Claude API (model ID
claude-haiku-5-5), AWS, Google Cloud and Microsoft Azure, Claude Code, and the Claude app on Free, Pro, Max, Team and Enterprise.
Anthropic's own benchmark table puts it far ahead of Haiku 4.5: 72.4% vs 15.7% on OSWorld 2.1 computer use (offline subset), 1620 vs 735 on GDPval-AA v2.1 knowledge work. Sonnet 5.5 still scores higher on every row, and Anthropic says plainly that Sonnet and Opus "remain better choices for complex agentic coding tasks." Haiku is the fast hands. The bigger models are still the planner. If you're choosing a default model for a business, that call lives in the best Claude model for business.
Claude Haiku 5.5 pricing
Per million tokens, from Anthropic's pricing docs, read October 8, 2026:
| Model | Input | Output | Cache hit | 5-min cache write |
|---|---|---|---|---|
| Haiku 5.5, prompts up to 100K tokens | $0.10 | $0.50 | $0.01 | $0.125 |
| Haiku 5.5, prompts over 100K tokens | $0.50 | $2.50 | $0.05 | $0.625 |
| Haiku 4.5 | $1 | $5 | $0.10 | $1.25 |
| Sonnet 5.5 | $2 | $10 | $0.10 (cut from $0.20 on Oct 7) | $2.50 |
| Opus 5.5 | $4 | $20 | $0.20 | $5 |
The Batch API halves Haiku 5.5 again: $0.05 / $0.25 under 100K, $0.25 / $1.25 over. Anthropic's headline is "around 75% less to run" than Haiku 4.5 on average. That's lower than the 90% rate cut for two reasons it states itself: some requests land over 100K, and the new tokenizer "uses slightly more tokens per task." Every other model's rates, plus the subscription side, are in Anthropic API pricing.
What does it cost on real work?
Rate-card arithmetic, no caching, and no adjustment for the tokenizer, so treat these as close estimates, not quotes.
Job 1: support ticket triage, 20,000 tickets a month
Each ticket: about 1,500 tokens in (the ticket plus your routing rules) and 300 out (a category, a priority, a draft first reply). That's 30 million input tokens and 6 million output a month.
| Model | Monthly cost |
|---|---|
| Haiku 5.5 | $6 |
| Haiku 4.5 | $60 |
| Sonnet 5.5 | $120 |
| Opus 5.5 | $240 |
Six dollars. For most businesses, that job just stopped being a line item.
Job 2: questions against long documents, 1,000 a month
Each request sends a 150,000-token document (a long contract, a quarter of call transcripts) and gets about 1,000 tokens back. That's 150 million input tokens, and every one of those prompts is over the 100K line.
| Setup | Monthly cost |
|---|---|
| Haiku 5.5, whole document per request (over-100K rate) | $77.50 |
| Haiku 5.5, same text split into chunks under 100K | about $15.50 |
| Haiku 4.5, whole document | $155 |
| Sonnet 5.5, whole document | $310 |
Same model, same text, five times the bill, purely because of where the prompt length lands. The chunked figure ignores overlap between chunks and the extra calls to merge answers, so the real saving is a bit smaller. The direction isn't.
The 100,000-token line is a design rule
Under 100K tokens, Haiku 5.5 is the cheapest Claude by a mile. Over it, it's still cheaper than Haiku 4.5, but only by half. So the build question changes from "which model?" to "how big is each request?"
- Keep standing instructions short. A 20-page system prompt on every call eats the budget you could spend on the actual document.
- Chunk long inputs on purpose. Split by section, ask per section, merge at the end. Haiku is fast enough that the extra calls don't hurt.
- Let a bigger model plan and Haiku do the reading. That's the subagent pattern Anthropic is pitching: Opus or Sonnet breaks down the job, Haiku runs the many small pieces.
Does this change your Claude plan?
Not the price. Pro, Max, Team and Enterprise cost what they did on October 7, and the full tier map is in Claude AI pricing. Two things did change for subscribers:
- Haiku 5.5 is in the model picker on every plan, Free included, for when you want speed over depth.
- Max and Team now come with a monthly API credit, announced the same day: $100 on Max 5x, $200 on Max 20x, and $20 per Standard Team seat ($100 per Premium) pooled up to $500 a month. A 10-person Team on Standard seats gets $200 a month, which covers Job 1 above more than thirty times over. Who qualifies, how to claim it and what it won't pay for are in Claude API credits.
Should your business switch to Haiku 5.5?
- Switch now if you run Haiku 4.5 on classification, extraction, summaries or routing. Same job shape, better model, roughly a tenth of the rate under 100K. Test on a few hundred of your real inputs first.
- Try it as a downgrade for anything you run on Sonnet that is narrow and repetitive. Many "we use Sonnet for everything" setups are paying 20 times the rate for work Haiku can do.
- Don't switch the planner in a multi-step agent, or complex coding work. Anthropic says the bigger models stay better there, and its own numbers back that up.
If the honest answer is "we don't run anything on the API, we just have a few people in the chat app," the launch matters less than the setup. Getting a team's Claude workspace connected to the tools it already uses is my Claude for Small Business install: from $4,500, about two weeks, roughly three hours of your time. Seats are paid to Anthropic.
Sources
- Anthropic, "Introducing Claude Haiku 5.5", October 7, 2026, anthropic.com/claude-haiku-5-5 (read October 8, 2026)
- Anthropic, Claude Haiku product page, anthropic.com/claude/haiku (read October 8, 2026)
- Anthropic, Pricing, platform.claude.com/docs/en/about-claude/pricing (read October 8, 2026)
- Anthropic Help Center, "Monthly API credits for Max and Team plans", support.claude.com/en/articles/17154008 (read October 8, 2026)
Get the Free AI Content Toolkit
The exact system I use to turn one idea into a month of content: atomization framework, voice template, prompt library, weekly system.
Frequently Asked Questions
- How much does Claude Haiku 5.5 cost?
- As of October 2026, Claude Haiku 5.5 is priced by prompt length on the Claude API. For prompts up to 100,000 tokens: $0.10 per million input tokens, $0.50 per million output tokens, and $0.01 for cache hits. For prompts over 100,000 tokens: $0.50 input, $2.50 output, $0.05 cache hits. The Batch API halves both sets. Haiku 4.5 was $1 / $5 at every length.
- Is Claude Haiku 5.5 free?
- In the Claude app, yes, as one of the models you can pick: Anthropic says Free, Pro, Max, Team and Enterprise users can select Haiku 5.5 on Claude.ai, web, iOS and Android, within each plan's usage limits. On the API it is pay-per-token, and Max and Team subscribers now get a monthly API credit that Haiku 5.5 can draw on ($100 on Max 5x, $200 on Max 20x, $20 per Standard Team seat pooled up to $500).
- Is Claude Haiku 5.5 better than Haiku 4.5?
- On Anthropic's own published benchmarks, by a wide margin: 72.4% vs 15.7% on OSWorld 2.1 (offline subset) and 1620 vs 735 on GDPval-AA v2.1, at roughly a tenth of the per-token price for prompts under 100,000 tokens. Customer tests quoted in the launch were in the same direction (AlphaSense 0.84 vs 0.76 on 400 document questions). Run your own prompts before you switch production work.
- Should I use Haiku 5.5 or Sonnet 5.5?
- Haiku 5.5 for narrow, repeated jobs: classification, routing, summaries, extraction, live support replies, and subagent work under a bigger model. Sonnet 5.5 or Opus 5.5 for long multi-step agent work and complex coding, which Anthropic itself says they remain better at. On a 20,000-ticket triage job the rate-card difference is about $6 a month on Haiku 5.5 vs $120 on Sonnet 5.5.
- Why does Haiku 5.5 cost more over 100,000 tokens?
- Anthropic prices Haiku 5.5 by prompt length: a request whose prompt is over 100,000 tokens pays $0.50 / $2.50 per million instead of $0.10 / $0.50. Anthropic says about 90% of requests to the previous Haiku were under that line. If your job routinely sends whole contracts or long transcripts, splitting them into chunks under 100,000 tokens keeps you on the cheap rate.
More on AI for Business
AI for HR (2026): What a 10 to 200 Person Company Can Hand It, What Stays Human, and Where Each HR Record Goes
How a small HR team should use AI: draft policies, job posts and onboarding from your own handbook, keep every decision about a person human, and sort each HR record into what can and can't go into an AI tool. A record-by-record data map and a one-week start plan.
ChatGPT for Finance Teams (2026): Month-End Close, Reconciliations, Variance Commentary, and What Never to Paste
How an in-house finance team at a 10 to 200 person company uses ChatGPT in 2026: where it fits in the month-end close, why it never does the reconciliation itself, a worked variance-commentary example, a three-tier map of what data goes where, and why the plan decision (Business at $20 a seat vs a personal Plus account) is really a data decision.
AI for Project Managers (2026): The Six Artifacts to Hand It, What You Still Check, and a One-Week Plan
How a project manager should actually use AI: hand it the six documents you rewrite every week (status report, RAID log, meeting actions, schedule risk, stakeholder update, change request), keep the judgment, and check each one against a short list. Copy-paste prompts and a one-week time audit included.
Claude API Credits (2026): The New Monthly Credit on Max and Team, How to Claim It, and What It Buys
As of October 2026, Claude Max and Team plans come with monthly Claude API credits: $100 on Max 5x, $200 on Max 20x, and $20 per Standard Team seat ($100 per Premium) pooled up to $500. Who qualifies, how to claim it, what it covers (not interactive Claude Code), what it buys on three real workloads, and the other ways to get Claude API credits.
Written by
Justin McKelvey
Fractional CTO & AI consultant in Austin, TX. 15 years building software, 50+ products shipped, $53M+ in client revenue generated. I help $1M–$50M founders ship production software and automate operations with AI, without hiring a full-time executive team.
Work with meIf this was useful, here are two ways I can help:
Before you go
Score your AI readiness in 3 minutes.
Free, no call. 30 yes/no questions show where AI actually pays off in your business (and where it doesn't yet).
Get my readiness score →