Justin McKelvey

Justin McKelvey

Fractional CTO · 15 years, 50+ products shipped

• AI for Business • 13 min read •

Which Claude Model Is Best for Business? Sonnet vs Opus vs Haiku (2026)

By Justin McKelvey, fractional CTO · 15+ years shipping products · Updated August 2026

Claude Sonnet 4.6 is the right default for roughly 90% of business workflows in 2026 — fast enough for real-time drafts, strong enough for serious reasoning, priced where it doesn't punish you for using it daily. Claude Opus 4.7 is reserved for the hardest reasoning tasks or one-shot heavy lifts (overkill and roughly 2.5x the cost per token for daily work). Claude Haiku is best for high-volume, simple tasks like classification, summaries, and batch drafts. Most businesses on Claude Team should set Sonnet 4.6 as the default and switch up to Opus only when a task actually needs it. Switching is two clicks. August 2026 update: the Claude 5 family has arrived — Fable 5, Opus 5, Sonnet 5 — and the same decision logic applies one generation up; details in the update section below.

Every week a client asks me some version of: "Which Claude model should I use?" The product page lists three, the chat dropdown has more, and the wrong default can quietly 5x your bill or starve a workflow that needed more horsepower. Here's the short, honest answer from someone who installs Claude into small businesses for a living.

Claude model comparison at a glance

Model Speed Cost (per 1M tokens) Context window Best for
Claude Sonnet 5 Fast (real-time) $2 in / $10 out 1M Daily business work — drafts, analysis, client emails, light coding, the 90% default
Claude Opus 5.5 Moderate (thinks longer) $4 in / $20 out 1M Hard reasoning, deep research, complex code refactors, long agent runs, one-shot heavy lifts
Claude Haiku 4.5 Fastest $1 in / $5 out 200K High-volume, simple tasks — classification, summaries, batch drafts, real-time chat
Claude Fable 5.1 Slower $10 in / $50 out 1M The escalation, when Opus 5.5 at higher effort still falls short
Sonnet 4.6 / Opus 4.7 / Opus 5 (older, still available) — $3 / $15 · $5 / $25 · $5 / $25 1M Only if a workflow is pinned to them; the newer models are cheaper or better

Correction (September 10, 2026), re-read from Anthropic's pricing page today: the Opus rate above is $5 / $25 per million tokens, not the $15 / $75 an earlier version of this table carried (that was the retired Opus 4.1 rate); Haiku is now Haiku 4.5 at $1 / $5 (Haiku 3.5's $0.80 / $4 is retired); and the Claude 5 generation prices at Sonnet 5 $2 / $10, Opus 5 $5 / $25, Haiku 4.5 $1 / $5, with the new top tier, Fable 5.1, at $10 / $50. Cache hits are one tenth of the input price on every model. The ratios in the sections below were corrected the same day: Opus is about 2.5x Sonnet per token, not 5x, and Haiku is a fifth of Opus, not a twentieth.

Correction (September 27, 2026), re-read from Anthropic's pricing page, model overview and plan table today: the table above now shows the current lineup. Anthropic released Claude Opus 5.5 on September 22, 2026 at $4 / $20 per million tokens, 20% under Opus 5, with cache hits at $0.20, and its model guide now says to start with Opus 5.5 for most workloads. That makes the current Opus 2x Sonnet 5 per token, not 2.5x; the 2.5x figures below are the Opus 5 and Opus 4.7 ratio and still hold for anyone pinned to those models. The recommendation doesn't change: Sonnet by default, Opus for the hard jobs, Haiku on volume. Only the Opus you reach for does. The full launch read is on Claude Opus 5.5.

Three models, three different jobs. Most teams quietly use whichever one is on top of the dropdown that week. That's the bug.

August 2026 update: the Claude 5 family is here — same logic, new names

Anthropic's Claude 5 generation began rolling out in 2026, and it adds something the 4.x lineup never had: a tier above Opus. The new flagship is Claude Fable 5 — Anthropic calls the tier "Mythos-class" — the most capable generally available Claude model, with Opus 5 and Sonnet 5 as the generational successors to the models in the table above, and Haiku 4.5 as the current small model. (There is also a Claude Mythos 5 — the same underlying model as Fable 5, offered to approved organizations without the additional dual-use safety measures — so for a business buyer, the top of the menu is Fable 5.)

Here's the part that matters for your bill and your defaults: nothing about the decision logic changes — it just shifts up one generation. The mid-tier model (now Sonnet 5) stays the right default for daily business work. The heavyweight (Opus 5, or Fable 5 where your plan offers it) stays reserved for tasks where reasoning depth is the product — strategy, hard analysis, one-shot heavy lifts. The small model stays on volume work. If your workspace defaults are set the way this guide describes, a generation upgrade is a dropdown change, not a re-architecture — Projects and Skills carry forward untouched.

Two practical notes from the first months of the 5 family. First, check current per-token rates on Anthropic's pricing page before moving any high-volume automation up a generation — the tier logic transfers, but exact numbers move with each release, and the cost trap described below gets more expensive one tier up. Second, resist the new-flagship reflex: the arrival of a tier above Opus doesn't mean your email triage suddenly needs it. The 90% of business work that ran perfectly on the mid-tier yesterday still runs perfectly on the mid-tier today.

September 2026 update: Claude Opus 5.5 is the Opus to use now

On September 22, 2026 Anthropic shipped Claude Opus 5.5, the first model in a Claude 5.5 family. The numbers that matter for this guide: $4 per million input tokens and $20 output, 20% under Opus 5; cache hits at $0.20, 60% under Opus 5; and, in Anthropic's words, it "performs at the level of Claude Fable 5.1 on most work." Its model guide now reads "start with Claude Opus 5.5 for most workloads," with Fable 5.1 kept for demanding reasoning and long-horizon agent work that Opus 5.5 at higher effort still can't handle.

What that does to the decision tree: the Opus branch gets cheaper and stronger, and the Fable branch gets narrower. On Pro, Max, Team and Enterprise, Opus is already in your plan, so there's no new bill. On the API, the gap between Sonnet and Opus drops from 2.5x to 2x per token. That still isn't a reason to make Opus the default. It's a reason to stop reaching for Fable first. Sonnet 5.5 followed on September 28 at Sonnet 5's exact per-token price ($2 / $10), so the Sonnet row did not move on price; Haiku 5.5 is still to come.

One plan detail that changed with it: Anthropic's plan table now lists Fable 5.1 on Team, but only on premium seats and capped at 50% of weekly limits. Standard Team seats get Opus, Sonnet and Haiku, and that's the right set for 90% of the work in this guide.

Sonnet 4.6: the right default for most business work

Sonnet 4.6 is the model you should set as your default in claude.ai, in Cowork, and in any Skill you build. Fast enough to feel real-time when you're drafting an email. Less than half the cost of Opus per token (Sonnet 5 is $2 / $10 per million against Opus 5's $5 / $25, as of September 2026). And the reasoning quality is genuinely close to Opus for most business tasks — proposals, newsletters, client deliverables, meeting summaries, lead scoring, basic code review. You will not feel the difference on the vast majority of work.

The places Sonnet 4.6 quietly shines: long-form writing with brand voice (especially when paired with a loaded Project), structured outputs like tables and JSON, multi-step reasoning that doesn't require novel research, and any workflow you run more than once a day. If you've been defaulting to Opus "just to be safe," you're paying 2.5x for a difference you can't measure on most tasks.

Opus 4.7: when the harder model is actually worth it

Opus 4.7 is the model you switch to when reasoning quality matters more than speed or cost. It thinks longer, holds more nuance, handles ambiguity better. For the right job, it's worth every cent. For the wrong job, it's a tax on your subscription.

The tasks where Opus actually earns its price tag: deep research where you're asking the model to weigh competing arguments, complex code refactors across multiple files, strategy documents where you want it to push back and find holes in your logic, legal or contract analysis where missing a clause matters, and any "one-shot" lift where you'd rather pay 2.5x once than re-prompt five times.

I use Opus for maybe 5% of my workflow: pricing strategy for a new offer, a board memo, a scope document for a six-figure engagement. Everything else stays on Sonnet. One pattern worth stealing: draft in Sonnet, then ask Opus to critique. You get most of Opus's value at a fraction of the cost because the critique is shorter than the draft.

Haiku: when "good enough" is actually the right answer

Haiku is the model nobody talks about, and it's the one that quietly saves businesses the most money on volume workflows. Fastest model Anthropic ships, half the cost of Sonnet 5 per token (Haiku 4.5 is $1 / $5 per million), and handles a surprising range of tasks well.

Use Haiku when you're doing something simple at high volume: classifying inbound leads as hot/warm/cold, summarizing call transcripts, drafting first-pass replies to common support tickets, tagging content, extracting structured data from emails, generating product description variants. It's also the right model for real-time chat experiences — customer-facing bots, internal lookup tools, anything where latency matters more than nuance.

The trap: don't use Haiku for anything that requires real reasoning. It's great at "do this simple thing fast," not "think through this carefully." If your output quality drops, jump back to Sonnet.

Cost comparison on a typical business workflow

Real numbers on a workflow most businesses have — triaging 500 inbound emails per week and drafting first-pass replies:

  • Haiku 4.5: roughly $2–$5/week. Near-instant. Quality is fine for templated replies and basic triage.
  • Sonnet 4.6 or Sonnet 5: roughly $6–$15/week. Noticeably better drafts that need less editing. Worth it if you're sending replies directly to clients.
  • Opus 4.7 or Opus 5: roughly $15–$38/week at the $5 / $25 rate (corrected September 2026 from an earlier $40–$75 estimate built on the retired $15 / $75 rate). Marginally better drafts that almost nobody can tell apart from Sonnet's on this task. Not worth it for inbox triage.

That's a 5x cost difference between Haiku and Opus on the same workflow, for an output difference your customers cannot perceive.

How to actually switch models (in claude.ai, Cowork, and Skills)

The mechanics are straightforward once you know where to look.

In claude.ai: there's a model dropdown above the message input. The default is whatever was selected last, which is why most people drift into whichever model was used most recently rather than the one that fits the task.

Inside Claude Cowork: each Skill can specify its own model. When you build a Skill, there's a model selector in the Skill settings. Set Haiku for high-volume classification skills, Sonnet for daily drafting skills, Opus for the rare reasoning-heavy ones. This is where most of the cost optimization lives.

Via the API: the model is set per-request via the `model` parameter (as of September 2026: `claude-sonnet-5`, `claude-opus-5-5`, `claude-haiku-4-5`, and `claude-fable-5-1` for the top tier). Set this once at the workflow level rather than guessing per-call. Inside a Project, the model selection carries through the conversation — you can switch mid-thread to, say, draft with Sonnet and then hit Opus for a critique pass on the same context.

When NOT to use Opus (the cost trap that catches almost everyone)

The single most common mistake I see when I audit a client's Claude setup: Opus is the default model on every workflow, and nobody remembers turning it on. This is the easiest way to 2.5x your Claude bill without realizing.

Don't use Opus 4.7 for: short questions you could answer in your head, formatting and rewriting tasks, summaries under a page, anything you're running on a schedule (cron jobs, automated reports, batch processing), real-time chat with customers, or "warm-up" prompts where you're still figuring out what you want.

The pattern that costs the most money: someone sets up a daily automation, leaves it on Opus by default, and it runs 30 times a month for six months before anyone audits the bill. The fix is always the same — downgrade to Sonnet, save 60%, see no quality drop.

The practical decision guide

Five branches that cover most business decisions:

  • If you're drafting client-facing copy (emails, proposals, newsletters): Sonnet 4.6.
  • If you're doing daily analysis, summaries, or "rewrite this in a friendlier tone": Sonnet 4.6.
  • If you're running a high-volume automation (classification, batch drafts, support triage): Haiku.
  • If you're making a real decision (strategy, pricing, board memo, contract review): Opus 4.7 for the heavy reasoning pass, Sonnet for the rest.
  • If you have no idea and just want to get unstuck: Sonnet 4.6. It is the right default 90% of the time.

This is the exact decision tree I install with clients when I set up Claude for their business — default Sonnet at the workspace level, Haiku on the high-volume Skills, Opus on the 2–3 Skills where reasoning depth actually moves the needle.

Set the right default and stop thinking about it

The honest takeaway: 90% of small businesses should set Sonnet 4.6 as the default everywhere — in claude.ai, in Cowork, in every Skill — and only switch up to Opus or down to Haiku when a specific task asks for it. Most people overpay for Opus they don't need, or underuse Haiku on volume work that would save them real money.

This is the same decision tree I configure when I install Claude for Small Business for clients. Default Sonnet, Haiku on the high-volume Skills, Opus on the 2–3 workflows where reasoning depth actually moves the needle. The setup takes an afternoon. The savings show up in the next billing cycle.

If you want the broader cost picture, I wrote a full Claude AI pricing breakdown — and the verdict on whether the $20 subscription itself is justified lives in Is Claude Pro worth it. Still deciding between Claude and another tool? The Claude vs ChatGPT for small business comparison is the head-to-head. Or skip the trial and error — my done-with-you Claude install ships with the model defaults already configured per Skill, so your business gets the right model on the right task from day one.

Claude Opus vs Sonnet: the math the pricing page does not do for you

Anthropic's pricing page gives you two numbers per model and stops. As of September 10, 2026 those numbers are Sonnet 5 at $2 per million input tokens and $10 per million output, Opus 5 at $5 / $25 (Opus 4.7 is priced the same), Haiku 4.5 at $1 / $5, and the new top tier, Fable 5.1, at $10 / $50. What the page does not do is the arithmetic that decides Opus vs Sonnet for a business. Three pieces of it. First, the ratio is 2.5x, not the 5x of the Opus 4.1 era, so the "just use Opus" tax got smaller this year; on a workflow that costs $10 a week on Sonnet, Opus is $25, not $50. Second, on a subscription you do not pay per token at all, but the same choice still costs you: the pricing page says every plan meters usage on a rolling five-hour window, that "how much you can do depends on the length and complexity of your conversations, the model you choose, and the features you use," and that there is no fixed message count. Opus on Pro drains the window faster than Sonnet does, so the model dropdown is a limits decision even when the bill is flat. Third, cache hits are one tenth of the input price on every model, so a Project with a long, stable system prompt makes the Opus premium smaller in practice than the list rate suggests. Fable is not on the Free plan and is metered separately on Pro and Max per the plan table, so it is not the default anywhere. September 27 update: the newest Opus, Opus 5.5, is $4 / $20, so the current ratio is 2x, not 2.5x, and a workflow that costs $10 a week on Sonnet 5 is $20 on Opus 5.5; the plan table also now lists Fable on Team premium seats at 50% of weekly limits. The decision rule survives the new numbers: Sonnet as the default, Opus on the two or three Skills where reasoning depth pays for itself, Haiku on volume.

Model choice is half the decision; the plan decides which models each seat gets. If you're buying for two or more people, my Claude Team plan breakdown covers seat prices, premium seats, and what admins control.

Free Resource Justin McKelvey

Get the Free AI Content Toolkit

The exact system I use to turn one idea into a month of content: atomization framework, voice template, prompt library, weekly system.

Frequently Asked Questions

Which Claude model is fastest?
Claude Haiku 4.5 is the fastest. As of September 2026, Anthropic's model overview rates the current lineup Haiku 4.5 "Fastest", Sonnet 5 "Fast", Opus 5.5 "Moderate" and Fable 5.1 "Slower". Sonnet 5 is fast enough to feel real-time for chat and drafting. Opus 5.5 thinks longer than Sonnet, though Anthropic says it writes output more than 30% faster than Opus 5, and Fable 5.1 is the slowest of the four.
Which Claude model is cheapest?
Claude Haiku 4.5 is the cheapest at $1 per million input tokens and $5 per million output tokens as of September 2026 — half the cost of Sonnet 5 ($2 in / $10 out), a third of Sonnet 4.6 ($3 / $15), and a fifth of Opus 5 or Opus 4.7 ($5 in / $25 out). For high-volume simple tasks, Haiku is the obvious pick. For complex work, the cost difference is worth paying for the better model. (Corrected September 10, 2026 from Anthropic's pricing page: an earlier version of this answer carried the retired Opus 4.1 rate of $15 / $75 and the retired Haiku 3.5 rate of $0.80 / $4.) Update September 27, 2026: the newest Opus, Opus 5.5, is $4 in / $20 out, so Haiku 4.5 is a quarter of the current Opus per token.
Which Claude model is best for writing?
Sonnet is the best Claude model for most business writing — proposals, newsletters, sales emails, client deliverables; as of September 2026 that means Sonnet 5. The writing quality is close enough to Opus that you can't reliably tell the drafts apart, and the speed makes it usable for daily work. Opus 5.5 is worth switching to for long-form pieces where you want the model to push back on your thinking, like strategy memos or pricing arguments, and Anthropic says it fixed the most common complaint about Opus 5's writing: it now puts the most important information first and follows your writing rules.
Which Claude model is best for coding?
Sonnet handles the vast majority of everyday coding work well — new features, refactors, debugging, code review. For hard architectural decisions, multi-file refactors that cross system boundaries, long agent runs, or a bug you've been stuck on for an hour, use Opus: as of September 2026 that's Opus 5.5, which Anthropic describes as built "for long-running agentic coding and knowledge work" and prices at $4 / $20 per million tokens, 20% under Opus 5. Haiku is fine for small scripts and one-liners but underpowered for real engineering work.
Can I use multiple Claude models in the same workflow?
Yes, and it's the smartest way to use Claude. A common pattern: draft with Sonnet 5, ask Opus 5.5 to critique the draft, then have Haiku 4.5 format the output. Each model does the part it's best at, and the total cost is lower than running everything through Opus. Inside Cowork, each Skill can specify its own model, so you can mix and match across a single business workflow without thinking about it.
What is the Claude 5 family and does it change this advice?
The Claude 5 family began rolling out in 2026: Claude Fable (now Fable 5.1) is the top tier — a Mythos-class model that sits above Opus in capability — alongside Opus and Sonnet 5, with Haiku 4.5 as the current small model. On September 22, 2026 Anthropic released Claude Opus 5.5, the first Claude 5.5 model, at $4 / $20 per million tokens, and Claude Sonnet 5.5 followed on September 28 at Sonnet 5's price of $2 / $10; Haiku 5.5 has not shipped yet. The decision logic in this guide transfers directly: a mid-tier default for daily work, the top tier only where reasoning depth pays for itself, the small model on volume. Your Projects and Skills carry forward automatically when you switch — nothing needs rebuilding.
Is Claude good for coding?
Yes — as of 2026, coding is arguably Claude's strongest suit, and it's the reason Claude Code became the default agentic coding tool for so many teams. The mid-tier model handles everyday engineering (features, refactors, debugging, review) well, and the top tier is genuinely strong on hard architectural work. For business owners the practical translation: Claude can build and maintain real internal tools, but treat AI-written code like work from a fast junior developer — it ships quicker than you expect and still needs review before you bet the business on it.
Is Claude Opus worth it over Sonnet?
For about 5% of business work, yes; for the rest, Sonnet is the better buy. As of September 27, 2026 Anthropic prices the newest Opus, Opus 5.5, at $4 per million input tokens and $20 per million output against Sonnet 5 at $2 / $10, so Opus costs 2x per token (2.5x on the older Opus 5 at $5 / $25), and on a Pro or Max subscription it also drains the same five-hour usage pool faster because the pricing page counts usage by model and conversation length, not by message. Opus earns that on strategy memos, contract review, multi-file refactors, and any one-shot task you would otherwise re-prompt five times. On drafting, summaries, triage, and anything scheduled, the quality difference is smaller than the bill difference.

More on AI for Business

ChatGPT Business vs Plus (2026): Which One a Small Business Should Actually Pay For

ChatGPT Plus vs Business, as of September 2026: Plus is $20 a month for one person and OpenAI may train on your chats unless you opt out; Business is $20 a seat billed annually or $25 monthly, two seats minimum, with no training by default, admin controls, shared projects for groups, and the Company Knowledge plugin. The side-by-side from OpenAI's pages, and three owner scenarios with a verdict each.

6 min

ChatGPT for Customer Service (2026): What Works, What Breaks, and What It Costs a Small Business

ChatGPT for customer service, as of September 2026: it drafts replies, answers policy questions from your own documents, and triages an inbox for $20 to $25 a seat a month on ChatGPT Business. It can't look up an order, issue a refund, or remember a customer unless you connect the system that holds them. The honest capability line, the plan you need, a setup in six steps, the failure modes, and the point where a real agent takes over.

7 min

ChatGPT Pulse (2026): What It Was, Why OpenAI Retired It, and How to Get the Morning Brief for Your Business

ChatGPT Pulse, the daily research cards OpenAI previewed for Pro users in September 2025, was sunset on June 17, 2026 with a 14-day wind-down. What it did, which plans ever had it, why scheduled tasks replaced it, and how an owner rebuilds the morning brief on a Plus or Business seat without handing ChatGPT the wrong things.

7 min

TRAIGA: What the Texas Responsible AI Governance Act Actually Requires of a Business (2026)

TRAIGA, the Texas Responsible AI Governance Act, took effect January 1, 2026. It applies to any business that operates in Texas, with no size exemption, but most of its duties bind state agencies. For a private company it is a short list of things you may not build AI to do on purpose, one disclosure duty for health care, and fines that start at $10,000. Here is the bill text in plain English.

9 min
Justin McKelvey, Fractional CTO and AI consultant in Austin, TX

Written by

Justin McKelvey

Fractional CTO & AI consultant in Austin, TX. 15 years building software, 50+ products shipped, $53M+ in client revenue generated. I help $1M–$50M founders ship production software and automate operations with AI, without hiring a full-time executive team.

Work with me

If this was useful, here are two ways I can help: