Justin McKelvey
Fractional CTO · 15 years, 50+ products shipped
Which Claude Model Is Best for Business? Sonnet vs Opus vs Haiku (2026)
By Justin McKelvey, fractional CTO · 15+ years shipping products · Updated August 2026
Claude Sonnet 4.6 is the right default for roughly 90% of business workflows in 2026 — fast enough for real-time drafts, strong enough for serious reasoning, priced where it doesn't punish you for using it daily. Claude Opus 4.7 is reserved for the hardest reasoning tasks or one-shot heavy lifts (overkill and roughly 2.5x the cost per token for daily work). Claude Haiku is best for high-volume, simple tasks like classification, summaries, and batch drafts. Most businesses on Claude Team should set Sonnet 4.6 as the default and switch up to Opus only when a task actually needs it. Switching is two clicks. August 2026 update: the Claude 5 family has arrived — Fable 5, Opus 5, Sonnet 5 — and the same decision logic applies one generation up; details in the update section below.
Every week a client asks me some version of: "Which Claude model should I use?" The product page lists three, the chat dropdown has more, and the wrong default can quietly 5x your bill or starve a workflow that needed more horsepower. Here's the short, honest answer from someone who installs Claude into small businesses for a living.
Claude model comparison at a glance
| Model | Speed | Cost (per 1M tokens) | Context window | Best for |
|---|---|---|---|---|
| Claude Sonnet 4.6 | Fast (real-time) | $3 in / $15 out | 200K (1M on Enterprise) | Daily business work — drafts, analysis, client emails, light coding, the 90% default |
| Claude Opus 4.7 | Slower (thinks longer) | $5 in / $25 out | 200K (1M on Enterprise) | Hard reasoning, deep research, complex code refactors, one-shot heavy lifts |
| Claude Haiku 4.5 | Fastest | $1 in / $5 out | 200K | High-volume, simple tasks — classification, summaries, batch drafts, real-time chat |
Correction (September 10, 2026), re-read from Anthropic's pricing page today: the Opus rate above is $5 / $25 per million tokens, not the $15 / $75 an earlier version of this table carried (that was the retired Opus 4.1 rate); Haiku is now Haiku 4.5 at $1 / $5 (Haiku 3.5's $0.80 / $4 is retired); and the Claude 5 generation prices at Sonnet 5 $2 / $10, Opus 5 $5 / $25, Haiku 4.5 $1 / $5, with the new top tier, Fable 5.1, at $10 / $50. Cache hits are one tenth of the input price on every model. The ratios in the sections below were corrected the same day: Opus is about 2.5x Sonnet per token, not 5x, and Haiku is a fifth of Opus, not a twentieth.
Three models, three different jobs. Most teams quietly use whichever one is on top of the dropdown that week. That's the bug.
August 2026 update: the Claude 5 family is here — same logic, new names
Anthropic's Claude 5 generation began rolling out in 2026, and it adds something the 4.x lineup never had: a tier above Opus. The new flagship is Claude Fable 5 — Anthropic calls the tier "Mythos-class" — the most capable generally available Claude model, with Opus 5 and Sonnet 5 as the generational successors to the models in the table above, and Haiku 4.5 as the current small model. (There is also a Claude Mythos 5 — the same underlying model as Fable 5, offered to approved organizations without the additional dual-use safety measures — so for a business buyer, the top of the menu is Fable 5.)
Here's the part that matters for your bill and your defaults: nothing about the decision logic changes — it just shifts up one generation. The mid-tier model (now Sonnet 5) stays the right default for daily business work. The heavyweight (Opus 5, or Fable 5 where your plan offers it) stays reserved for tasks where reasoning depth is the product — strategy, hard analysis, one-shot heavy lifts. The small model stays on volume work. If your workspace defaults are set the way this guide describes, a generation upgrade is a dropdown change, not a re-architecture — Projects and Skills carry forward untouched.
Two practical notes from the first months of the 5 family. First, check current per-token rates on Anthropic's pricing page before moving any high-volume automation up a generation — the tier logic transfers, but exact numbers move with each release, and the cost trap described below gets more expensive one tier up. Second, resist the new-flagship reflex: the arrival of a tier above Opus doesn't mean your email triage suddenly needs it. The 90% of business work that ran perfectly on the mid-tier yesterday still runs perfectly on the mid-tier today.
Sonnet 4.6: the right default for most business work
Sonnet 4.6 is the model you should set as your default in claude.ai, in Cowork, and in any Skill you build. Fast enough to feel real-time when you're drafting an email. Less than half the cost of Opus per token (Sonnet 5 is $2 / $10 per million against Opus 5's $5 / $25, as of September 2026). And the reasoning quality is genuinely close to Opus for most business tasks — proposals, newsletters, client deliverables, meeting summaries, lead scoring, basic code review. You will not feel the difference on the vast majority of work.
The places Sonnet 4.6 quietly shines: long-form writing with brand voice (especially when paired with a loaded Project), structured outputs like tables and JSON, multi-step reasoning that doesn't require novel research, and any workflow you run more than once a day. If you've been defaulting to Opus "just to be safe," you're paying 2.5x for a difference you can't measure on most tasks.
Opus 4.7: when the harder model is actually worth it
Opus 4.7 is the model you switch to when reasoning quality matters more than speed or cost. It thinks longer, holds more nuance, handles ambiguity better. For the right job, it's worth every cent. For the wrong job, it's a tax on your subscription.
The tasks where Opus actually earns its price tag: deep research where you're asking the model to weigh competing arguments, complex code refactors across multiple files, strategy documents where you want it to push back and find holes in your logic, legal or contract analysis where missing a clause matters, and any "one-shot" lift where you'd rather pay 2.5x once than re-prompt five times.
I use Opus for maybe 5% of my workflow: pricing strategy for a new offer, a board memo, a scope document for a six-figure engagement. Everything else stays on Sonnet. One pattern worth stealing: draft in Sonnet, then ask Opus to critique. You get most of Opus's value at a fraction of the cost because the critique is shorter than the draft.
Haiku: when "good enough" is actually the right answer
Haiku is the model nobody talks about, and it's the one that quietly saves businesses the most money on volume workflows. Fastest model Anthropic ships, half the cost of Sonnet 5 per token (Haiku 4.5 is $1 / $5 per million), and handles a surprising range of tasks well.
Use Haiku when you're doing something simple at high volume: classifying inbound leads as hot/warm/cold, summarizing call transcripts, drafting first-pass replies to common support tickets, tagging content, extracting structured data from emails, generating product description variants. It's also the right model for real-time chat experiences — customer-facing bots, internal lookup tools, anything where latency matters more than nuance.
The trap: don't use Haiku for anything that requires real reasoning. It's great at "do this simple thing fast," not "think through this carefully." If your output quality drops, jump back to Sonnet.
Cost comparison on a typical business workflow
Real numbers on a workflow most businesses have — triaging 500 inbound emails per week and drafting first-pass replies:
- Haiku 4.5: roughly $2–$5/week. Near-instant. Quality is fine for templated replies and basic triage.
- Sonnet 4.6 or Sonnet 5: roughly $6–$15/week. Noticeably better drafts that need less editing. Worth it if you're sending replies directly to clients.
- Opus 4.7 or Opus 5: roughly $15–$38/week at the $5 / $25 rate (corrected September 2026 from an earlier $40–$75 estimate built on the retired $15 / $75 rate). Marginally better drafts that almost nobody can tell apart from Sonnet's on this task. Not worth it for inbox triage.
That's a 5x cost difference between Haiku and Opus on the same workflow, for an output difference your customers cannot perceive.
How to actually switch models (in claude.ai, Cowork, and Skills)
The mechanics are straightforward once you know where to look.
In claude.ai: there's a model dropdown above the message input. The default is whatever was selected last, which is why most people drift into whichever model was used most recently rather than the one that fits the task.
Inside Claude Cowork: each Skill can specify its own model. When you build a Skill, there's a model selector in the Skill settings. Set Haiku for high-volume classification skills, Sonnet for daily drafting skills, Opus for the rare reasoning-heavy ones. This is where most of the cost optimization lives.
Via the API: the model is set per-request via the `model` parameter (e.g. `claude-sonnet-4-6`, `claude-opus-4-7`, `claude-haiku-4`). Set this once at the workflow level rather than guessing per-call. Inside a Project, the model selection carries through the conversation — you can switch mid-thread to, say, draft with Sonnet and then hit Opus for a critique pass on the same context.
When NOT to use Opus (the cost trap that catches almost everyone)
The single most common mistake I see when I audit a client's Claude setup: Opus is the default model on every workflow, and nobody remembers turning it on. This is the easiest way to 2.5x your Claude bill without realizing.
Don't use Opus 4.7 for: short questions you could answer in your head, formatting and rewriting tasks, summaries under a page, anything you're running on a schedule (cron jobs, automated reports, batch processing), real-time chat with customers, or "warm-up" prompts where you're still figuring out what you want.
The pattern that costs the most money: someone sets up a daily automation, leaves it on Opus by default, and it runs 30 times a month for six months before anyone audits the bill. The fix is always the same — downgrade to Sonnet, save 60%, see no quality drop.
The practical decision guide
Five branches that cover most business decisions:
- If you're drafting client-facing copy (emails, proposals, newsletters): Sonnet 4.6.
- If you're doing daily analysis, summaries, or "rewrite this in a friendlier tone": Sonnet 4.6.
- If you're running a high-volume automation (classification, batch drafts, support triage): Haiku.
- If you're making a real decision (strategy, pricing, board memo, contract review): Opus 4.7 for the heavy reasoning pass, Sonnet for the rest.
- If you have no idea and just want to get unstuck: Sonnet 4.6. It is the right default 90% of the time.
This is the exact decision tree I install with clients when I set up Claude for their business — default Sonnet at the workspace level, Haiku on the high-volume Skills, Opus on the 2–3 Skills where reasoning depth actually moves the needle.
Set the right default and stop thinking about it
The honest takeaway: 90% of small businesses should set Sonnet 4.6 as the default everywhere — in claude.ai, in Cowork, in every Skill — and only switch up to Opus or down to Haiku when a specific task asks for it. Most people overpay for Opus they don't need, or underuse Haiku on volume work that would save them real money.
This is the same decision tree I configure when I install Claude for Small Business for clients. Default Sonnet, Haiku on the high-volume Skills, Opus on the 2–3 workflows where reasoning depth actually moves the needle. The setup takes an afternoon. The savings show up in the next billing cycle.
If you want the broader cost picture, I wrote a full Claude AI pricing breakdown — and the verdict on whether the $20 subscription itself is justified lives in Is Claude Pro worth it. Still deciding between Claude and another tool? The Claude vs ChatGPT for small business comparison is the head-to-head. Or skip the trial and error — my done-with-you Claude install ships with the model defaults already configured per Skill, so your business gets the right model on the right task from day one.
Claude Opus vs Sonnet: the math the pricing page does not do for you
Anthropic's pricing page gives you two numbers per model and stops. As of September 10, 2026 those numbers are Sonnet 5 at $2 per million input tokens and $10 per million output, Opus 5 at $5 / $25 (Opus 4.7 is priced the same), Haiku 4.5 at $1 / $5, and the new top tier, Fable 5.1, at $10 / $50. What the page does not do is the arithmetic that decides Opus vs Sonnet for a business. Three pieces of it. First, the ratio is 2.5x, not the 5x of the Opus 4.1 era, so the "just use Opus" tax got smaller this year; on a workflow that costs $10 a week on Sonnet, Opus is $25, not $50. Second, on a subscription you do not pay per token at all, but the same choice still costs you: the pricing page says every plan meters usage on a rolling five-hour window, that "how much you can do depends on the length and complexity of your conversations, the model you choose, and the features you use," and that there is no fixed message count. Opus on Pro drains the window faster than Sonnet does, so the model dropdown is a limits decision even when the bill is flat. Third, cache hits are one tenth of the input price on every model, so a Project with a long, stable system prompt makes the Opus premium smaller in practice than the list rate suggests. Fable is not on the Free plan and is metered separately on Pro and Max per the plan table, so it is not the default anywhere. The decision rule survives the new numbers: Sonnet as the default, Opus on the two or three Skills where reasoning depth pays for itself, Haiku on volume.
Get the Free AI Content Toolkit
The exact system I use to turn one idea into a month of content — atomization framework, voice template, prompt library, weekly system.
Frequently Asked Questions
- Which Claude model is fastest?
- Claude Haiku is the fastest, with near-instant response times even on long inputs. Sonnet 4.6 is fast enough to feel real-time for chat and drafting. Opus 4.7 is the slowest because it spends more time reasoning — expect 2-4x the latency of Sonnet on the same prompt.
- Which Claude model is cheapest?
- Claude Haiku 4.5 is the cheapest at $1 per million input tokens and $5 per million output tokens as of September 2026 — half the cost of Sonnet 5 ($2 in / $10 out), a third of Sonnet 4.6 ($3 / $15), and a fifth of Opus 5 or Opus 4.7 ($5 in / $25 out). For high-volume simple tasks, Haiku is the obvious pick. For complex work, the cost difference is worth paying for the better model. (Corrected September 10, 2026 from Anthropic's pricing page: an earlier version of this answer carried the retired Opus 4.1 rate of $15 / $75 and the retired Haiku 3.5 rate of $0.80 / $4.)
- Which Claude model is best for writing?
- Sonnet 4.6 is the best Claude model for most business writing — proposals, newsletters, sales emails, client deliverables. The writing quality is close enough to Opus that you can't reliably tell the drafts apart, and the speed makes it usable for daily work. Opus 4.7 is worth switching to for long-form pieces where you want the model to push back on your thinking, like strategy memos or pricing arguments.
- Which Claude model is best for coding?
- Sonnet 4.6 is the default model inside Claude Code and handles the vast majority of coding work well — new features, refactors, debugging, code review. Opus 4.7 is worth switching to for hard architectural decisions, multi-file refactors that cross system boundaries, or when you've been stuck on a bug for an hour. Haiku is fine for small scripts and one-liners but underpowered for real engineering work.
- Can I use multiple Claude models in the same workflow?
- Yes, and it's the smartest way to use Claude. A common pattern: draft with Sonnet 4.6, ask Opus 4.7 to critique the draft, then have Haiku format the output. Each model does the part it's best at, and the total cost is lower than running everything through Opus. Inside Cowork, each Skill can specify its own model, so you can mix and match across a single business workflow without thinking about it.
- What is the Claude 5 family and does it change this advice?
- The Claude 5 family began rolling out in 2026: Claude Fable 5 is the new top tier — a Mythos-class model that sits above Opus in capability — alongside Opus 5 and Sonnet 5, with Haiku 4.5 as the current small model. The decision logic in this guide transfers directly: a mid-tier default for daily work, the top tier only where reasoning depth pays for itself, the small model on volume. Your Projects and Skills carry forward automatically when you switch — nothing needs rebuilding.
- Is Claude good for coding?
- Yes — as of 2026, coding is arguably Claude's strongest suit, and it's the reason Claude Code became the default agentic coding tool for so many teams. The mid-tier model handles everyday engineering (features, refactors, debugging, review) well, and the top tier is genuinely strong on hard architectural work. For business owners the practical translation: Claude can build and maintain real internal tools, but treat AI-written code like work from a fast junior developer — it ships quicker than you expect and still needs review before you bet the business on it.
- Is Claude Opus worth it over Sonnet?
- For about 5% of business work, yes; for the rest, Sonnet is the better buy. As of September 2026 Anthropic prices Opus 5 at $5 per million input tokens and $25 per million output against Sonnet 5 at $2 / $10, so Opus costs 2.5x per token, and on a Pro or Max subscription it also drains the same five-hour usage pool faster because the pricing page counts usage by model and conversation length, not by message. Opus earns that on strategy memos, contract review, multi-file refactors, and any one-shot task you would otherwise re-prompt five times. On drafting, summaries, triage, and anything scheduled, the quality difference is smaller than the bill difference.
More on AI for Business
GPT-6 Astra (2026): What It Costs, Who Actually Gets It, and GPT-6 vs Claude Opus 5 for a Business
GPT-6 Astra, as of September 2026: $10 per million input tokens and $50 output on OpenAI's API, twice Claude Opus 5 and the same as Claude Fable 5.1, with a 1,050,000-token window and a 272K line that reprices the whole request. On ChatGPT Plus it runs 5 to 45 Codex messages per five hours. Every number from OpenAI's and Anthropic's own pages, and what a business should actually do about it.
Gemini vs Copilot (2026): Which Is Better for a Small Business, What Each Costs, and Why the Answer Is Usually the Suite You Already Pay For
Gemini vs Copilot for a business, as of September 2026: Gemini if your company runs on Google Workspace, because it is included in the Business plans from $7 a user; Microsoft 365 Copilot if you run on Microsoft 365 and will pay $21 to $30 a seat on top of it for the version that reads your files, mail, and meetings. The consumer apps are free at both. Prices from Google's pricing page and the Microsoft canon on this site, plus the third option most owner-led companies end up choosing.
Is DeepSeek Safe? (2026): What Its Privacy Policy Actually Says, Where Your Data Goes, and the Two Ways to Use It Anyway
Is DeepSeek safe? For anything you would not post publicly, no, and the reason is in DeepSeek's own privacy policy: your prompts, device data, and IP are stored in the People's Republic of China by a Hangzhou-registered company and used to train its models unless you opt out. The model itself is a different question, because the weights are MIT-licensed and can run on a host you choose. Here is the policy, read line by line, and the rule for a business.
Kimi vs Claude (2026): The $0 Chat, the $3 Flagship API, and the Terms That Decide It for a Business
Kimi vs Claude for a business, as of September 2026: Kimi's free tier is the most generous chat in the category and its K3 flagship carries a 1M-token window at $3 in / $15 out per million tokens; Claude's Sonnet 5 is $2 / $10 with a 200k window, and its Team plan ships no-training terms and admin controls at $20 a seat. On price the two are closer than the hype says. The decision is the terms, the workspace, and what you put in the box.
Written by
Justin McKelvey
Fractional CTO & AI consultant in Austin, TX. 15 years building software, 50+ products shipped, $53M+ in client revenue generated. I help $1M–$50M founders ship production software and automate operations with AI — without hiring a full-time executive team.
Work with meIf this was useful, here are two ways I can help:
Before you go
Score your AI readiness in 3 minutes.
Free, no call. 30 yes/no questions show where AI actually pays off in your business — and where it doesn't yet.
Get my readiness score →