In this article
- 1 The short answer, by the job in front of you
- 2 Sonnet vs Opus: start with Sonnet, and there are four models, not two
- 3 Why you keep hitting your limit
- 4 The control nobody mentions: Effort
- 5 My actual routing rule
- 6 When the heavy model is the wrong tool
- 7 Where Opus genuinely earns it
- 8 The test that settles it in ten minutes
- 9 What changes on Free, Pro and Max
- 10 How I checked this
- 11 So which one should you open?
Quick answer: Use Sonnet. Anthropic’s own guidance says that if you are not sure which model to pick, start there, and for writing, analysis and everyday multi-step work it is the right call.
Save Opus for problems you have already watched Sonnet struggle with. The question is not which model is smarter. It is which is the cheapest one that still gets your job right.
And the question itself is out of date. There are not two models. There are four.
That matters more than it sounds, because picking the heavy one by reflex is the single most common reason people hit a limit and think Claude is broken.
The short answer, by the job in front of you
| The job | Pick | Why |
|---|---|---|
| Summarize this, pull the dates out, quick lookup | Haiku 4.5 | Instant, and the lightest on your limit |
| Write it, analyze it, work through it, most things | Sonnet 5 | Anthropic’s stated default. Start here |
| You already tried Sonnet and it missed things | Opus 5 | Reasoning specialist. Costs more of your limit |
| Long project, many connected steps, few check-ins | Fable 5.1 | Heaviest, slowest, and on Pro it costs credits |
| You are hitting limits constantly | Go down a model | Not up. This is almost always the fix |
Sonnet vs Opus: start with Sonnet, and there are four models, not two
Whether you searched “Opus vs Sonnet” or “Sonnet vs Opus”, the answer is the same, and Anthropic states it outright: start with Sonnet.
The other half of the answer is that the two-model question is out of date. Here is Anthropic’s own table, read September 5, 2026.
| Model | Rate limit use | Anthropic’s stated best for |
|---|---|---|
| Haiku 4.5 | Lightest | “Quick answers, summaries, and simple extraction” |
| Sonnet 5 | Moderate | “Coding, writing, analysis, and multi-step workflows… your versatile default” |
| Opus 5 | Heavy | “Deep research and complex reasoning that genuinely needs sustained thinking” |
| Fable 5.1 | Heaviest | “Your largest, most critical projects: long, complex tasks” |
Read the middle column, not the right one. That is a price list.
Every ranking page treats this as a quality ranking where Fable is the good one and Haiku is the compromise. Anthropic does not describe it that way. It describes four tools with four costs, and it names Sonnet as the one to reach for when you have not thought about it.
There is a fifth Claude model, Mythos, built for cybersecurity and biology research. It is restricted to vetted organisations through Anthropic’s trusted access programmes, so it will not appear in your picker and it is not part of this decision.
Why you keep hitting your limit
This is the section I would send to most people instead of the rest of the page.
Anthropic says it directly: “if you use Opus or Fable on a task that Sonnet or Haiku could handle, you may be using more of your limit unnecessarily.”
Your limit is not a message count. It is a token budget running across two windows at once, a rolling five-hour session and a weekly cap.
A heavy model on a light task spends more of that budget for an answer that was not better. Anthropic does not publish how much more, and I am not going to invent a multiplier. The one number it does print is on the Effort control below.
The cleanest version of this I have seen came from a non-coder on r/claude. u/TeachMeThings3209067 wrote:
“Yeah I was using Opus 4.6 frequently and hitting my usage limits. Then I realised sonnet was actually good enough for what I wanted to do which was just regular analysis… I am literally just analysing transcripts from meetings.”
They fixed a limit problem by going down a model. Nothing else changed.
The original poster in that thread, u/LinkDaSquid, landed in the same place:
“I even asked Claude itself if I should use Opus instead of Sonnett, and it said to just keep using Sonnett. As a result, my weekly limit barely gets above like 15%.”
The control nobody mentions: Effort
Open your model picker and look under the model name. There is a second setting, and it changes your limit consumption on the model you already chose.
My own Claude picker, September 5, 2026. Note the warning label on Max. Anthropic prints the cost right there and almost no comparison page mentions this setting exists.
Five levels: Low, Medium, High as the default, Extra, and Max. The app flags Max with “1.5x or more usage” in orange, which is Anthropic telling you the price before you pay it.
That gives you a cheaper move than switching models. If Sonnet is close but not quite getting there, raise Effort before you jump to Opus. If Opus is working but draining you, drop Effort before you drop the model.
My actual routing rule
This is how the work gets assigned in my week. It is not a recommendation for your setup, it is what I do.
Fable is for complex skills. Editing my video skill, generating content for my website, running my SEO work. Anything long and genuinely complicated, that is Fable, hands down.
Opus is my everyday driver. Making shorts for YouTube, planning, thinking something through. When I want to plan something out, that is Opus. No reason to spend Fable’s budget on it.
Sonnet and Haiku I do not personally use much anymore. I want to be straight about that rather than pretend to a routing rule I do not run.
One caveat that matters if you are on Pro: I am on Max 20x. Running Fable and Opus as freely as I do is affordable at $200 and would not be at $20.
But the history is the useful part. Before Fable existed, my split was Opus and Sonnet: Opus did the planning and the main execution, and Sonnet did all the grunt work, the research, the fetch-and-summarize. That split was good. There are just so many models now that I stopped reaching for the light end.
If you are on Pro and watching your limit, that old split is better advice than what I currently do.
When the heavy model is the wrong tool
Here is the scar, and it is a cheap one to avoid.
Say you are writing a script. You are not going to use Fable for that. It might nail it, but it is burning tokens for no reason, and it is not even going to be faster.
It will probably be slower.
That is not a feel. Anthropic says the same thing in its own model guide: Fable “takes time to think through problems before answering, so responses take longer, and it uses the most of your rate limit.” You pay twice, in waiting and in allowance, for an answer a lighter model would have handed you.
One genuine exception worth knowing: for biology and security topics, Anthropic states that “Claude answers these topics with Opus even if you’ve picked Fable.” If that is your work, picking Opus yourself is simpler than being quietly rerouted.
Where Opus genuinely earns it
I have spent this page talking you down the ladder, so let me be fair to the top of it.
Opus is not Sonnet with a bigger bill. Anthropic describes it as built for problems that need sustained thinking over time, and its own worked example is analyzing complex research papers: long specialised documents, methodology critique, conclusions you will act on.
The best description of the gap came from u/roselan on r/claude:
“Everyone can make a good sandwich, but Opus is the 3 Michelin stars Chef.”
Which is exactly right, and also exactly why you should not order from him every day.
u/Far-Pomelo-1483 gave the compressed version: “Only use opus if sonnet can’t do it.”
The test that settles it in ten minutes
Do not take my routing rule or anybody else’s. Anthropic publishes a method and it is better than an opinion.
- Pick a task you already know the answer to. A report you have actually read, a document you wrote.
- Run it on the lighter model first.
- Start a fresh chat and run the identical prompt on the heavier one.
- Compare where the answers differ, not how long they are.
That last instruction is Anthropic’s own, and it is the part people get wrong. A longer answer feels better and usually is not. You are checking whether the lighter model missed anything you would have caught yourself.
If it did not miss anything, that task is Sonnet-shaped forever. Bank it and stop paying for the upgrade.
What changes on Free, Pro and Max
Model access is gated by plan, and one row surprises people.
| Plan | Haiku | Sonnet | Opus | Fable |
|---|---|---|---|---|
| Free | Yes | Yes | No | No |
| Pro, $20 | Yes | Yes | Yes | Usage credits only |
| Max 5x and 20x | Yes | Yes | Yes | Included, up to 50% of weekly limits |
| Team standard seat | Yes | Yes | Yes | Usage credits only |
| Team premium seat | Yes | Yes | Yes | Included, up to 50% of weekly limits |
The Fable row is the one to notice. On Pro you can use it, but it is not inside your plan limits, so it bills separately on top of the $20. On Max it is included up to half your weekly allowance. Verified against Anthropic’s pricing page on September 5, 2026.
How I checked this
The model names, rate-limit ordering, stated use cases, the Effort guidance and the biology and security behavior all come from Anthropic’s model selection guide, read September 5, 2026. Plan gating and the Fable rows come from claude.com/pricing, read the same day. The picker screenshot is my own machine on that date.
I have not benchmarked these models against each other and I am not going to pretend otherwise. Anthropic publishes no speed or quality numbers for chat, so anything you read that calls this a measured contest is somebody’s impression dressed up as data. The routing rule above is mine and I have labeled it as mine.
The 60-second version, from my Shorts: Claude's 4 models in plain English (and the one you're wasting).
So which one should you open?
- You are not sure: Sonnet. That is Anthropic’s answer and it is the right one.
- You do not write code and your work is documents, analysis and writing: Sonnet, and you will rarely need more.
- Sonnet tried and visibly missed things: Opus, for that task specifically, not as your new default.
- You are hitting your limit constantly: go down a model or drop Effort before you spend anything.
- You are running one long complex build with few check-ins: Fable, and check whether your plan includes it before you start.
If the limits themselves are the real problem, Claude usage limits explained covers the two ceilings, and Claude Pro vs Max covers whether more headroom is worth buying. If you have not paid for Claude at all yet, and Opus and Fable are the reason you are considering it, start with is Claude Pro worth it. If you are new to the current lineup, Claude 5 explained covers the models themselves.
This post is part of Claude at Work, the hub for using Claude at your job without code.
Published September 5, 2026. Model lineup, rate-limit ordering, plan gating and the Effort control checked that day against Anthropic’s model guide and pricing page, both linked inline. Anthropic ships new models often, so check the picker before trusting any list of four.
Written by
Chris AlarconChris Alarcon builds Ship Lean: the boring Claude and AI setups that actually work, handed to people who don’t code. He runs his one-person operation on these systems, around a full-time job, and shares every workflow, prompt, and tool combo in public. Start with the 15 prompts he uses every day.
Work With Me
Setting this up for your team?
I do that. The right plan per seat, the boring setup that makes it stick, and the first three chores handed off. A 2-hour workshop or a done-with-you build, for teams of 2-50.
See how it works →The 15 AI prompts I actually use every day
I tested hundreds and kept the 15 that survived at my desk. You get them right away, then one short Ship Lean note every Tuesday.