Articles/7 min read

Choosing your Claude model: the simple version

Choosing your Claude model: the simple version explains Fable, Opus, Sonnet and Haiku and gives a practical escalation ladder for cost and capability.

Worth Knowing?

✅Yes, if you have a zero idea of where your AI budget is going and have a willing staff member who can keep an eye on token use

❌No, if you have endless buckets of money to throw at AI or you know exactly what you need to use and when

01What Changed?

Claude now has four distinct models that match specific jobs instead of one size that tries to do everything. Fable, Opus, Sonnet and Haiku are the names and each has a clear role, price and context window. Fable and Opus share a 1,000,000 token context and are designed for heavy multi-step work and deep reasoning respectively. Sonnet also has a 1,000,000 token context and sits as the versatile daily default with an introductory price through 31 August 2026 ( personally use this as default). Haiku drops to a 200,000 token context with a 64,000 reply length and is built for high-volume quick tasks. Concrete figures matter because billing now reads like a taxi meter. Prices are listed per million tokens: Fable at $10 in and $50 out, Opus at $5 in and $25 out, Sonnet at $2 in and $10 out through 31 August 2026 before moving to $3 in and $15 out, and Haiku at $1 in and $5 out. Plans matter too because Fable requires a paid plan while Opus needs Pro and above and Sonnet and Haiku are available on Free and above.

02Why Should I Care?

Why should this model split change how you work with Claude? The answer is practical: choice saves money and reduces friction when results matter. You can stop overpaying for routine tasks and stop trusting lighter models with legal level reasoning when precision counts. You will notice cost differences fast when you run high-volume jobs because Haiku costs about a fifth of Sonnet for simple outputs and Sonnet remains the best low-friction option for most office work. Opus is the step up for documents that need cross-checking or careful logic and Fable is the reserve for projects that are long, multi-step or require handing a full project over to the model. You’ll also get better results more consistently because you pick for the task instead of defaulting to one model. That becomes the operational skill that saves money and preserves quality and it scales whether you’re a solo consultant or a small team of five.

03Who Benefits?

Who will notice the most immediate gains from this clarity? Solo founders, consultants, executive assistants and small agency teams will benefit first because they run a wide range of tasks and tight budgets. You as a business owner will save on repetitive work by using Haiku for tagging, short replies and high-volume lookups. You as an EA or VA will find Sonnet handles most daily workflows like drafting emails, research, and light automation and you can keep costs predictable on the Free plan (but limited). You as an analyst or technical consultant will reach for Opus when reasoning must hold up under scrutiny, and you’ll reserve Fable for large, multi-step projects that you'd rather hand off than micromanage.

04Is It Worth Using?

Is adopting the model ladder worth the learning curve? Yes when you treat it like a three-rule escalation and no if you ignore the pricing and context differences. The ladder is: Sonnet first, Haiku for volume, Opus for hard reasoning, and Fable for very large complex projects. You’ll get immediate ROI because Sonnet handles the majority of tasks and is available on the Free plan. Haiku saves money on repetitive volume work so your monthly token spend drops. Opus is costlier but worth it when the alternative is manual review by a human at a higher hourly rate. Fable is expensive but designed for handing off whole projects so it can replace external contractor fees on long efforts. You should test this over two weeks with three representative tasks: one routine task on Haiku, one multi-step workflow on Sonnet, and one complex reasoning task on Opus. Track time saved and token spend, and adjust the routing rule. That small experiment reveals whether the new model split returns value to your specific workflows.

05Pros

1. Clear task-to-model mapping that reduces guesswork when choosing which Claude to use. You’ll spend less time testing and more time executing. 2. Better cost control because Haiku covers high-volume simple work at a much lower rate while Sonnet covers daily needs on a Free tier so you can keep baseline costs small. 3. Predictable upgrade path with Opus and Fable so you only pay for advanced capabilities when a task genuinely needs them and you can justify the spend. 4. Large context windows (1,000,000 tokens) on three models so long documents, project histories and extended multi-step threads stay in memory without pulling together complicated workarounds. 5. Simpler operational rules for teams. You’ll be able to write a one-page SOP for model routing and onboarding new people will take minutes instead of hours.

06Cons

1. Pricing complexity increases because you now manage in and out token rates and a context window that’s separate from billing so tracking cost needs new monitoring practices. 2. Plan requirements add friction because Opus is locked to Pro and Fable is paid-only so some teams will need to upgrade plans to access higher tiers. 3. The ladder needs active governance. You’ll save money only if someone enforces routing rules and audits token spend on a weekly or monthly cadence. 4. Sonnet price moves after 31 August 2026 which changes baseline economics so any long-term budget should model both current and future rates. 5. Small teams might default to Opus for tasks Sonnet could handle which creates unnecessary expense unless you define clear escalation tests.

07Our Verdict

Get the most bang for buck by starting on Sonnet for everyday tasks and routing high-volume simple work to Haiku to save money. You should step up to Opus only when Sonnet fails a clear accuracy or reasoning test and reserve Fable for long, multi-step projects you want the model to manage end to end. One person will need to be assigned (quick, don't make eye contact) to own the routing rule and a simple spend dashboard that tracks in and out token usage. That governance step is cheap but decisive because it prevents wasted spend and preserves output quality. The approach works for basically anyone from solo operators to small agencies and avoids the trap of overpaying for capability you don’t require.

08FAQs

What exactly is a token and why does it matter?

A token is roughly three quarters of a word and acts like a taxi meter that measures input and output. You pay per million tokens so longer inputs or longer replies increase cost directly.

When should I use Opus instead of Sonnet?

Use Opus when Sonnet struggles with accuracy or complex cross-referencing that must hold up under scrutiny. Test a task on Sonnet and upgrade to Opus only if the result fails a specific correctness checklist.

Can I keep using Sonnet after 31 August 2026?

Yes you can keep using Sonnet but the price moves from $2 in and $10 out to $3 in and $15 out. Model availability and features remain, so update your budgets accordingly.

How should small teams route requests between the models?

Use a simple rule: Sonnet by default, Haiku for high-volume simple tasks, Opus for tasks that need careful reasoning, and Fable for large multi-step projects. Assign one owner to enforce the rule and review token spend weekly.

Does the large context window cost extra?

No the context window is separate from token billing. The context window is the amount Claude can hold in memory at once and you don’t pay extra for having a 1,000,000 token context beyond the model’s listed token rates.

I can help you map these four models to your specific workflows and set a two-week test plan so you see token spend and quality changes. Tell me your three most common tasks and team size and I’ll draft a routing SOP you can use immediately.

Work with me

More Updates

I'll tell you when something actually changes and whether it's worth your time.