Learn

Claude Opus 5 vs Fable 5.1 vs Sonnet 5: Which Should You Use?

Choose Sonnet 5 for fast, economical daily work, Opus 5 for demanding agentic and professional tasks, and Fable 5.1 only when measured capability gains justify higher cost, retention rules, and stricter safeguards.

Start with the selection criteria. Use this page when you know the category and need a practical framework for narrowing the field.

UpdatedSeptember 3, 2026
Browse tool profiles

Editorial guide

Guide

Start with the criteria, tradeoffs, and shortlist logic before you open individual tools.

The practical Claude choice is no longer a three-way contest between Fable 5, Opus 4.8, and Mythos 5. For normal app, Claude Code, and API buyers, the current decision is Sonnet 5 versus Opus 5 versus Fable 5.1. Opus 4.8 matters as a migration baseline, while Mythos 5.1 remains an access-gated route rather than a self-serve option.

The short answer is to start with Sonnet 5 when responsiveness and volume dominate, move to Opus 5 when difficult work makes failed attempts or extra human review expensive, and reserve Fable 5.1 for workloads where a matched evaluation proves that its additional capability is worth the higher price and stricter operating boundary. Anthropic's own model guide starts complex agentic coding and enterprise work on Opus 5, but that positioning is a starting hypothesis, not proof that one model wins every workload.

Start with the job, not the model name

Sonnet 5 is the throughput choice. Anthropic positions it as the best combination of speed and intelligence, lists it as fast relative to the other current models, and prices it below Opus 5. It is the sensible baseline for interactive assistants, routine coding, high-volume extraction, and agent loops where latency or token spend compounds across many calls.

Opus 5 is the quality-first daily model. Anthropic positions it for complex agentic coding and enterprise work, while keeping its standard API price at the Opus 4.8 level. It is the stronger starting point when a task is ambiguous, long-running, tool-heavy, or expensive to review, but it should still be compared with Sonnet 5 on the actual work.

Fable 5.1 is the capability-ceiling route. Anthropic describes it as its most capable widely released model and targets long-running agents, but it costs twice Opus 5 per input and output token. It also carries model-specific retention requirements and stricter safeguards. That makes Fable a specialized escalation, not the automatic premium default.

Anthropic publishes benchmark and safety evaluations for all three models. Those results are useful vendor evidence about intended positioning, but they are not independent comparative testing. Use them to form an evaluation plan, not to declare a universal benchmark winner.

Specs that change the decision

Decision factor

Sonnet 5

Opus 5

Fable 5.1

Official positioning

Best combination of speed and intelligence

Complex agentic coding and enterprise work

Next-generation intelligence for long-running agents

Standard API price

US$2 input / US$10 output per million tokens

US$5 input / US$25 output per million tokens

US$10 input / US$50 output per million tokens

Cache reads per million tokens

US$0.20 per million cached input tokens

US$0.50 per million cached input tokens

US$0.25 per million cached input tokens

Comparative latency

Fast

Moderate

Slower

Context window

1 million tokens

1 million tokens

1 million tokens

Maximum output

128,000 tokens

128,000 tokens

128,000 tokens

Adaptive thinking

Supported; on by default

Supported; on by default

Always on

Retention boundary

Standard platform controls; not designated a Covered Model

No model-specific retention requirement for general access

Normally 30-day retention; explicitly authorized eligible customers may use transitional ZDR

The shared context and output limits remove a common but misleading selection shortcut: choosing the more expensive model does not buy a larger advertised window among these three. The real differences are capability profile, latency orientation, token price, effort behavior, access, and governance.

Per-token price is also not the same as cost per completed job. A slower or more expensive model can be economical if it needs fewer retries, fewer tool calls, or less review. A cheaper model can become costly if a long agent loop repeatedly fails. Measure input tokens, output tokens, elapsed time, retries, tool calls, and human correction together.

Sonnet 5 now has standard US$2 input / US$10 output pricing per million tokens; the scheduled September increase was cancelled. Fable 5.1 retains Fable 5’s US$10/US$50 base rates but lowers cache reads from US$1 to US$0.25. That cache category is cheaper than Opus 5’s US$0.50, although Fable’s uncached input and output remain twice Opus. Compare the actual token mix, not only the base-price ratio.

Choose Sonnet 5 for speed and scale

Choose Sonnet 5 when most tasks are repeatable, latency-sensitive, or high-volume. It is the first model to try for customer-facing chat, broad internal assistants, routine coding changes, extraction pipelines, and agents whose economics depend on many inexpensive steps. Its lower standard price leaves more room for retries, caching, or parallel attempts.

Sonnet 5 is also the easiest family baseline for evaluating whether more capability changes the outcome. Run it at the same effort level and with the same tools as Opus 5. If it meets the quality threshold, the faster latency orientation and lower price usually make the selection straightforward.

Do not promote Sonnet solely because a vendor benchmark approaches a premium model. Match the task, tools and effort, then compare accepted results under the current standard US$2/US$10 rate. There is no longer a separate scheduled higher Sonnet rate to budget for.

Choose Opus 5 for difficult everyday work

Choose Opus 5 when the task is difficult enough that judgment, persistence, or self-correction matters more than the lowest token rate. Complex repository work, multi-step research, professional document analysis, and enterprise agents with costly failure modes are better candidates than routine transformations.

Opus 5 is the middle route: half Fable 5.1’s base input/output price, moderate rather than slower comparative latency, and no Fable-specific retention requirement. Anthropic recommends starting most workloads on Opus 5. This is a workload starting point, not a claim that Fable cannot be bought on Pro or that all app and Code defaults match.

For latency-sensitive premium work, Anthropic offers Opus 5 Fast mode at about 2.5 times the default speed. That route costs twice the base Opus API price and can consume usage credits in Claude Code, so compare it with Sonnet 5 rather than treating speed as a free toggle.

Choose Fable 5.1 only for measured headroom

Choose Fable 5.1 when long-horizon autonomy, unusually difficult reasoning, or frontier-level capability is the primary constraint and a representative evaluation shows a material gain over Opus 5. The evaluation should use the same prompt, tools, effort, stopping conditions, and review rubric for both models.

Fable 5.1’s standard US$10 input / US$50 output rates are twice Opus 5’s US$5/US$25, but its US$0.25 cache-read rate is below Opus’s US$0.50. Cache writes, uncached input, output, retries and review work all affect total cost. Prefer Fable only when that complete task economics and quality assessment supports it.

Fable normally requires 30-day retention. Eligible customers explicitly authorized by Anthropic can use transitional ZDR while Enterprise Frontier Safeguards rolls out; confirm the agreement rather than assuming every account qualifies. EFS is planned in phases beginning later in the fall; it is not already universally deployed. Privacy settings for consumer chat are not an API ZDR agreement.

Stricter safety classifiers can also refuse or reroute some requests. In Claude surfaces, Anthropic may fall back from a flagged Fable request to another model; API users need to handle refusal details or configure an eligible fallback path. For cybersecurity, biology, or other sensitive workflows, confirm which model actually answers and whether the resulting behavior remains acceptable.

Effort changes the model you are buying

Fable 5.1, Opus 5, and Sonnet 5 all support the low, medium, high, xhigh, and max effort levels. High is the API default. Effort affects thinking, visible output, and tool calls, so it changes both capability and cost; it is a behavioral signal rather than a strict token budget.

Start a family comparison at high effort, then sweep downward or upward only where the workload justifies it. Sonnet at high may beat Opus at low for one task, while Opus at high may avoid retries that make Sonnet more expensive in practice. A model comparison without matched effort is partly an effort comparison.

Fable's adaptive thinking is always on. Opus 5 and Sonnet 5 also use adaptive thinking by default, but Opus 5 cannot disable thinking at xhigh or max. Applications that hard-code older thinking settings need a compatibility check before a model change.

Max effort is not a default buying recommendation. It allows unconstrained token spending for the deepest work. Use it only when a measured quality gain matters more than the added latency and cost, and set output limits with enough room for thinking and tool use.

Claude app, Claude Code, API, and cloud routes

Model access and billing depend on the surface and account. Fable 5 and 5.1 share up to 50% of regular weekly usage on Max and Premium seats; Pro and Standard seats use paid usage credits from the start. This is not an extra allowance for each model. The earlier Fable 5 promotion did not include Fable 5.1 and is not renewed by this release. Direct API use and usage-based Enterprise follow their own metered rates.

Claude Code's model selector supports Sonnet 5, Opus 5, Fable 5.1, and Opus 4.8. Use the /model menu or a model flag to make the choice explicit, then check /status so an evaluation does not accidentally run against a different model or billing route. Fable 5.1 requires Claude Code 2.1.255 or later; the fable alias points to it unless overridden, but neither Fable model is an account-type default.

A paid Claude app plan does not include direct Claude API access. Console and API usage are billed separately, and Claude Code can use either a supported subscription allocation or API credits depending on authentication and account state. Keep app entitlement, Claude Code usage, and direct API spend as separate budget lines.

For developers, all three current models are available through the Claude API, Claude in Amazon Bedrock, Claude Platform on AWS, Google Cloud, and Microsoft Foundry. Model IDs, feature support, fallback behavior, data processing, and regional availability can differ by route. Anthropic is the processor for its native API, Claude Platform on AWS, and Claude in Microsoft Foundry; cloud-provider terms govern Amazon Bedrock and Google Cloud processing.

Opus 4.8 and Mythos 5.1 are boundaries, not peer choices

Keep Opus 4.8 as a compatibility baseline for existing integrations rather than the current default comparison. Fable 5.1 needs its own API migration check: use claude-fable-5-1 and remove forced tool choices. Older models cannot read its thinking blocks; on a model switch the API drops incompatible blocks and the older model re-plans. Separately, editing earlier turns can invalidate thinking signatures and produce a 400; default enforcement differs for accounts created before versus on/after August 31, 2026. Test the integration without implying that every first-party Claude Code user must rewrite Messages history.

Mythos 5.1 shares Fable 5.1's specifications and pricing but is not a general buyer option. Anthropic limits it to approved Project Glasswing customers for defensive cybersecurity workflows, with no self-serve signup. Unless an organization already has that access path, compare Sonnet 5, Opus 5, and Fable 5.1 instead.

Run a workload evaluation before switching

Build a small evaluation set from work that is valuable, difficult, and representative. Include routine tasks that expose needless premium spend, difficult tasks that expose quality differences, and sensitive tasks that can trigger safeguards or governance constraints.

For each candidate, keep prompts, tools, context, effort and stopping rules matched. Record accepted output, human corrections, retries, tool calls, token mix, elapsed time and fallback events. Use current standard Sonnet pricing and each model’s distinct cache-read/write rates; do not revive the cancelled Sonnet price increase.

Choose the least expensive route that reliably clears the quality, latency, and governance thresholds. Start with Sonnet 5 for throughput work, Opus 5 for difficult daily work, and Fable 5.1 only where its measured headroom pays for the higher price and operating constraints. Re-run the evaluation when prompts, tools, effort settings, or cloud routes change.

Evidence boundary

Official sources

Editorial guidance grounded in official product sources.

FAQ

Common questions

Which Claude 5 model should most people start with?

Start with Sonnet 5 for latency-sensitive, repeatable, or high-volume work. Start with Opus 5 when complex agentic coding, difficult analysis, or expensive review makes quality the primary constraint. Anthropic itself points complex agentic and enterprise users to Opus 5; Fable 5.1 is better treated as an escalation that must prove extra value on a matched workload evaluation.

Is Fable 5.1 worth twice Opus 5's API price?

Only a representative workload can establish value. Fable 5.1 has twice Opus 5’s standard input/output rates, but its cache reads cost US$0.25 versus Opus’s US$0.50 per million tokens. Compare total task cost, successful completion and review effort; vendor benchmarks do not guarantee the result.

Do Fable 5.1, Opus 5, and Sonnet 5 have different context limits?

No. Anthropic lists a 1-million-token context window and a 128,000-token maximum output for all three current models. Context size therefore does not decide among them; capability, latency, effort, price, access, retention, and safeguard behavior do.

Can I use Sonnet 5, Opus 5, and Fable 5.1 in Claude Code?

Yes, subject to the account and provider. Fable 5.1 requires Claude Code 2.1.255 or later and is selected through the fable alias or full model ID. Neither Fable model is an account-type default; check the actual model, subscription allowance and usage-credit/API route.

Can eligible ZDR customers use Fable 5.1?

Fable normally requires 30-day retention. Eligible customers explicitly authorized by Anthropic can use transitional ZDR while Enterprise Frontier Safeguards rolls out; confirm the agreement rather than assuming every account qualifies.

Should Opus 4.8 or Mythos 5.1 still be part of this choice?

Opus 4.8 belongs in an upgrade and compatibility decision because Opus 5 changes default thinking behavior while keeping the same standard token price. Mythos 5.1 shares Fable 5.1's specifications and pricing but is invitation-only for approved Project Glasswing customers and has no self-serve signup. Neither should displace the current Sonnet 5, Opus 5, and Fable 5.1 selection.

Next steps

Take the next evaluation step

Use these next pages to evaluate the strongest candidates, supporting profiles, or follow-up guides against the selection criteria.

View all tools