Claude models in 2026: Opus, Sonnet, Haiku, or Fable

Model choice is the largest single lever on both cost and quota consumption, and most people set it once and never touch it again. The default is not always the right answer, in either direction.

By the Continuum team. We build a workbench that runs Claude Code, Codex, and their peers, so the model rates quoted here are the ones our own cost analytics ship with.

The short version

The Claude Code default depends on the account, not on one universal pick. Max, Team premium, Enterprise pay-as-you-go, and the Anthropic API start on Opus 5. Pro, Team standard, and Enterprise subscription seats start on Sonnet 5. Fable 5 is never the default. On Pro and Team standard it bills usage credits from the first token. On Max and Team premium it can use up to 50 percent of the same weekly pool, not extra. Switch per task with /model.

What you need to know
  • Default is account-specific. Max and API start on Opus 5. Pro and Team standard start on Sonnet 5. Fable 5 is never the default.
  • Haiku 4.5 is genuinely capable on mechanical edits at $1 and $5 per million.
  • Opus 5 is $5 and $25 per million. It earns that on reasoning, not on volume.
  • Fable 5 on Pro and Team standard is credits from the first token. On Max it is up to 50 percent of the same weekly pool.
  • Aliases resolve differently per provider. opus is not the same model everywhere.
  • opusplan plans on Opus and executes on Sonnet, which is the cheapest good habit here.

Claude models: the August 2026 lineup

Current Claude models from Anthropic's models overview, opened 19 August 2026. Input and output are Claude API list rates per million tokens. Reliable knowledge cutoff is the date through which the model's knowledge is most extensive; training cutoff can be later. Haiku 4.5 is reliable through Feb 2025 and trains through Jul 2025.
ModelAPI IDInput / outputContextMax outReliable cutoffTrain cutoffThinkingEffort
Fable 5claude-fable-5$10 / $501M128kJan 2026Jan 2026Adaptive, always on. Extended thinking: noYes, default high
Claude Opus 5claude-opus-5$5 / $251M128kMay 2026May 2026Adaptive: yes. Extended thinking: noYes, default high
Sonnet 5claude-sonnet-5$2 / $101M128kJan 2026Jan 2026Adaptive: yes. Extended thinking: noYes, default high
Haiku 4.5claude-haiku-4-5-20251001 (alias claude-haiku-4-5)$1 / $5200k64kFeb 2025Jul 2025Extended thinking: yes. Adaptive: noNone
Cloud IDs from the same models overview, opened 19 August 2026. Copied from the live Copy model ID controls. Claude Platform on AWS uses the Claude API IDs, not the Bedrock-style IDs.
ModelAWS Bedrock IDGoogle Cloud ID
Fable 5anthropic.claude-fable-5claude-fable-5
Opus 5anthropic.claude-opus-5claude-opus-5
Sonnet 5anthropic.claude-sonnet-5claude-sonnet-5
Haiku 4.5anthropic.claude-haiku-4-5-20251001-v1:0claude-haiku-4-5@20251001

Claude Opus 5 is the coding flagship at $5 input and $25 output per million on the API. Fable 5 sits above it for runs larger than one sitting. How Fable is billed on a Claude plan is in the table below, from Anthropic's Fable plan article. The Fable API table is also on Claude Code pricing. Session and weekly windows shared across models live on Claude Code limits. How those Claude models sit against GPT-5.6 and Gemini is on best coding models.

Prompt-cache rates per million tokens from Anthropic pricing, opened 19 August 2026. Multipliers: 5-minute write 1.25x base input, 1-hour write 2x, cache hit 0.1x.
ModelBase input5m cache write1h cache writeCache hitOutput
Fable 5$10$12.50$20$1$50
Opus 5$5$6.25$10$0.50$25
Sonnet 5$2$2.50$4$0.20$10
Haiku 4.5$1$1.25$2$0.10$5
Fast mode, research preview, from the same pricing page opened 19 August 2026. First-party Claude API only. Not on Claude Platform on AWS or partner clouds. Not with the Batch API. Stacks with cache and residency multipliers.
ModelFast inputFast output
Opus 5$10$50
Opus 4.8$10$50
US-only inference meter from Anthropic pricing, opened 19 August 2026. For Claude 4.6 and later, inference_geo: "us" is 1.1x on every token category (input, output, cache write, cache read). Global is the default and stays at list price. Fast mode stacks on top of this meter. Partner Bedrock and Google Cloud regional prices are separate.
ModelUS-only input / output (1.1x)
Fable 5$11 / $55
Opus 5$5.50 / $27.50
Sonnet 5$2.20 / $11
  • Output is priced five times input on every current model, so the output column is where the difference actually bites.
  • Cache reads bill at 10 percent of the base input rate, which is why a long cached session is far cheaper than the input column alone suggests.
  • The Batch API halves both input and output, but it is asynchronous and irrelevant to interactive coding.
  • Haiku 4.5 does not support effort levels. Fable 5, Opus 5, and Sonnet 5 all do, and default to high.

Coding-agent cards with harness, DeepSWE cost, and when not to use it: Opus 5.5, Sonnet 5.5, Opus 5, and Fable 5.

What your account defaults to

The default alias is not one model. Claude Code documentation, opened 19 August 2026, maps it by account type. Fable 5 is never the default on any of these rows. Sessions use Fable only after you choose it with /model fable, a model setting, or the best alias where Fable is available.

Account-type defaults from Claude Code model configuration, opened 19 August 2026.
Account typeDefault model
Max, Team premium, Enterprise pay-as-you-go, Anthropic APIOpus 5
Claude Platform on AWS, Amazon Bedrock, Google Cloud Agent PlatformOpus 5
Pro, Team standard, Enterprise subscription seatsSonnet 5
Microsoft FoundrySonnet 4.5

Fable 5 on your Claude plan

A promotion that included Fable 5 in Pro weekly limits ended 19 July 2026 at 11:59:59 PM PT. Starting 20 July 2026, access depends on the plan. This table is from Anthropic's Fable plan article and claude.com/pricing, opened 19 August 2026. Claude Code needs v2.1.170 or later.

How Fable 5 bills after 19 July 2026, from the Anthropic support article Claude Fable 5 on your plan, opened 19 August 2026.
PlanIncluded in weekly limit?CapWhat happens at the capHow to keep working
FreeNot availableFable 5 is paid-plan onlyYou cannot start a Fable 5 sessionUpgrade to a paid plan
Max, Team premium, seat-based Enterprise premiumYesUp to 50% of the same weekly pool, not extraThe Fable 5 weekly slice is used upUsage credits, or switch to another model
Pro, Team standardNoCredits from the first request. Eligible seats get a one-time credit. Promo that included Fable in Pro weekly limits ended 19 Jul 2026 11:59:59 PM PTThere is no weekly Fable slice. Further Fable 5 uses creditsCredits, or upgrade to Max
Seat-based Enterprise standardNoCredits only if the org enabled them. No one-time promo creditFable 5 is blocked if credits are offAsk an admin to enable credits, or switch model
Usage-based Enterprise / Claude APIOwn meter, not the weekly-limit tableBilled at standard API ratesThe weekly-limit table does not applyKeep working on the API meter

Fable 5 is not available under zero data retention. The /model picker either omits it or shows it disabled. Source: Claude Code model configuration, opened 19 August 2026.

When Fable 5 bills usage credits, the /model picker shows "Requires usage credits" on the Fable 5 row. Interactive sessions get a consent prompt before that request bills. Members of Enterprise plans with organization billing do not see the prompt. After you choose to continue on Fable 5 using usage credits, the prompt does not return. In non-interactive mode with the -p flag, and through the Agent SDK, Claude Code never shows the consent prompt and bills without asking. Source: Claude Code model configuration, opened 19 August 2026.

Older Claude models still available

Still-available models from Anthropic's models overview, opened 19 August 2026. Prefer current models for new work. Bedrock and Google Cloud IDs from the same page.
ModelAPI IDAWS Bedrock IDGoogle Cloud IDInput / outputContextMax outReliable cutoff
Opus 4.8claude-opus-4-8anthropic.claude-opus-4-8claude-opus-4-8$5 / $251M128kJan 2026
Opus 4.7claude-opus-4-7anthropic.claude-opus-4-7claude-opus-4-7$5 / $251M128kJan 2026
Opus 4.6claude-opus-4-6anthropic.claude-opus-4-6-v1claude-opus-4-6$5 / $251M128kMay 2025
Sonnet 4.6claude-sonnet-4-6anthropic.claude-sonnet-4-6claude-sonnet-4-6$3 / $151M128kAug 2025
Sonnet 4.5claude-sonnet-4-5-20250929 (alias claude-sonnet-4-5)anthropic.claude-sonnet-4-5-20250929-v1:0claude-sonnet-4-5@20250929$3 / $15200k64kJan 2025
Opus 4.5claude-opus-4-5-20251101 (alias claude-opus-4-5)anthropic.claude-opus-4-5-20251101-v1:0claude-opus-4-5@20251101$5 / $25200k64kMay 2025

Aliases, and why yours may not mean what you think

Claude Code takes either a full model name or an alias. Aliases track the recommended version and change over time, which is convenient until you are on a provider where they resolve somewhere else.

The aliases, per Claude Code documentation in August 2026.
AliasWhat it does
defaultClears any override and reverts to the recommended model for your account
bestFable 5.1 where your organisation has access, otherwise the latest Opus
fableFable 5.1, for the hardest and longest-running tasks
opusThe latest Opus for your provider (Opus 5.5 on the Claude API)
sonnetThe latest Sonnet for your provider
haikuThe fast Haiku model for simple work
opusplanOpus during plan mode, Sonnet for execution
sonnet[1m] / opus[1m] / opusplan[1m]sonnet[1m] and opus[1m] select the 1M window. opusplan[1m] forces 1M on both the plan and the execute phases
Where the two common aliases land, by provider. Source: Claude Code model configuration, opened 19 August 2026.
Provideropussonnet
Anthropic APIOpus 5Sonnet 5
Claude Platform on AWSOpus 5Sonnet 4.6
Amazon Bedrock, Google Cloud's Agent PlatformOpus 5Sonnet 4.5
Microsoft FoundryOpus 4.6Sonnet 4.5
Environment pins from Claude Code model configuration, opened 19 August 2026. Each redirects the matching family alias to a specific model ID.
VariableFamily it pins
ANTHROPIC_DEFAULT_OPUS_MODELOpus
ANTHROPIC_DEFAULT_SONNET_MODELSonnet
ANTHROPIC_DEFAULT_HAIKU_MODELHaiku
ANTHROPIC_DEFAULT_FABLE_MODELFable

1M context by plan

Claude Code's extended-context page, opened 19 August 2026, splits 1M access by plan and by model. On the Anthropic API, Fable 5, Sonnet 5, and Opus 4.7 and later always run with the 1M window. Sonnet 5 is native 1M on that API: no credit gate, no 200K variant, and no [1m] suffix to select.

Opus and Sonnet 4.6 1M access from Claude Code model configuration, opened 19 August 2026. Sonnet 5 is not in this credit-gate table because it is native 1M on the Anthropic API.
PlanOpus 1MSonnet 4.6 1M
Max, Team, EnterpriseIncluded with the subscriptionUsage credits
ProUsage creditsUsage credits
API and pay-as-you-goFull accessFull access
  • Max, Team, and Enterprise auto-upgrade Opus to 1M with no extra configuration.
  • Pro needs usage credits for Opus 1M.
  • Sonnet 5 on the Anthropic API is native 1M on every plan. There is no credit gate.
  • opusplan[1m] forces 1M on both the plan-mode Opus phase and the execute-mode Sonnet phase. Use it when you are not on an auto-upgrade tier and want both phases at 1M.
  • The 1M window uses standard model pricing. There is no extra rate past 200K.

Effort levels

Effort controls adaptive reasoning: how often and how deeply the model thinks on each step. Available levels come from Claude Code model configuration, opened 19 August 2026. Models not in this table do not support effort, except Haiku 4.5, which is listed so the none row is explicit. Default is high except Opus 4.7, which defaults to xhigh.

Effort levels by model, from Claude Code model configuration, opened 19 August 2026.
ModelLevelsDefault
Fable 5low, medium, high, xhigh, maxhigh
Opus 5, Sonnet 5, Opus 4.8low, medium, high, xhigh, maxhigh
Opus 4.7low, medium, high, xhigh, maxxhigh
Opus 4.6, Sonnet 4.6low, medium, high, max (no xhigh)high
Haiku 4.5nonenone. No effort control

Thinking cannot be turned off on Fable 5. The session toggle, alwaysThinkingEnabled, and MAX_THINKING_TOKENS=0 have no effect there. Source: Claude Code model configuration, opened 19 August 2026.

Automatic model fallback

Fable 5 and Opus 5 run with safety classifiers for cybersecurity and biology content. When a classifier flags a request and the flagged category has a fallback model, Claude Code re-runs the request on that model and shows a notice in the transcript. After a fallback, the session continues on the fallback model. To return to your original model, run /model. Category-based fallback requires Claude Code v2.1.219 or later. Before v2.1.219, every flagged Fable 5 request re-ran on your provider's default Opus model, and Opus 5 was not a fallback source. Source: Claude Code model configuration, opened 19 August 2026.

Automatic model fallback. Claude Code v2.1.219+. Source: Claude Code model configuration, opened 19 August 2026.
Selected modelFlagWhat happens
Fable 5Biology-flaggedRe-run on Opus 5
Fable 5Cybersecurity-flaggedRe-run on Opus 4.8
Opus 5Cybersecurity-flaggedRe-run on Opus 4.8
Opus 5Biology-flaggedEnds in a refusal. Opus 5 has no biology fallback

To decide what happens each time a request is flagged, rather than switching automatically, run /config and turn off Switch models when a message is flagged, or set switchModelsOnFlag to false in your settings file. A flagged request then pauses the session with two options: switch to the fallback model, or edit the prompt and retry on the current model. Source: Claude Code model configuration, opened 19 August 2026.

  • A biology flag on Opus 5 has no fallback. Claude Code does not show the prompt and the request ends with the refusal.
  • When the fallback target is blocked by availableModels, Claude Code does not show the prompt. The flagged request ends with the refusal.
  • If both models flag the same request, edit the prompt and retry, or start a new session.
  • On mobile Claude Code on the web, editing and retrying is not supported. Switch models, or continue from a desktop browser or the desktop app.
  • In non-interactive mode and SDK integrations that cannot show the prompt, a flagged request ends the turn with a refusal.

Fallback model chains

This is a different mechanism from the classifier table above. When the primary model is overloaded, unavailable, or returns another non-retryable server error, Claude Code can switch to a fallback model instead of failing the request. Authentication, billing, rate-limit, request-size, and transport errors never trigger a switch. Those follow their normal retry and error handling. Source: Claude Code model configuration, opened 19 August 2026.

Configure one or more fallback models and Claude Code tries them in order, showing a notice when it switches. The switch lasts for the current turn only, so your next message tries the primary model first again. Claude Code caps chains at three models after duplicate removal and ignores extra entries. Set a chain for one session with --fallback-model (comma-separated list). To persist a chain, set fallbackModel in settings as an array. The flag takes precedence over the setting.

Session-only availability chain.
claude --fallback-model sonnet,haiku
Persisted availability chain in settings.
{
  "fallbackModel": ["claude-sonnet-5", "claude-haiku-4-5"]
}

The chain also covers compaction, but Claude Code will not fall back to a model with a smaller context window than the primary. If every fallback is smaller, compaction shows the original error and you can retry. Entries outside availableModels are dropped when the chain is read.

Setting and switching it

Four places, in descending priority.
# 1. in session, effective from the next turn
/model sonnet
/model claude-opus-5-5
/model               # opens the picker; Enter saves as default, s is session-only

# 2. at launch, this session only
claude --model opus

# 3. environment, this session only
ANTHROPIC_MODEL=claude-haiku-4-5 claude

# 4. settings file, persistent
#   { "model": "claude-sonnet-5" }
  • As of v2.1.153, /model in the picker: Enter saves the default for new sessions. s is this session only. A model set with /model in non-interactive -p applies to the current session only and is not saved.
  • Picker prices appear only when Claude Code talks to the Anthropic API, directly or through a proxying LLM gateway. On Amazon Bedrock and the Claude apps gateway, the rows show no price. The price is a display label only.
  • The picker asks for confirmation when the conversation already has output, because the next response re-reads the full history without cached context. That confirmation is a real cost warning, not a formality.
  • Resumed sessions keep the model they were using when the transcript was saved, so another terminal switching models cannot change yours on resume.
  • To run different models in different terminals at once, launch each with its own --model flag rather than switching with /model.
  • A subagent can pin its own model in frontmatter, and a skill can override the model for the turn it runs in.

A routing table

TaskModelWhy
Rename a symbol across 30 filesHaiku 4.5Mechanical; the spec is already complete
Write tests to a stated specHaiku 4.5Pattern-following
Generate a commit messageHaiku 4.5Trivial and high volume
Agent-team teammatesSonnet 5Anthropic recommends it for coordination work
Fix a test you have already locatedSonnet 5Bounded, needs some judgement
Add a feature across a few filesSonnet 5The normal case
Refactor a moduleSonnet 5Still the normal case
Debug something already attempted onceOpus 5Sonnet failing is itself the signal
Design an approach for a large changeOpus 5A wrong direction is the expensive outcome
Review a security-sensitive diffOpus 5One real catch pays for many reviews
A task larger than one sittingFable 5Built for long autonomous runs

What the gap costs on one real task

List rates per million tokens are hard to reason about, because an agent session is mostly cache reads. Here is the same task priced four ways, so the ratio is concrete rather than notional.

The task: 300,000 input tokens across the session, of which 90 percent arrive as cache reads, plus 20,000 output tokens. That is a realistic shape for a bounded feature or a substantial debugging run.

Arithmetic at August 2026 Claude API rates.
ModelFresh inputCache readsOutputTotal
Haiku 4.5$0.03$0.03$0.10$0.16
Sonnet 5$0.06$0.05$0.20$0.31
Opus 5$0.15$0.14$0.50$0.79
Fable 5$0.30$0.27$1.00$1.57

On a subscription the same ratios apply to your usage window rather than to a dollar figure: a session or weekly allowance is consumed roughly in proportion to token cost, and the windows are shared across models. Switching to Sonnet with /model after hitting an Opus-specific limit keeps you working; it does not restore a session or weekly window, because those are not per model.

Escalate on evidence, not on anxiety

Anthropic publishes two pieces of guidance that read differently, and both are correct in context. The platform models page says to start with Opus 5 for complex agentic coding and enterprise work. The Claude Code cost guidance says Sonnet handles most coding tasks and to reserve Opus for architectural decisions and multi-step reasoning. The first is answering "which model is capable of this"; the second is answering "which model should run your Tuesday".

The efficient pattern resolves both. It is not choosing a model per project, it is starting on the cheaper capable model your account actually has and escalating the moment a task has shown you it is hard.

  1. If your account defaults to Sonnet 5, stay there for routine work. Pro and Team standard start on Sonnet. Max and API start on Opus 5, so drop to Sonnet with /model sonnet when the task does not need Opus.
  2. Escalate to Opus 5 when Sonnet has produced a wrong answer twice, when the task is architectural, or when you cannot describe the fix yourself.
  3. Drop to Haiku 4.5 the moment a task turns mechanical, which frequently happens partway through: Opus decides the approach, Haiku applies it across thirty files.
  4. Escalate for review even when the work was done cheaply. A fresh Opus read of a finished diff in a subagent is one of the best uses of the expensive model there is.
  5. Or let opusplan do it. It runs Opus during plan mode and Sonnet for execution, which is the same discipline with no manual switching.

Model choice also interacts with reasoning effort, and the two multiply. High effort on Opus 5 is many times the cost of the default on Sonnet 5 for work the second combination often completes the same way. Decide the model first, then the effort, and change one dial at a time.

Questions people ask

Which Claude model is best for coding?

It depends on the account and the task. Pro and Team standard default to Sonnet 5, which is the right everyday coding model. Max and the Anthropic API default to Opus 5. Haiku 4.5 is for mechanical, fully specified work. Claude Opus 5 is for hard reasoning, architecture, and bugs that have already defeated one attempt. Fable 5 is for work that spans more than one sitting, and it is never the default.

What is Claude Opus 5?

Claude Opus 5 is Anthropic's coding flagship at $5 input and $25 output per million tokens on the API. Use it for hard reasoning, architecture, and bugs that have already defeated one attempt. It is the default on Max, Team premium, Enterprise pay-as-you-go, and the Anthropic API. It is not the default on Pro or Team standard.

What is Fable 5?

Fable 5 is for work larger than a single sitting. It is never the default. On Pro and Team standard it bills usage credits from the first token. On Max and Team premium it can use up to 50 percent of the same weekly pool, not extra. The promotion that included Fable in Pro weekly limits ended 19 July 2026. Claude Code needs v2.1.170 or later. API rates are on Claude Code pricing.

How do I change the model in Claude Code?

Run /model followed by an alias or model name inside a session, or /model alone to open the picker. Enter saves it as your default for new sessions; s applies it to this session only. It takes effect from the next turn and does not lose your conversation.

What does opusplan do?

It is an alias that uses Opus during plan mode and then switches to Sonnet for execution. It buys the expensive model where the decision is made and the cheaper one where the typing happens, with no manual switching.

Is Opus worth the extra cost?

For the hard few percent of work, clearly. As an everyday default on work Sonnet completes identically, no: at $5 and $25 per million it consumes quota several times faster. Max and API accounts start on Opus 5 anyway; switch down with /model sonnet when the task does not need it.

Is Haiku good enough for real work?

For mechanical, fully specified edits, yes, and at $1 and $5 per million it is very cheap. For anything requiring judgement about unfamiliar code, no. Note it also has a 200k context window rather than 1M, and does not support effort levels.

Does the model affect my rate limit?

Yes, substantially. Subscription usage windows are consumed roughly in proportion to token cost, so Opus exhausts an allowance several times faster than Sonnet for the same work. Switching models with /model does not reset a session or weekly window, because those are shared across models.

Why does my opus alias give me a different model than a colleague?

Aliases resolve per provider. On the Anthropic API opus is Opus 5; on Microsoft Foundry it is Opus 4.6. Pin the full model name, or set ANTHROPIC_DEFAULT_OPUS_MODEL, if you need them to match.

Should I use different models in one session?

Yes, that is the most efficient pattern. Opus to decide an approach, Sonnet or Haiku to apply it, Opus again in a fresh subagent to review the finished diff.

What model does Claude Code start on?

It depends on the account. Max, Team premium, Enterprise pay-as-you-go, and the Anthropic API start on Opus 5. Claude Platform on AWS, Amazon Bedrock, and Google Cloud Agent Platform also start on Opus 5. Pro, Team standard, and Enterprise subscription seats start on Sonnet 5. Microsoft Foundry starts on Sonnet 4.5. Fable 5 is never the default. An organization default set by an admin replaces the row.

Is Fable 5 included on Pro?

No. On Pro and Team standard, Fable 5 bills usage credits from the first token. A promotion that included Fable 5 in Pro weekly limits ended 19 July 2026 at 11:59:59 PM PT. Eligible Pro and Team standard seats get a one-time credit. Free does not include Fable 5.

Will this use my regular usage limits?

It depends on your plan. On the Max plan, premium seats on the Team plan, and premium seats on the seat-based Enterprise plan, Fable 5 counts toward your plan's usage limits, and you can use up to 50% of your weekly usage limits on Fable 5 at no extra cost. On the Pro plan, standard seats on the Team plan, and standard seats on the seat-based Enterprise plan, Fable 5 runs on usage credits rather than your plan's usage limits.

Sources

Every figure above was read from these pages on August 2026. Vendors reprice without notice; if you find a stale number, tell us.

  1. Claude Code: model configuration (opened 19 August 2026)
  2. Anthropic models overview (opened 19 August 2026)
  3. Anthropic pricing (opened 19 August 2026)
  4. Claude Fable 5 on your plan (opened 19 August 2026)
  5. Claude plans and pricing (opened 19 August 2026)
  6. Project Glasswing (opened 19 August 2026)
  7. Claude Code: manage costs effectively
Try it

Spend, split
by model.

Continuum shows which model consumed what, which is usually more persuasive than any routing advice. Get Plus when a Claude window is spent and the job is not.

Plus is $25/mo hosted overflow · cancel anytime