The Claude Code default depends on the account, not on one universal pick. Max, Team premium, Enterprise pay-as-you-go, and the Anthropic API start on Opus 5. Pro, Team standard, and Enterprise subscription seats start on Sonnet 5. Fable 5 is never the default. On Pro and Team standard it bills usage credits from the first token. On Max and Team premium it can use up to 50 percent of the same weekly pool, not extra. Switch per task with /model.
- Default is account-specific. Max and API start on Opus 5. Pro and Team standard start on Sonnet 5. Fable 5 is never the default.
- Haiku 4.5 is genuinely capable on mechanical edits at $1 and $5 per million.
- Opus 5 is $5 and $25 per million. It earns that on reasoning, not on volume.
- Fable 5 on Pro and Team standard is credits from the first token. On Max it is up to 50 percent of the same weekly pool.
- Aliases resolve differently per provider.
opusis not the same model everywhere. opusplanplans on Opus and executes on Sonnet, which is the cheapest good habit here.
Claude models: the August 2026 lineup
| Model | API ID | Input / output | Context | Max out | Reliable cutoff | Train cutoff | Thinking | Effort |
|---|---|---|---|---|---|---|---|---|
| Fable 5 | claude-fable-5 | $10 / $50 | 1M | 128k | Jan 2026 | Jan 2026 | Adaptive, always on. Extended thinking: no | Yes, default high |
| Claude Opus 5 | claude-opus-5 | $5 / $25 | 1M | 128k | May 2026 | May 2026 | Adaptive: yes. Extended thinking: no | Yes, default high |
| Sonnet 5 | claude-sonnet-5 | $2 / $10 | 1M | 128k | Jan 2026 | Jan 2026 | Adaptive: yes. Extended thinking: no | Yes, default high |
| Haiku 4.5 | claude-haiku-4-5-20251001 (alias claude-haiku-4-5) | $1 / $5 | 200k | 64k | Feb 2025 | Jul 2025 | Extended thinking: yes. Adaptive: no | None |
| Model | AWS Bedrock ID | Google Cloud ID |
|---|---|---|
| Fable 5 | anthropic.claude-fable-5 | claude-fable-5 |
| Opus 5 | anthropic.claude-opus-5 | claude-opus-5 |
| Sonnet 5 | anthropic.claude-sonnet-5 | claude-sonnet-5 |
| Haiku 4.5 | anthropic.claude-haiku-4-5-20251001-v1:0 | claude-haiku-4-5@20251001 |
Claude Opus 5 is the coding flagship at $5 input and $25 output per million on the API. Fable 5 sits above it for runs larger than one sitting. How Fable is billed on a Claude plan is in the table below, from Anthropic's Fable plan article. The Fable API table is also on Claude Code pricing. Session and weekly windows shared across models live on Claude Code limits. How those Claude models sit against GPT-5.6 and Gemini is on best coding models.
| Model | Base input | 5m cache write | 1h cache write | Cache hit | Output |
|---|---|---|---|---|---|
| Fable 5 | $10 | $12.50 | $20 | $1 | $50 |
| Opus 5 | $5 | $6.25 | $10 | $0.50 | $25 |
| Sonnet 5 | $2 | $2.50 | $4 | $0.20 | $10 |
| Haiku 4.5 | $1 | $1.25 | $2 | $0.10 | $5 |
| Model | Fast input | Fast output |
|---|---|---|
| Opus 5 | $10 | $50 |
| Opus 4.8 | $10 | $50 |
inference_geo: "us" is 1.1x on every token category (input, output, cache write, cache read). Global is the default and stays at list price. Fast mode stacks on top of this meter. Partner Bedrock and Google Cloud regional prices are separate.| Model | US-only input / output (1.1x) |
|---|---|
| Fable 5 | $11 / $55 |
| Opus 5 | $5.50 / $27.50 |
| Sonnet 5 | $2.20 / $11 |
- Output is priced five times input on every current model, so the output column is where the difference actually bites.
- Cache reads bill at 10 percent of the base input rate, which is why a long cached session is far cheaper than the input column alone suggests.
- The Batch API halves both input and output, but it is asynchronous and irrelevant to interactive coding.
- Haiku 4.5 does not support effort levels. Fable 5, Opus 5, and Sonnet 5 all do, and default to
high.
Coding-agent cards with harness, DeepSWE cost, and when not to use it: Opus 5.5, Sonnet 5.5, Opus 5, and Fable 5.
What your account defaults to
The default alias is not one model. Claude Code documentation, opened 19 August 2026, maps it by account type. Fable 5 is never the default on any of these rows. Sessions use Fable only after you choose it with /model fable, a model setting, or the best alias where Fable is available.
| Account type | Default model |
|---|---|
| Max, Team premium, Enterprise pay-as-you-go, Anthropic API | Opus 5 |
| Claude Platform on AWS, Amazon Bedrock, Google Cloud Agent Platform | Opus 5 |
| Pro, Team standard, Enterprise subscription seats | Sonnet 5 |
| Microsoft Foundry | Sonnet 4.5 |
Fable 5 on your Claude plan
A promotion that included Fable 5 in Pro weekly limits ended 19 July 2026 at 11:59:59 PM PT. Starting 20 July 2026, access depends on the plan. This table is from Anthropic's Fable plan article and claude.com/pricing, opened 19 August 2026. Claude Code needs v2.1.170 or later.
| Plan | Included in weekly limit? | Cap | What happens at the cap | How to keep working |
|---|---|---|---|---|
| Free | Not available | Fable 5 is paid-plan only | You cannot start a Fable 5 session | Upgrade to a paid plan |
| Max, Team premium, seat-based Enterprise premium | Yes | Up to 50% of the same weekly pool, not extra | The Fable 5 weekly slice is used up | Usage credits, or switch to another model |
| Pro, Team standard | No | Credits from the first request. Eligible seats get a one-time credit. Promo that included Fable in Pro weekly limits ended 19 Jul 2026 11:59:59 PM PT | There is no weekly Fable slice. Further Fable 5 uses credits | Credits, or upgrade to Max |
| Seat-based Enterprise standard | No | Credits only if the org enabled them. No one-time promo credit | Fable 5 is blocked if credits are off | Ask an admin to enable credits, or switch model |
| Usage-based Enterprise / Claude API | Own meter, not the weekly-limit table | Billed at standard API rates | The weekly-limit table does not apply | Keep working on the API meter |
Fable 5 is not available under zero data retention. The /model picker either omits it or shows it disabled. Source: Claude Code model configuration, opened 19 August 2026.
When Fable 5 bills usage credits, the /model picker shows "Requires usage credits" on the Fable 5 row. Interactive sessions get a consent prompt before that request bills. Members of Enterprise plans with organization billing do not see the prompt. After you choose to continue on Fable 5 using usage credits, the prompt does not return. In non-interactive mode with the -p flag, and through the Agent SDK, Claude Code never shows the consent prompt and bills without asking. Source: Claude Code model configuration, opened 19 August 2026.
Older Claude models still available
| Model | API ID | AWS Bedrock ID | Google Cloud ID | Input / output | Context | Max out | Reliable cutoff |
|---|---|---|---|---|---|---|---|
| Opus 4.8 | claude-opus-4-8 | anthropic.claude-opus-4-8 | claude-opus-4-8 | $5 / $25 | 1M | 128k | Jan 2026 |
| Opus 4.7 | claude-opus-4-7 | anthropic.claude-opus-4-7 | claude-opus-4-7 | $5 / $25 | 1M | 128k | Jan 2026 |
| Opus 4.6 | claude-opus-4-6 | anthropic.claude-opus-4-6-v1 | claude-opus-4-6 | $5 / $25 | 1M | 128k | May 2025 |
| Sonnet 4.6 | claude-sonnet-4-6 | anthropic.claude-sonnet-4-6 | claude-sonnet-4-6 | $3 / $15 | 1M | 128k | Aug 2025 |
| Sonnet 4.5 | claude-sonnet-4-5-20250929 (alias claude-sonnet-4-5) | anthropic.claude-sonnet-4-5-20250929-v1:0 | claude-sonnet-4-5@20250929 | $3 / $15 | 200k | 64k | Jan 2025 |
| Opus 4.5 | claude-opus-4-5-20251101 (alias claude-opus-4-5) | anthropic.claude-opus-4-5-20251101-v1:0 | claude-opus-4-5@20251101 | $5 / $25 | 200k | 64k | May 2025 |
Aliases, and why yours may not mean what you think
Claude Code takes either a full model name or an alias. Aliases track the recommended version and change over time, which is convenient until you are on a provider where they resolve somewhere else.
| Alias | What it does |
|---|---|
default | Clears any override and reverts to the recommended model for your account |
best | Fable 5.1 where your organisation has access, otherwise the latest Opus |
fable | Fable 5.1, for the hardest and longest-running tasks |
opus | The latest Opus for your provider (Opus 5.5 on the Claude API) |
sonnet | The latest Sonnet for your provider |
haiku | The fast Haiku model for simple work |
opusplan | Opus during plan mode, Sonnet for execution |
sonnet[1m] / opus[1m] / opusplan[1m] | sonnet[1m] and opus[1m] select the 1M window. opusplan[1m] forces 1M on both the plan and the execute phases |
| Provider | opus | sonnet |
|---|---|---|
| Anthropic API | Opus 5 | Sonnet 5 |
| Claude Platform on AWS | Opus 5 | Sonnet 4.6 |
| Amazon Bedrock, Google Cloud's Agent Platform | Opus 5 | Sonnet 4.5 |
| Microsoft Foundry | Opus 4.6 | Sonnet 4.5 |
| Variable | Family it pins |
|---|---|
ANTHROPIC_DEFAULT_OPUS_MODEL | Opus |
ANTHROPIC_DEFAULT_SONNET_MODEL | Sonnet |
ANTHROPIC_DEFAULT_HAIKU_MODEL | Haiku |
ANTHROPIC_DEFAULT_FABLE_MODEL | Fable |
1M context by plan
Claude Code's extended-context page, opened 19 August 2026, splits 1M access by plan and by model. On the Anthropic API, Fable 5, Sonnet 5, and Opus 4.7 and later always run with the 1M window. Sonnet 5 is native 1M on that API: no credit gate, no 200K variant, and no [1m] suffix to select.
| Plan | Opus 1M | Sonnet 4.6 1M |
|---|---|---|
| Max, Team, Enterprise | Included with the subscription | Usage credits |
| Pro | Usage credits | Usage credits |
| API and pay-as-you-go | Full access | Full access |
- Max, Team, and Enterprise auto-upgrade Opus to 1M with no extra configuration.
- Pro needs usage credits for Opus 1M.
- Sonnet 5 on the Anthropic API is native 1M on every plan. There is no credit gate.
opusplan[1m]forces 1M on both the plan-mode Opus phase and the execute-mode Sonnet phase. Use it when you are not on an auto-upgrade tier and want both phases at 1M.- The 1M window uses standard model pricing. There is no extra rate past 200K.
Effort levels
Effort controls adaptive reasoning: how often and how deeply the model thinks on each step. Available levels come from Claude Code model configuration, opened 19 August 2026. Models not in this table do not support effort, except Haiku 4.5, which is listed so the none row is explicit. Default is high except Opus 4.7, which defaults to xhigh.
| Model | Levels | Default |
|---|---|---|
| Fable 5 | low, medium, high, xhigh, max | high |
| Opus 5, Sonnet 5, Opus 4.8 | low, medium, high, xhigh, max | high |
| Opus 4.7 | low, medium, high, xhigh, max | xhigh |
| Opus 4.6, Sonnet 4.6 | low, medium, high, max (no xhigh) | high |
| Haiku 4.5 | none | none. No effort control |
Thinking cannot be turned off on Fable 5. The session toggle, alwaysThinkingEnabled, and MAX_THINKING_TOKENS=0 have no effect there. Source: Claude Code model configuration, opened 19 August 2026.
Automatic model fallback
Fable 5 and Opus 5 run with safety classifiers for cybersecurity and biology content. When a classifier flags a request and the flagged category has a fallback model, Claude Code re-runs the request on that model and shows a notice in the transcript. After a fallback, the session continues on the fallback model. To return to your original model, run /model. Category-based fallback requires Claude Code v2.1.219 or later. Before v2.1.219, every flagged Fable 5 request re-ran on your provider's default Opus model, and Opus 5 was not a fallback source. Source: Claude Code model configuration, opened 19 August 2026.
| Selected model | Flag | What happens |
|---|---|---|
| Fable 5 | Biology-flagged | Re-run on Opus 5 |
| Fable 5 | Cybersecurity-flagged | Re-run on Opus 4.8 |
| Opus 5 | Cybersecurity-flagged | Re-run on Opus 4.8 |
| Opus 5 | Biology-flagged | Ends in a refusal. Opus 5 has no biology fallback |
To decide what happens each time a request is flagged, rather than switching automatically, run /config and turn off Switch models when a message is flagged, or set switchModelsOnFlag to false in your settings file. A flagged request then pauses the session with two options: switch to the fallback model, or edit the prompt and retry on the current model. Source: Claude Code model configuration, opened 19 August 2026.
- A biology flag on Opus 5 has no fallback. Claude Code does not show the prompt and the request ends with the refusal.
- When the fallback target is blocked by
availableModels, Claude Code does not show the prompt. The flagged request ends with the refusal. - If both models flag the same request, edit the prompt and retry, or start a new session.
- On mobile Claude Code on the web, editing and retrying is not supported. Switch models, or continue from a desktop browser or the desktop app.
- In non-interactive mode and SDK integrations that cannot show the prompt, a flagged request ends the turn with a refusal.
Fallback model chains
This is a different mechanism from the classifier table above. When the primary model is overloaded, unavailable, or returns another non-retryable server error, Claude Code can switch to a fallback model instead of failing the request. Authentication, billing, rate-limit, request-size, and transport errors never trigger a switch. Those follow their normal retry and error handling. Source: Claude Code model configuration, opened 19 August 2026.
Configure one or more fallback models and Claude Code tries them in order, showing a notice when it switches. The switch lasts for the current turn only, so your next message tries the primary model first again. Claude Code caps chains at three models after duplicate removal and ignores extra entries. Set a chain for one session with --fallback-model (comma-separated list). To persist a chain, set fallbackModel in settings as an array. The flag takes precedence over the setting.
claude --fallback-model sonnet,haiku
{
"fallbackModel": ["claude-sonnet-5", "claude-haiku-4-5"]
}
The chain also covers compaction, but Claude Code will not fall back to a model with a smaller context window than the primary. If every fallback is smaller, compaction shows the original error and you can retry. Entries outside availableModels are dropped when the chain is read.
Setting and switching it
# 1. in session, effective from the next turn
/model sonnet
/model claude-opus-5-5
/model # opens the picker; Enter saves as default, s is session-only
# 2. at launch, this session only
claude --model opus
# 3. environment, this session only
ANTHROPIC_MODEL=claude-haiku-4-5 claude
# 4. settings file, persistent
# { "model": "claude-sonnet-5" }
- As of v2.1.153,
/modelin the picker: Enter saves the default for new sessions.sis this session only. A model set with/modelin non-interactive-papplies to the current session only and is not saved. - Picker prices appear only when Claude Code talks to the Anthropic API, directly or through a proxying LLM gateway. On Amazon Bedrock and the Claude apps gateway, the rows show no price. The price is a display label only.
- The picker asks for confirmation when the conversation already has output, because the next response re-reads the full history without cached context. That confirmation is a real cost warning, not a formality.
- Resumed sessions keep the model they were using when the transcript was saved, so another terminal switching models cannot change yours on resume.
- To run different models in different terminals at once, launch each with its own
--modelflag rather than switching with/model. - A subagent can pin its own model in frontmatter, and a skill can override the model for the turn it runs in.
A routing table
| Task | Model | Why |
|---|---|---|
| Rename a symbol across 30 files | Haiku 4.5 | Mechanical; the spec is already complete |
| Write tests to a stated spec | Haiku 4.5 | Pattern-following |
| Generate a commit message | Haiku 4.5 | Trivial and high volume |
| Agent-team teammates | Sonnet 5 | Anthropic recommends it for coordination work |
| Fix a test you have already located | Sonnet 5 | Bounded, needs some judgement |
| Add a feature across a few files | Sonnet 5 | The normal case |
| Refactor a module | Sonnet 5 | Still the normal case |
| Debug something already attempted once | Opus 5 | Sonnet failing is itself the signal |
| Design an approach for a large change | Opus 5 | A wrong direction is the expensive outcome |
| Review a security-sensitive diff | Opus 5 | One real catch pays for many reviews |
| A task larger than one sitting | Fable 5 | Built for long autonomous runs |
What the gap costs on one real task
List rates per million tokens are hard to reason about, because an agent session is mostly cache reads. Here is the same task priced four ways, so the ratio is concrete rather than notional.
The task: 300,000 input tokens across the session, of which 90 percent arrive as cache reads, plus 20,000 output tokens. That is a realistic shape for a bounded feature or a substantial debugging run.
| Model | Fresh input | Cache reads | Output | Total |
|---|---|---|---|---|
| Haiku 4.5 | $0.03 | $0.03 | $0.10 | $0.16 |
| Sonnet 5 | $0.06 | $0.05 | $0.20 | $0.31 |
| Opus 5 | $0.15 | $0.14 | $0.50 | $0.79 |
| Fable 5 | $0.30 | $0.27 | $1.00 | $1.57 |
On a subscription the same ratios apply to your usage window rather than to a dollar figure: a session or weekly allowance is consumed roughly in proportion to token cost, and the windows are shared across models. Switching to Sonnet with /model after hitting an Opus-specific limit keeps you working; it does not restore a session or weekly window, because those are not per model.
Escalate on evidence, not on anxiety
Anthropic publishes two pieces of guidance that read differently, and both are correct in context. The platform models page says to start with Opus 5 for complex agentic coding and enterprise work. The Claude Code cost guidance says Sonnet handles most coding tasks and to reserve Opus for architectural decisions and multi-step reasoning. The first is answering "which model is capable of this"; the second is answering "which model should run your Tuesday".
The efficient pattern resolves both. It is not choosing a model per project, it is starting on the cheaper capable model your account actually has and escalating the moment a task has shown you it is hard.
- If your account defaults to Sonnet 5, stay there for routine work. Pro and Team standard start on Sonnet. Max and API start on Opus 5, so drop to Sonnet with
/model sonnetwhen the task does not need Opus. - Escalate to Opus 5 when Sonnet has produced a wrong answer twice, when the task is architectural, or when you cannot describe the fix yourself.
- Drop to Haiku 4.5 the moment a task turns mechanical, which frequently happens partway through: Opus decides the approach, Haiku applies it across thirty files.
- Escalate for review even when the work was done cheaply. A fresh Opus read of a finished diff in a subagent is one of the best uses of the expensive model there is.
- Or let
opusplando it. It runs Opus during plan mode and Sonnet for execution, which is the same discipline with no manual switching.
Model choice also interacts with reasoning effort, and the two multiply. High effort on Opus 5 is many times the cost of the default on Sonnet 5 for work the second combination often completes the same way. Decide the model first, then the effort, and change one dial at a time.
Questions people ask
Which Claude model is best for coding?
It depends on the account and the task. Pro and Team standard default to Sonnet 5, which is the right everyday coding model. Max and the Anthropic API default to Opus 5. Haiku 4.5 is for mechanical, fully specified work. Claude Opus 5 is for hard reasoning, architecture, and bugs that have already defeated one attempt. Fable 5 is for work that spans more than one sitting, and it is never the default.
What is Claude Opus 5?
Claude Opus 5 is Anthropic's coding flagship at $5 input and $25 output per million tokens on the API. Use it for hard reasoning, architecture, and bugs that have already defeated one attempt. It is the default on Max, Team premium, Enterprise pay-as-you-go, and the Anthropic API. It is not the default on Pro or Team standard.
What is Fable 5?
Fable 5 is for work larger than a single sitting. It is never the default. On Pro and Team standard it bills usage credits from the first token. On Max and Team premium it can use up to 50 percent of the same weekly pool, not extra. The promotion that included Fable in Pro weekly limits ended 19 July 2026. Claude Code needs v2.1.170 or later. API rates are on Claude Code pricing.
How do I change the model in Claude Code?
Run /model followed by an alias or model name inside a session, or /model alone to open the picker. Enter saves it as your default for new sessions; s applies it to this session only. It takes effect from the next turn and does not lose your conversation.
What does opusplan do?
It is an alias that uses Opus during plan mode and then switches to Sonnet for execution. It buys the expensive model where the decision is made and the cheaper one where the typing happens, with no manual switching.
Is Opus worth the extra cost?
For the hard few percent of work, clearly. As an everyday default on work Sonnet completes identically, no: at $5 and $25 per million it consumes quota several times faster. Max and API accounts start on Opus 5 anyway; switch down with /model sonnet when the task does not need it.
Is Haiku good enough for real work?
For mechanical, fully specified edits, yes, and at $1 and $5 per million it is very cheap. For anything requiring judgement about unfamiliar code, no. Note it also has a 200k context window rather than 1M, and does not support effort levels.
Does the model affect my rate limit?
Yes, substantially. Subscription usage windows are consumed roughly in proportion to token cost, so Opus exhausts an allowance several times faster than Sonnet for the same work. Switching models with /model does not reset a session or weekly window, because those are shared across models.
Why does my opus alias give me a different model than a colleague?
Aliases resolve per provider. On the Anthropic API opus is Opus 5; on Microsoft Foundry it is Opus 4.6. Pin the full model name, or set ANTHROPIC_DEFAULT_OPUS_MODEL, if you need them to match.
Should I use different models in one session?
Yes, that is the most efficient pattern. Opus to decide an approach, Sonnet or Haiku to apply it, Opus again in a fresh subagent to review the finished diff.
What model does Claude Code start on?
It depends on the account. Max, Team premium, Enterprise pay-as-you-go, and the Anthropic API start on Opus 5. Claude Platform on AWS, Amazon Bedrock, and Google Cloud Agent Platform also start on Opus 5. Pro, Team standard, and Enterprise subscription seats start on Sonnet 5. Microsoft Foundry starts on Sonnet 4.5. Fable 5 is never the default. An organization default set by an admin replaces the row.
Is Fable 5 included on Pro?
No. On Pro and Team standard, Fable 5 bills usage credits from the first token. A promotion that included Fable 5 in Pro weekly limits ended 19 July 2026 at 11:59:59 PM PT. Eligible Pro and Team standard seats get a one-time credit. Free does not include Fable 5.
Will this use my regular usage limits?
It depends on your plan. On the Max plan, premium seats on the Team plan, and premium seats on the seat-based Enterprise plan, Fable 5 counts toward your plan's usage limits, and you can use up to 50% of your weekly usage limits on Fable 5 at no extra cost. On the Pro plan, standard seats on the Team plan, and standard seats on the seat-based Enterprise plan, Fable 5 runs on usage credits rather than your plan's usage limits.
Sources
Every figure above was read from these pages on August 2026. Vendors reprice without notice; if you find a stale number, tell us.
- Claude Code: model configuration (opened 19 August 2026)
- Anthropic models overview (opened 19 August 2026)
- Anthropic pricing (opened 19 August 2026)
- Claude Fable 5 on your plan (opened 19 August 2026)
- Claude plans and pricing (opened 19 August 2026)
- Project Glasswing (opened 19 August 2026)
- Claude Code: manage costs effectively