Claude Code runs a rolling 5-hour window for burst protection, a separate weekly cap for total volume, and a separate allowance for Opus. The session and weekly limits are shared across every model, so switching model does not restore access; the Opus limit is the one exception. Consumption tracks tokens, not turns, so model choice, context size, and tool output volume dominate. Anthropic doubled the 5-hour limits on 6 May 2026 and has been running a 50% weekly increase that ClaudeDevs extended again on 29 August 2026, now until 14 September. Starting 14 September, ClaudeDevs says standard weekly limits are permanently 25% higher than the old baseline for Pro, Max, Team, and seat-based Enterprise. They said that is a 17% reduction vs the boosted level. The 5-hour doubling does not expire with it. Fable 5 sits on top of that weekly pool: on Max and premium seats you can spend up to 50% of the weekly limit on Fable 5, which is a slice of the same cap, not extra capacity. Run /usage for your own bars. Do not trust a blog for the clock.
- There are three limits, not one: a rolling 5-hour session window, a weekly cap, and a separate Opus allowance.
- The session and weekly limits are shared across models. Switching with
/modeldoes not restore access. The exceptions are the model-scoped ones: the Opus allowance, and on Max and premium seats the Fable 5 slice of the weekly pool. - The 5-hour window is rolling: capacity returns gradually, and the error tells you the reset time.
- Anthropic doubled the 5-hour limits for Pro, Max, Team, and seat-based Enterprise on 6 May 2026 and removed the peak-hours reduction.
- The Claude Code weekly +50% promo until 14 September 2026 is extra weekly capacity. Starting 14 September, ClaudeDevs says standard weekly limits are permanently +25% vs the old baseline (they said that is a 17% cut vs the boosted week). Fable 5's "up to 50% of your weekly limit" is a slice of that same pool, not a second increase.
- On Pro and Team standard, Fable 5 is not inside plan limits: it bills usage credits from the first request. Select it with
/model fableon Claude Code v2.1.170 or later. - Consumption tracks tokens, not turns. One Opus run over a large repo can cost more than fifty short Haiku messages.
- Your status line can render
rate_limits.five_hour.used_percentageandrate_limits.seven_day.used_percentagecontinuously, so you see the wall before you hit it.
Three limits, not one
When Claude Code stops accepting work, it names which ceiling you hit. The three messages look almost identical and mean very different things.
You've hit your session limit · resets 3:45pm
You've hit your weekly limit · resets Mon 12:00am
You've hit your Opus limit · resets 3:45pm
| Session (5-hour) | Weekly | Opus | |
|---|---|---|---|
| Purpose | Burst protection | Total volume budget | Protects the most expensive model |
| Shape | Rolling, refills gradually | Resets on a weekly schedule | Model-scoped allowance |
| Shared across models | Yes | Yes | No, Opus requests only |
Does /model help? | No | No | Yes. Switch and keep working |
| Typical trigger | One intense afternoon | A sustained heavy week | Leaving Opus as your default |
One more model-scoped ceiling sits inside the weekly cap rather than beside it. On Max, Team premium, and seat-based Enterprise premium, Fable 5 can draw at most 50% of the weekly pool. Like the Opus allowance, switching away from it keeps you working; unlike the Opus allowance, it is a slice of the weekly budget rather than a separate one, so it can never give you more total room. Claude Fable 5 in Claude Code has the plan-by-plan detail.
The rolling window is the part people get wrong
It is not a bucket that empties at a fixed time. It is a trailing measurement of what you consumed over the past five hours. If you burned most of your allowance between 9am and 11am, that consumption ages out gradually through the early afternoon rather than all at once. Sitting and waiting for a "reset moment" is the wrong mental model; capacity is already returning while you wait, and the error prints the exact time it clears.
What actually consumes your allowance
Limits are measured in compute, which tracks tokens processed, not messages sent. Five things dominate, roughly in order of size.
Model choice
Opus-class models cost several times what Haiku-class models cost for identical work. This is the largest single lever available to you, and it is one keystroke.
Long context
Claude Code sends your full conversation with every request, and each tool use sends another request carrying that batch of results. A one-line question in a session that has been open all day still draws usage for the whole conversation.
Cache misses
Your first message after a break longer than the cache lifetime reprocesses the full context at the full input rate instead of a tenth of it. The lifetime is an hour on a subscription, and drops to five minutes once you are drawing on usage credits or using an API key.
Tool output volume
A grep that returns 4,000 lines, a test suite that prints its whole log, a read of a 6,000-line module: all of it lands in context and is re-sent on every later turn. Agents that run noisy commands burn allowance quickly.
Parallel work
Two agents on one account draw from one pool. Agent teams are worse: Anthropic documents them at roughly 7x the tokens of a standard session when teammates run in plan mode, because each teammate keeps its own context window.
Reading the gauges before you get blocked
The failure mode worth designing around is not the block itself. It is being blocked twenty minutes into a refactor with a half-applied change on disk. Three surfaces tell you where you stand.
1. /usage, on demand
On a Pro, Max, Team, or Enterprise plan, /usage shows plan usage bars plus a breakdown of what is consuming them: usage attributed to skills, subagents, plugins, and individual MCP servers, each as a percentage. It also raises behavior flags for anything accounting for 10% or more of recent usage, such as long context or cache misses. Press d or w to switch between the last 24 hours and the last 7 days.
2. The status line, continuously
A custom status line receives session JSON on stdin, and that JSON carries the live rate-limit state. This is the cheapest way to keep both windows in front of you at all times.
{
"rate_limits": {
"five_hour": { "used_percentage": 23.5, "resets_at": 1738425600 },
"seven_day": { "used_percentage": 41.2, "resets_at": 1738857600 }
},
"context_window": {
"used_percentage": 8,
"remaining_percentage": 92,
"context_window_size": 200000
}
}
#!/bin/bash
# save as ~/.claude/statusline.sh, chmod +x, then point settings.json at it
input=$(cat)
h5=$(echo "$input" | jq -r '.rate_limits.five_hour.used_percentage // 0')
d7=$(echo "$input" | jq -r '.rate_limits.seven_day.used_percentage // 0')
ctx=$(echo "$input" | jq -r '.context_window.used_percentage // 0')
printf '5h %.0f%% | 7d %.0f%% | ctx %.0f%%' "$h5" "$d7" "$ctx"
3. Every device, all the time
What changed in 2026
The limits landscape moved several times, which is why advice written in 2025 misleads now.
| Date | What changed |
|---|---|
| 2025 | Weekly caps introduced alongside the existing 5-hour window, in response to a small number of accounts running agents continuously. |
| 6 May 2026 | 5-hour limits doubled for Pro, Max, Team, and seat-based Enterprise plans. |
| 6 May 2026 | The peak-hours reduction was removed for Pro and Max Claude Code accounts, so busy periods no longer shrink your allowance. |
| 13 May 2026 | Weekly limits raised 50% as a promotion, initially through 13 July 2026. |
| 20 Jul 2026 | Fable 5 plan split: included for up to 50% of the weekly pool on Max and premium seats; usage credits from the first request on Pro and Team standard. Not extra weekly capacity. |
| 19 Aug 2026 | ClaudeDevs extended the 50% weekly promotion through 31 August 2026. Free plans and consumption-based Enterprise seats stay excluded. The 5-hour limit is unchanged by it. |
| 29 Aug 2026 | ClaudeDevs: the +50% weekly promotion stays until 14 September 2026. Then standard weekly limits are permanently +25% vs the old baseline (they said that is a 17% cut vs the boosted week). 5-hour doubling from 6 May is unchanged. Free plans and consumption-based Enterprise were not named. |
Claude Fable 5 in Claude Code
Fable 5 was Anthropic's most capable model when these plan rules were published. Fable 5.1 now holds that lineup slot; the rest of this section preserves the Fable 5-specific limits and July 2026 plan history.
It is not the default on any account type. /model fable and the best alias now select Fable 5.1; the plan rules below are the Fable 5 history they were published against. It is offered on paid Claude plans (Pro, Max, Team, Enterprise) and not on Free, though on seat-based Enterprise standard seats it only runs where the organisation enabled usage credits. The per-token rates live on the Claude Code pricing guide, which already carries the Fable 5 API table; this page is the plan-limit explainer. The rest of the claude models lineup, including Claude Opus 5, is on the claude models guide.
How to select it
Fable 5 requires Claude Code v2.1.170 or later. Older versions do not show it in the picker and cannot select it. Run claude update, then /model fable. You can also launch with claude --model fable. Choosing it in /model saves it as the selected model in your user settings, so later sessions start on Fable 5 until you change models. The best alias uses Fable 5.1 where your organisation has access, otherwise the latest Opus.
# Requires Claude Code v2.1.170 or later
claude update
/model fable
# Or start a session already on Fable
claude --model fable
The 20 July 2026 plan split
Until 19 July 2026 at 11:59:59 PM PT, a promotion let paid seats spend up to 50% of the weekly subscription limit on Fable 5 at no extra cost. Starting 20 July 2026, Anthropic split access by plan and seat. Fable 5 is still available. How it is billed is not.
| Plan / seat | How Fable 5 is billed |
|---|---|
| Max; Team premium; seat-based Enterprise premium | Included as a standard part of the plan. You can use up to 50% of your weekly usage limits on Fable 5 at no extra cost. It draws from the same weekly pool as every other model and uses that pool faster. When you reach the Fable 5 portion, keep going on usage credits, or switch model to stay inside the plan limits. |
| Pro; Team standard | Not included in plan usage limits. Pay-as-you-go usage credits from the first request. Eligible Pro and Team standard seats qualify for a one-time credit. |
| Seat-based Enterprise standard | Not included. Usage credits only if the organisation enabled them. These seats do not qualify for the one-time promotional credit. |
| Usage-based Enterprise; Claude API | Billed at standard API rates. The dollar table is on Claude Code pricing. |
| Free | Not available. |
On Max, Team premium, and seat-based Enterprise premium, that 50% is of the weekly limit, not 50% more. Other models draw from the same weekly pool. You can never use more than the weekly limit. Anthropic's own FAQ answers the obvious question with no: you do not get 50% more weekly capacity for Fable 5. The Claude Code weekly +50% promo until 14 September 2026 raises that pool for eligible plans; after that, ClaudeDevs says the pool stays permanently +25% vs the old baseline. It does not give Fable 5 its own extra bucket.
The Pro and standard-seat credits path
On Pro and Team standard seats, Fable 5 is not inside the plan's usage limits at all. The first Fable request bills pay-as-you-go usage credits. Eligible Pro and Team standard seats qualify for a one-time credit to cover the change; seat-based Enterprise standard seats do not. When Fable bills to credits, the /model picker shows "Requires usage credits" on the Fable 5 row.
Interactive sessions prompt for one-time consent before a Fable 5 request bills credits. You can continue on Fable with credits, switch back to your default model, or dismiss the prompt (the picker keeps the current model; a mid-session dismiss continues the turn on the default). Members of Enterprise plans with organisation billing skip that prompt. After you continue on credits, Claude Code does not ask again. Non-interactive runs with -p, and the Agent SDK, bill without asking.
How to stop hitting them
Match the model to the task
Use the biggest model for architecture, hard debugging, and anything where being wrong is expensive. Use a mid-tier model for implementation against a clear spec. Use the smallest for mechanical work: renames, boilerplate, running and reading tests. For subagents, set model: haiku in the subagent config.
# Inside a session
/model # switch model, saved as the default for new sessions
/effort low # low, medium, high, xhigh, max
# Or start pinned
claude --model sonnet --effort medium
Turn thinking down, not off
Extended thinking is on by default and its tokens bill as output tokens. For simple work, lowering the effort level with /effort is the cheap fix. On models with a fixed thinking budget you can also set MAX_THINKING_TOKENS, for example MAX_THINKING_TOKENS=8000; adaptive-reasoning models ignore nonzero budgets, so use effort levels there instead.
Clear rather than compact between unrelated tasks
/compact has to read the conversation it summarises, so compacting a large context is itself a large request. When the next thing you type is unrelated to the last thing, /clear is both cheaper and better: it costs nothing. Rename the session first if you want to find it again.
/rename tax-rounding-fix # label it before you leave
/clear # free, drops everything
/compact Focus on the API surface and test output # keeps continuity, costs a pass
/resume # come back to it later
Keep tool output small
Tell the agent to pipe noisy commands through head, run focused tests rather than whole suites, and grep with tight patterns. A PreToolUse hook that rewrites test commands to show only failures turns tens of thousands of tokens into hundreds, deterministically, without relying on the model to remember.
Move long instructions out of CLAUDE.md
CLAUDE.md loads into context at session start, every session. Detailed workflow instructions sitting there are paid for even when you are doing something unrelated. Move them into skills, which load only when invoked, and keep CLAUDE.md under about 200 lines.
Separate the accounts if you run agents in parallel
If you routinely run two or three sessions at once, one subscription is structurally the wrong shape. Two Max 5x accounts cost the same as one Max 20x and give you two independent windows.
Plan limits versus a 429
One more distinction, because the fixes are unrelated. A plan limit is a subscription window closing. A 429 is an API rate limit on a key, a Bedrock project, or a Google Cloud project.
API Error: Request rejected (429) · this may be a temporary capacity issue.
If it persists, check https://status.claude.com.
- Run
/statusand confirm the active credential is the one you expect. A strayANTHROPIC_API_KEYin your environment routes requests through a low-tier key instead of your subscription, and this is the single most common cause of a surprise 429. - Check your provider console for the active tier and request a higher one if needed.
- Reduce concurrency: lower
CLAUDE_CODE_MAX_TOOL_USE_CONCURRENCY, avoid many parallel subagents, or switch to a smaller model for high-volume scripted runs.
If it is a plan limit and you need to finish today, /usage-credits buys overflow on Pro and Max, or sends a request to your admin on Team and Enterprise. Remember the cache-lifetime cost: once you are on credits, the prompt cache drops from an hour to five minutes.
Questions people ask
What is a session limit in Claude?
The session limit is Claude Code's rolling 5-hour allowance, and it is the burst-protection ceiling rather than your total budget. It measures what you consumed over the trailing five hours, so it is not a bucket that empties at a fixed time: capacity returns gradually as earlier usage ages out. Hitting it prints "You've hit your session limit" with the reset time. It is shared across every model, so switching with /model does not restore access, and it runs alongside a separate weekly cap and a separate Opus allowance.
What should I do when Claude says I am approaching my 5-hour limit?
Treat it as the cue to change what you are spending on, not to stop working. Anthropic does not publish the percentage that triggers a warning, so the reliable signals are the ones you can read continuously: /usage inside a session, the Desktop app's usage ring, and rate_limits.five_hour.used_percentage in a custom status line. With headroom running out, drop to a smaller model for mechanical work, compact or restart a long conversation so you stop resending a large context on every turn, and hold Opus back for the part that actually needs it. Because the window is rolling, capacity starts returning on its own while you keep going at a lower burn rate.
What are the Claude Code Pro limits?
Pro at $20/mo is the baseline tier and carries the same three ceilings as every paid plan: a rolling 5-hour session window, a weekly cap, and a separate Opus allowance. Anthropic publishes the mechanism rather than numeric token figures, and expresses the higher tiers as multipliers of Pro, so Max 5x at $100 and Max 20x at $200 give you 5x and 20x the Pro allowance on both the session window and the weekly cap. Pro is comfortable for intermittent single-session work and is the tier people outgrow first by running parallel sessions, since sessions under one account share one allowance. Claude Code is not included on the free plan.
When does the Claude Code 5-hour limit reset?
The error prints the exact reset time, for example "resets 3:45pm". It is a rolling window measuring consumption over the trailing five hours, so capacity also returns gradually as earlier usage ages out rather than all at once.
Why did I hit a limit when I barely used it today?
Two likely causes. Either your trailing five hours include a heavy session from earlier, or you hit the separate weekly cap. Both counters run simultaneously, and one big fan-out can exhaust the weekly allowance before the session window resets.
Does switching to Sonnet or Haiku restore access?
Not for the session or weekly limits: those are shared across all models. It does work for the Opus-specific limit, which is why the "You have hit your Opus limit" message is the one where /model actually helps.
Do Max plans remove rate limits?
No. Max 5x and Max 20x multiply your allowance by 5x and 20x relative to Pro on both the 5-hour window and the weekly cap. The limits still exist, they just arrive later.
How do I check my Claude Code usage?
Run /usage in a session for plan bars, an attribution breakdown by skill, subagent, plugin, and MCP server, and behavior flags. For a continuous view, add rate_limits.five_hour.used_percentage and rate_limits.seven_day.used_percentage to a custom status line.
Do parallel sessions share one limit?
Yes, if they run under one account. Three simultaneous sessions consume roughly three times as fast, and agent teams run about 7x a standard session in plan mode. Separate accounts have separate windows.
Does the API have these limits?
No. API usage is metered per token with no 5-hour or weekly plan windows. It is subject to per-organisation rate limits instead, which surface as HTTP 429 rather than a plan-limit message.
Can I buy more usage without upgrading my plan?
Yes. Run /usage-credits after signing in with a claude.ai subscription. On Pro and Max it opens billing settings; on Team and Enterprise without billing access it sends a request to your admins. It is not available under API-key authentication.
What is Claude Fable 5 in Claude Code?
Fable 5 was the most capable model when these July 2026 plan rules were published; Fable 5.1 now occupies that current lineup slot. This answer retains the older model's plan history: Max and premium seats could use up to 50% of the weekly pool on Fable 5, while Pro and Team standard billed usage credits from the first request.
How do Claude Fable 5 limits work?
Fable 5 is offered on paid plans and not on Free. Starting 20 July 2026, Max, Team premium, and seat-based Enterprise premium include Fable 5 as a standard part of the plan: you can use up to 50% of the weekly usage limits on Fable 5 at no extra cost, drawn from the same weekly pool, which Fable 5 uses faster than other models. Pro and Team standard bill Fable 5 to usage credits from the first request; eligible seats qualify for a one-time credit. Seat-based Enterprise standard seats need organisation-enabled credits and do not get that promotional credit. Usage-based Enterprise and the Claude API bill at standard API rates. In Claude Code, Fable 5 needs v2.1.170 or later and /model fable; it is not the default.
Will I get 50% more weekly limit for Fable 5 on Max or premium seats?
No. You can use up to 50% of the weekly limit on Fable 5. Other models draw from the same weekly pool, and you can never use more than the weekly limit. That is a different 50% from the Claude Code weekly +50% promo that ClaudeDevs extended until 14 September 2026. Starting 14 September, they say standard weekly limits are permanently +25% vs the old baseline (a 17% cut vs the boosted week). That raise applies to Pro, Max, Team, and seat-based Enterprise and leaves the 5-hour window unchanged.
How do I use Fable 5 in Claude Code?
Update to Claude Code v2.1.170 or later, then run /model fable. Fable 5 is not the default. On plans where it bills to usage credits, the picker shows "Requires usage credits". Interactive sessions ask for one-time consent before billing those credits; Enterprise organisation billing skips the prompt. Non-interactive -p runs and the Agent SDK bill without asking.
Sources
Every figure above was read from these pages on August 2026. Vendors reprice without notice; if you find a stale number, tell us.
- Claude Code: usage limit errors the exact limit messages and what each means
- Claude Code: manage costs effectively /usage breakdown, cache lifetime, agent team costs
- Claude Code: status line reference the rate_limits JSON fields
- Claude plans and pricing plan multipliers
- ClaudeDevs on X +50% weekly until 14 September 2026, then permanent +25% vs the old baseline; 29 August 2026
- Claude Fable 5 on your plan plan split, 50% of weekly pool, credits on Pro and standard seats; read 19 August 2026
- Claude Code: model configuration /model fable, v2.1.170, usage-credit consent; read 19 August 2026