Models/Anthropic/Claude Opus 5

Claude Opus 5

Opus 5 is Anthropic’s previous Opus. Its historical card, benchmark rows, and pricing remain, but Continuum now offers Opus 5.5 for new sessions.

01

Identity

1,000,000Context in
128,000Context out
2026-07-24 (OpenRouter created)Released
ProprietaryWeights
Canonical name
Claude Opus 5
Aliases
opus-5, claude-opus-5, continuum/claude-opus-5
Continuum hosted id
Not listed as a Continuum hosted id
OpenRouter slug
anthropic/claude-opus-5
Hugging Face
None (closed or unpublished)
Modalities
text, image, file
Plan mapping
Claude Max / Team / API. Delisted from Continuum new-session menus on 22 Sep 2026 in favor of Opus 5.5; the opus alias now selects Opus 5.5.
Closed weights Not on Continuum host list Vendor lab

Same 1M / 128k window as Fable 5 and Sonnet 5 on the OpenRouter catalog. Effort ladder low through max, default high.

02

Should I use this for coding agents

Use Opus when Sonnet already failed, or when the decision is architectural. DeepSWE official puts Opus at 74% Pass@1 ±4% at max, $11.84 per task, 118k output, 99 steps: the top row on that board the day we fetched it.

DeepSWEmini-swe-agent
74%±4%
bench
DeepSWE
version
v1.1
split
public 113 tasks
harness
mini-swe-agent
effort
max
n
113
metric
Pass@1
value
74%
ci
±4%
$/task
$11.84
tokens
118k out
steps
99
independent as of 2026-08-13 source
Artificial Analysispublished integer
63
bench
Artificial Analysis
version
Intelligence Index
harness
published integer
metric
Intelligence Index
value
63
independent as of 2026-08-19 source
Vals Indexvals.ai card
67.21%
bench
Vals Index
version
hero index
harness
vals.ai card
metric
Vals Index
value
67.21%
independent as of 2026-08-19 source

Ranked view: Best coding models: the independent leaderboard puts this row and every other card that carries an independent coding score on one board, with the confidence intervals left visible.

03

When not to use it

Trust this before you buy

  • Do not leave Opus selected for mechanical edits. Sonnet 5 at $2 / $10 finishes most coding at a fraction of the quota burn.
  • Sibling: Fable 5.1 if the run must stay coherent across sittings. Sonnet 5 is the working default. Haiku 4.5 (not carded here) for fully specified mechanical work.
  • Closed weights. Anthropic’s data-retention policy does not allow zero data retention. You cannot air-gap the parameters.
  • Official DeepSWE row is max. Default effort is high. If you have not tried Sonnet at high, do not start here.

Get Plus for the current Continuum-hosted Opus 5.5 replacement. Not a Mac download.

04

Artifact

Weights

Anthropic does not publish Opus 5 weights. No official HF repo. Get Plus is hosted inference.

  • Official Anthropic pricing.md, fetched 19 Aug 2026: Opus 5 $5 input / $6.25 5-minute cache write / $10 1-hour cache write / $0.50 cache hits / $25 output per million.
  • OpenRouter catalog 19 Aug 2026: $5 / $25, cache read $0.50 /M, cache write $6.25 /M, 1-hour cache write $10 /M, web search $0.01 per call. Same table as official, plus the search field.
  • Anthropic’s 22 Sep 2026 models overview says to start with Opus 5.5 for most workloads and names Fable 5.1 as the highest available capability. That is vendor guidance, not a DeepSWE ranking.

What you can open

Lab card and OpenRouter listing only. No Hugging Face Files button, because there is no official repo to point at.

Fetched OpenRouter catalog on 2026-08-19. hugging_face_id was empty.

05

Continuum serving

Continuum hosted claude-opus-5 in the 19 Aug 2026 snapshot. It is now delisted from new-session menus in favor of claude-opus-5-5; it was retired from Continuum inference on 22 Sep 2026. The old rate stays for history, and Opus 5 still runs on your own Claude subscription.

Continuum hosted id not on the host list

We list first-party and OpenRouter, plus Continuum only when the live public allowlist named the id. This is not a 15-host routing table.

06

Economics

DeepSWE official: $11.84 per task at max (118k out, 99 steps). That is the hero economics number.

Official Anthropic + OpenRouter catalog, 19 Aug 2026: $5 / $25 per million. Cache read $0.50 /M. Cache write $6.25 /M (1-hour write $10 /M). Web search $0.01 per call on the catalog. Sonnet 5 is $2 / $10. Fable 5.1 is $10 / $50 with $0.25 /M cache reads.

Price this model Opens the pricing calculator preloaded with Claude Opus 5.

On a Claude plan? The plan picker models the weekly caps.

Artificial Analysis footnote, 19 Aug 2026: about $2.34 per AA task, 56.9 tok/s. The AA tile is Intelligence Index 63. AA did not publish a DeepSWE-style CI or step count we will copy onto the chip. Do not average AA cost with DeepSWE $11.84/task.

07

Provenance

ClaimSourceAs of
DeepSWE 74% ±4% at maxDeepSWE official board2026-08-13
Official Opus 5 $5 / $6.25 5m write / $10 1h write / $0.50 hits / $25 outAnthropic pricing.md2026-08-19
OR catalog $5 / $25, cache read $0.50, write $6.25, 1h write $10, search $0.01OpenRouter catalog2026-08-19
Hosted id claude-opus-5 (delisted from new-session menus on 22 Sep 2026; the current Opus is claude-opus-5-5)GET /v1/chat/hosted/models/public2026-09-22
2.68T week tokens (+89% WoW)OpenRouter rankings This Week2026-08-19
OpenRouter id anthropic/claude-opus-5, context 1,000,000OpenRouter /api/v1/models2026-08-19
08

Compare, FAQ, and Get Plus

Same official DeepSWE harness (mini-swe-agent, public 113 tasks), fetched 2026-08-19. Different effort labels are the lab's own setting on that board, shown here rather than normalized.

Continue through Anthropic's model family with Claude Sonnet 5.

FAQ

Does Opus 5.5 replace Opus 5?

Yes for new sessions. Continuum delists Opus 5 from every picker and retired it from Continuum inference on 22 Sep 2026. Its rate stays so historic usage prices correctly, and it still runs on your own Claude subscription. The benchmark rows on this page are for Opus 5, not Opus 5.5.

Is Opus worth it over Sol?

On DeepSWE official they are inside each other’s confidence interval (74% ±4 vs 73% ±3) and Opus costs more per task ($11.84 vs $8.39). Pick on product fit and which subscription you already pay for, not on a one-point gap.

Why does the lab blog disagree with DeepSWE or Scale?

Lab posts pick a harness, an effort, a split, and sometimes a private eval set. DeepSWE publishes the official mini-swe-agent row with cost, tokens, and steps. Scale SWE-bench Pro only counts when the public shared-harness board has a row we can fetch. A higher lab number is usually a different test, not a better one.

OpenAI-compatible call

No Continuum hosted id is verified for this model, so this page does not invent a curl target. Use the first-party API or the OpenRouter slug anthropic/claude-opus-5.

Get Plus for the current Continuum-hosted Opus 5.5 replacement. Not a Mac download.