Models/Qwen/Qwen3.8 2.4T A95B

Qwen3.8 2.4T A95B

Qwen3.8 2.4T A95B is an open-weight sparse mixture-of-experts model from Qwen and the open-weight variant of [Qwen3.8 Max](/qwen/qwen3.8-max), with 95 billion active parameters out of 2.4 trillion total. It is...

01

Identity

1,048,576Context in
262,144Context out
2026-08-12 (OpenRouter created 1786551702)Released
OpenWeights
Canonical name
Qwen3.8 2.4T A95B
Aliases
qwen/qwen3.8-2.4t-a95b, Qwen/Qwen3.8-2.4T-A95B
Continuum hosted id
Not listed as a Continuum hosted id
OpenRouter slug
qwen/qwen3.8-2.4t-a95b
Hugging Face
Qwen/Qwen3.8-2.4T-A95B
Modalities
text
Open weights Not on Continuum host list Open

Context window on the 2026-08-19 OpenRouter row: 1,048,576 tokens.

Max completion tokens on that row: 262,144.

Input / output on that row: $2.00 / $6.00 per 1M tokens.

Architecture fields: text->text · Qwen.

Hugging Face id on that row: Qwen/Qwen3.8-2.4T-A95B.

02

Should I use this for coding agents

Qwen3.8 2.4T A95B is an open-weight sparse mixture-of-experts model from Qwen and the open-weight variant of [Qwen3.8 Max](/qwen/qwen3.8-max), with 95 billion active parameters out of 2.4 trillion total.

When you want the 2026-08-12 catalog SKU, not a later rename.

DeepSWE

No official DeepSWE mini-swe-agent row for this identity on the public board we fetched.

omitted, not invented as of 2026-08-13 source
Artificial Analysispublished integer
58
bench
Artificial Analysis
version
Intelligence Index
harness
published integer
metric
Intelligence Index
value
58
independent as of 2026-08-19 source
Vals Index

We did not open a vals.ai card for this identity. Chip omitted.

omitted, not invented as of 2026-08-19 source

Ranked view: Best coding models: the independent leaderboard puts this row and every other card that carries an independent coding score on one board, with the confidence intervals left visible.

Named benches, printed only

Copied from AA or vals HTML we opened. 0.0% placeholder bars are omitted. Do not average these into the hero tiles, and do not invent CI, $/task, or steps for Artificial Analysis.

BenchPrintedNoteSourceURLAs of
GDPval-AA v2 61% Printed on the AA Intelligence Evaluations grid for Qwen3.8 2.4T A95B. Not a DeepSWE chip. Artificial Analysis Qwen3.8 2.4T A95B source 2026-08-19
τ³-Banking 49.1% Printed on the AA Intelligence Evaluations grid for Qwen3.8 2.4T A95B. Not a DeepSWE chip. Artificial Analysis Qwen3.8 2.4T A95B source 2026-08-19
Terminal-Bench v2.1 82% Printed on the AA Intelligence Evaluations grid for Qwen3.8 2.4T A95B. Not a DeepSWE chip. Artificial Analysis Qwen3.8 2.4T A95B source 2026-08-19
SciCode 51.6% Printed on the AA Intelligence Evaluations grid for Qwen3.8 2.4T A95B. Not a DeepSWE chip. Artificial Analysis Qwen3.8 2.4T A95B source 2026-08-19
Humanity's Last Exam 42.4% Printed on the AA Intelligence Evaluations grid for Qwen3.8 2.4T A95B. Not a DeepSWE chip. Artificial Analysis Qwen3.8 2.4T A95B source 2026-08-19
GPQA Diamond 93.5% Printed on the AA Intelligence Evaluations grid for Qwen3.8 2.4T A95B. Not a DeepSWE chip. Artificial Analysis Qwen3.8 2.4T A95B source 2026-08-19
CritPt 20% Printed on the AA Intelligence Evaluations grid for Qwen3.8 2.4T A95B. Not a DeepSWE chip. Artificial Analysis Qwen3.8 2.4T A95B source 2026-08-19
AA-Omniscience Accuracy 31.3% Printed on the AA Intelligence Evaluations grid for Qwen3.8 2.4T A95B. Not a DeepSWE chip. Artificial Analysis Qwen3.8 2.4T A95B source 2026-08-19
AA-LCR 75.3% Printed on the AA Intelligence Evaluations grid for Qwen3.8 2.4T A95B. Not a DeepSWE chip. Artificial Analysis Qwen3.8 2.4T A95B source 2026-08-19
03

When not to use it

Trust this before you buy

  • When you need a Continuum-hosted SKU from this lab, use the hosted card on this hub instead of this catalog row.
  • $2.00 in / $6.00 out per 1M on the 2026-08-19 OpenRouter row. Context 1,048,576 in / 262,144 out.
  • Need a denser published sibling on this hub? Qwen3.8 Max is the card already on the cluster.

Plus is $25/mo with $25 weekly hosted usage. The Mac app stays free with your own keys. Get Plus is not Download for Mac.

04

Artifact and Hugging Face downloads

Artifact

SPDX / license
other
Total params
HF safetensors.total: 2,446,182,725,504 stored tensors
Native precision
HF tensors BF16
Repo size
HF API usedStorage 4,892,378,458,656 bytes
HF created
2026-08-08T01:50:52.000Z
Official HF repo
Qwen/Qwen3.8-2.4T-A95B
HF downloads
12,699
HF likes
1,092

Official weights id on the 2026-08-19 OpenRouter row: Qwen/Qwen3.8-2.4T-A95B. Community GGUF is a quant, not a second Continuum card.

Official weights

This is the model repo, not Download for Mac. Downloads and likes from the Hugging Face API on 2026-08-19.

Hub file tree

Safetensor shards
213
Shard bytes
4,892,365,649,336 bytes
Other file bytes
23,091,916 bytes
Tree file count
224
Hub usedStorage
4,892,378,458,656 bytes

usedStorage and the sum of .safetensors sizes can differ (LFS pointers, non-shard files, duplicate copies). Cited separately. Source: recursive tree API, 2026-08-19.

PathBytes
model-00001-of-00213.safetensors34,359,738,528
model-00002-of-00213.safetensors17,179,869,336
model-00003-of-00213.safetensors34,359,738,528
model-00004-of-00213.safetensors17,179,869,336
model-00005-of-00213.safetensors34,359,738,528
model-00006-of-00213.safetensors17,179,869,336
model-00007-of-00213.safetensors34,359,738,528
model-00008-of-00213.safetensors17,179,869,336
model-00009-of-00213.safetensors34,359,738,528
model-00010-of-00213.safetensors17,179,869,336
model-00011-of-00213.safetensors3,985,468,128
model-00012-of-00213.safetensors34,359,738,528
model-00013-of-00213.safetensors17,179,869,336
model-00014-of-00213.safetensors34,359,738,528
model-00015-of-00213.safetensors17,179,869,336
model-00016-of-00213.safetensors34,359,738,528
model-00017-of-00213.safetensors17,179,869,336
model-00018-of-00213.safetensors34,359,738,528
model-00019-of-00213.safetensors17,179,869,336
model-00020-of-00213.safetensors3,972,721,928
model-00021-of-00213.safetensors34,359,738,528
model-00022-of-00213.safetensors17,179,869,336
model-00023-of-00213.safetensors34,359,738,528
model-00024-of-00213.safetensors17,179,869,336
model-00025-of-00213.safetensors4,068,475,024
model-00026-of-00213.safetensors2,810,632,576
model-00027-of-00213.safetensors34,359,738,528
model-00028-of-00213.safetensors17,179,869,336
model-00029-of-00213.safetensors34,359,738,528
model-00030-of-00213.safetensors17,179,869,336
model-00031-of-00213.safetensors34,359,738,528
model-00032-of-00213.safetensors17,179,869,336
model-00033-of-00213.safetensors34,359,738,528
model-00034-of-00213.safetensors17,179,869,336
model-00035-of-00213.safetensors34,359,738,528
model-00036-of-00213.safetensors17,179,869,336
model-00037-of-00213.safetensors3,981,110,080
model-00038-of-00213.safetensors34,359,738,528
model-00039-of-00213.safetensors17,179,869,336
model-00040-of-00213.safetensors34,359,738,528
model-00041-of-00213.safetensors17,179,869,336
model-00042-of-00213.safetensors34,359,738,528
model-00043-of-00213.safetensors17,179,869,336
model-00044-of-00213.safetensors34,359,738,528
model-00045-of-00213.safetensors17,179,869,336
model-00046-of-00213.safetensors3,939,167,440
model-00047-of-00213.safetensors34,359,738,528
model-00048-of-00213.safetensors17,179,869,336
model-00049-of-00213.safetensors34,359,738,528
model-00050-of-00213.safetensors17,179,869,336
model-00051-of-00213.safetensors2,810,632,616
model-00052-of-00213.safetensors34,359,738,528
model-00053-of-00213.safetensors17,179,869,336
model-00054-of-00213.safetensors34,359,738,528
model-00055-of-00213.safetensors17,179,869,336
model-00056-of-00213.safetensors34,359,738,528
model-00057-of-00213.safetensors17,179,869,336
model-00058-of-00213.safetensors34,359,738,528
model-00059-of-00213.safetensors17,179,869,336
model-00060-of-00213.safetensors3,909,971,448
model-00061-of-00213.safetensors34,359,738,528
model-00062-of-00213.safetensors17,179,869,336
model-00063-of-00213.safetensors34,359,738,528
model-00064-of-00213.safetensors17,179,869,336
tokenizer.json12,809,320
vocab.json6,722,759
merges.txt3,353,259
model.safetensors.index.json137,777
README.md35,755
tokenizer_config.json16,438
chat_template.jinja7,495
config.json3,951
LICENSE3,390
.gitattributes1,570
generation_config.json202

Showing the first 64 of 213 safetensor shards, plus 11 other files. The snapshot was capped at 64 shards when the tree API was read on 2026-08-19, so the remaining 149 are on the Files tab, not on this page. Totals above are the full tree API sum, not this slice. Byte sizes are Hub tree API size fields, never invented.

05

Continuum serving

OpenRouter slug qwen/qwen3.8-2.4t-a95b on the 2026-08-19 catalog.

First party: https://qwen.ai/.

Not on the Continuum host list fetched 2026-08-19.

Continuum hosted id not on the host list

We list first-party and OpenRouter, plus Continuum only when the live public allowlist named the id. This is not a 15-host routing table.

06

Economics

OpenRouter 2026-08-19: $2.00 in / $6.00 out per 1M tokens. Context 1,048,576 in / 262,144 out.

No separate internal-reasoning price on that row.

Price this model Opens the pricing calculator preloaded with Qwen3.8 2.4T A95B.
07

Provenance

ClaimSourceAs of
OpenRouter id qwen/qwen3.8-2.4t-a95b; context 1048576; created 1786551702; $2.00 / $6.00 per 1MOpenRouter /api/v1/models2026-08-19
AA Intelligence Index 58; 9 printed Index benchesArtificial Analysis model page2026-08-19
OpenRouter id qwen/qwen3.8-2.4t-a95b, context 1,048,576OpenRouter /api/v1/models2026-08-19
HF downloads 12,699Hugging Face API Qwen/Qwen3.8-2.4T-A95B2026-08-19
HF safetensors.total: 2,446,182,725,504 stored tensors; HF API usedStorage 4,892,378,458,656 bytes; created 2026-08-08T01:50:52.000Z; HF tensors BF16Hugging Face API Qwen/Qwen3.8-2.4T-A95B2026-08-19
213 safetensor shards; shard bytes 4,892,365,649,336; tree files 224Hugging Face tree API Qwen/Qwen3.8-2.4T-A95B2026-08-19
08

Compare, FAQ, and Get Plus

No other card in this cluster shares a published DeepSWE official row we can put next to this one. We will not compare on vendor-blog numbers.

Continue through Qwen's model family with Qwen3.7 Flash.

FAQ

Why does the lab blog disagree with DeepSWE or Scale?

Lab posts pick a harness, an effort, a split, and sometimes a private eval set. DeepSWE publishes the official mini-swe-agent row with cost, tokens, and steps. Scale SWE-bench Pro only counts when the public shared-harness board has a row we can fetch. A higher lab number is usually a different test, not a better one.

OpenAI-compatible call

No Continuum hosted id is verified for this model, so this page does not invent a curl target. Use the first-party API or the OpenRouter slug qwen/qwen3.8-2.4t-a95b.

Plus is $25/mo with $25 weekly hosted usage. The Mac app stays free with your own keys. Get Plus is not Download for Mac.