No official DeepSWE mini-swe-agent row for this identity on the public board we fetched.
Context window on the 2026-08-19 OpenRouter row: 131,072 tokens.
Max completion tokens on that row: 131,072.
Input / output on that row: $0.05 / $0.08 per 1M tokens.
Architecture fields: text->text · Llama3 · llama3.
Hugging Face id on that row: meta-llama/Meta-Llama-3.1-8B-Instruct.
Meta's latest class of model (Llama 3.1) launched with a variety of sizes & flavors.
When you want the 2024-07-23 catalog SKU, not a later rename.
No independent board lists this exact SKU as of 2026-08-19.
Closest benchmarked sibling: Muse Spark 1.2, 55% ±2% Pass@1 at xhigh on the DeepSWE official board. That is a score for the sibling, not for this SKU.
Boards checked: DeepSWE official (2026-08-13), Artificial Analysis Intelligence Index, vals.ai hero index. Every card that does carry a row is on the coding leaderboard.
No official DeepSWE mini-swe-agent row for this identity on the public board we fetched.
No Intelligence Index integer fetched for this identity. Chip omitted.
We did not open a vals.ai card for this identity. Chip omitted.
Plus is $25/mo with $25 weekly hosted usage. The Mac app stays free with your own keys. Get Plus is not Download for Mac.
Official weights id on the 2026-08-19 OpenRouter row: meta-llama/Meta-Llama-3.1-8B-Instruct. Community GGUF is a quant, not a second Continuum card.
Hugging Face Hub API 2026-08-19: gated=manual.
This is the model repo, not Download for Mac. Downloads and likes from the Hugging Face API on 2026-08-19.
usedStorage and the sum of .safetensors sizes can differ (LFS pointers, non-shard files, duplicate copies). Cited separately. Source: recursive tree API, 2026-08-19.
| Path | Bytes |
|---|---|
model-00001-of-00004.safetensors | 4,976,698,672 |
model-00002-of-00004.safetensors | 4,999,802,720 |
model-00003-of-00004.safetensors | 4,915,916,176 |
model-00004-of-00004.safetensors | 1,168,138,808 |
original/consolidated.00.pth | 16,060,617,592 |
tokenizer.json | 9,085,657 |
original/tokenizer.model | 2,183,982 |
tokenizer_config.json | 55,351 |
README.md | 44,044 |
model.safetensors.index.json | 23,950 |
LICENSE | 7,627 |
USE_POLICY.md | 4,691 |
.gitattributes | 1,519 |
config.json | 855 |
special_tokens_map.json | 296 |
original/params.json | 199 |
generation_config.json | 184 |
All 17 files the tree API returned are listed: every one of the 4 safetensor shards plus 13 other files. Files tab: https://huggingface.co/meta-llama/Meta-Llama-3.1-8B-Instruct/tree/main.
OpenRouter slug meta-llama/llama-3.1-8b-instruct on the 2026-08-19 catalog.
First party: https://www.llama.com/.
Not on the Continuum host list fetched 2026-08-19.
We list first-party and OpenRouter, plus Continuum only when the live public allowlist named the id. This is not a 15-host routing table.
OpenRouter 2026-08-19: $0.05 in / $0.08 out per 1M tokens. Context 131,072 in / 131,072 out.
No separate internal-reasoning price on that row.
| Claim | Source | As of |
|---|---|---|
| OpenRouter id meta-llama/llama-3.1-8b-instruct; context 131072; created 1721692800; $0.05 / $0.08 per 1M | OpenRouter /api/v1/models | 2026-08-19 |
| OpenRouter id meta-llama/llama-3.1-8b-instruct, context 131,072 | OpenRouter /api/v1/models | 2026-08-19 |
| HF downloads 7,199,331 | Hugging Face API meta-llama/Meta-Llama-3.1-8B-Instruct | 2026-08-19 |
| HF safetensors.total: 8,030,261,248 stored tensors; HF API usedStorage 32,123,357,950 bytes; created 2024-07-18T08:56:00.000Z; HF tensors BF16 | Hugging Face API meta-llama/Meta-Llama-3.1-8B-Instruct | 2026-08-19 |
| 4 safetensor shards; shard bytes 16,060,556,376; tree files 17 | Hugging Face tree API meta-llama/Meta-Llama-3.1-8B-Instruct | 2026-08-19 |
No other card in this cluster shares a published DeepSWE official row we can put next to this one. We will not compare on vendor-blog numbers.
Continue through Meta's model family with Muse Glimmer 30B.
Lab posts pick a harness, an effort, a split, and sometimes a private eval set. DeepSWE publishes the official mini-swe-agent row with cost, tokens, and steps. Scale SWE-bench Pro only counts when the public shared-harness board has a row we can fetch. A higher lab number is usually a different test, not a better one.
No Continuum hosted id is verified for this model, so this page does not invent a curl target. Use the first-party API or the OpenRouter slug meta-llama/llama-3.1-8b-instruct.