Models/Meta/Llama 3.1 8B Instruct

Llama 3.1 8B Instruct

Meta's latest class of model (Llama 3.1) launched with a variety of sizes & flavors. This 8B instruct-tuned version is fast and efficient. It has demonstrated strong performance compared to...

01

Identity

131,072Context in
131,072Context out
2024-07-23 (OpenRouter created 1721692800)Released
OpenWeights
Canonical name
Llama 3.1 8B Instruct
Aliases
meta-llama/llama-3.1-8b-instruct, meta-llama/Meta-Llama-3.1-8B-Instruct
Continuum hosted id
Not listed as a Continuum hosted id
OpenRouter slug
meta-llama/llama-3.1-8b-instruct
Hugging Face
meta-llama/Meta-Llama-3.1-8B-Instruct
Modalities
text
Open weights Not on Continuum host list Open

Context window on the 2026-08-19 OpenRouter row: 131,072 tokens.

Max completion tokens on that row: 131,072.

Input / output on that row: $0.05 / $0.08 per 1M tokens.

Architecture fields: text->text · Llama3 · llama3.

Hugging Face id on that row: meta-llama/Meta-Llama-3.1-8B-Instruct.

02

Should I use this for coding agents

Meta's latest class of model (Llama 3.1) launched with a variety of sizes & flavors.

When you want the 2024-07-23 catalog SKU, not a later rename.

No independent row for this SKU

No independent board lists this exact SKU as of 2026-08-19.

Closest benchmarked sibling: Muse Spark 1.2, 55% ±2% Pass@1 at xhigh on the DeepSWE official board. That is a score for the sibling, not for this SKU.

Boards checked: DeepSWE official (2026-08-13), Artificial Analysis Intelligence Index, vals.ai hero index. Every card that does carry a row is on the coding leaderboard.

DeepSWE

No official DeepSWE mini-swe-agent row for this identity on the public board we fetched.

omitted, not invented as of 2026-08-13 source
Artificial Analysis

No Intelligence Index integer fetched for this identity. Chip omitted.

omitted, not invented as of 2026-08-19 source
Vals Index

We did not open a vals.ai card for this identity. Chip omitted.

omitted, not invented as of 2026-08-19 source
03

When not to use it

Trust this before you buy

  • When you need a Continuum-hosted SKU from this lab, use the hosted card on this hub instead of this catalog row.
  • $0.05 in / $0.08 out per 1M on the 2026-08-19 OpenRouter row. Context 131,072 in / 131,072 out.
  • Need a denser published sibling on this hub? Llama 4 Maverick is the card already on the cluster.
  • When the job needs a published DeepSWE, Artificial Analysis, or vals.ai chip: this page does not invent one. Those boards did not list this exact SKU on 2026-08-19.

Plus is $25/mo with $25 weekly hosted usage. The Mac app stays free with your own keys. Get Plus is not Download for Mac.

04

Artifact and Hugging Face downloads

Artifact

SPDX / license
llama3.1
Total params
HF safetensors.total: 8,030,261,248 stored tensors
Native precision
HF tensors BF16
Repo size
HF API usedStorage 32,123,357,950 bytes
HF created
2024-07-18T08:56:00.000Z
Official HF repo
meta-llama/Meta-Llama-3.1-8B-Instruct
HF downloads
7,199,331
HF likes
6,628

Official weights id on the 2026-08-19 OpenRouter row: meta-llama/Meta-Llama-3.1-8B-Instruct. Community GGUF is a quant, not a second Continuum card.

Hugging Face Hub API 2026-08-19: gated=manual.

Official weights

This is the model repo, not Download for Mac. Downloads and likes from the Hugging Face API on 2026-08-19.

Hub file tree

Safetensor shards
4
Shard bytes
16,060,556,376 bytes
Other file bytes
16,072,025,947 bytes
Tree file count
17
Hub usedStorage
32,123,357,950 bytes

usedStorage and the sum of .safetensors sizes can differ (LFS pointers, non-shard files, duplicate copies). Cited separately. Source: recursive tree API, 2026-08-19.

PathBytes
model-00001-of-00004.safetensors4,976,698,672
model-00002-of-00004.safetensors4,999,802,720
model-00003-of-00004.safetensors4,915,916,176
model-00004-of-00004.safetensors1,168,138,808
original/consolidated.00.pth16,060,617,592
tokenizer.json9,085,657
original/tokenizer.model2,183,982
tokenizer_config.json55,351
README.md44,044
model.safetensors.index.json23,950
LICENSE7,627
USE_POLICY.md4,691
.gitattributes1,519
config.json855
special_tokens_map.json296
original/params.json199
generation_config.json184

All 17 files the tree API returned are listed: every one of the 4 safetensor shards plus 13 other files. Files tab: https://huggingface.co/meta-llama/Meta-Llama-3.1-8B-Instruct/tree/main.

05

Continuum serving

OpenRouter slug meta-llama/llama-3.1-8b-instruct on the 2026-08-19 catalog.

First party: https://www.llama.com/.

Not on the Continuum host list fetched 2026-08-19.

Continuum hosted id not on the host list

We list first-party and OpenRouter, plus Continuum only when the live public allowlist named the id. This is not a 15-host routing table.

06

Economics

OpenRouter 2026-08-19: $0.05 in / $0.08 out per 1M tokens. Context 131,072 in / 131,072 out.

No separate internal-reasoning price on that row.

Price this model Opens the pricing calculator preloaded with Llama 3.1 8B Instruct.
07

Provenance

ClaimSourceAs of
OpenRouter id meta-llama/llama-3.1-8b-instruct; context 131072; created 1721692800; $0.05 / $0.08 per 1MOpenRouter /api/v1/models2026-08-19
OpenRouter id meta-llama/llama-3.1-8b-instruct, context 131,072OpenRouter /api/v1/models2026-08-19
HF downloads 7,199,331Hugging Face API meta-llama/Meta-Llama-3.1-8B-Instruct2026-08-19
HF safetensors.total: 8,030,261,248 stored tensors; HF API usedStorage 32,123,357,950 bytes; created 2024-07-18T08:56:00.000Z; HF tensors BF16Hugging Face API meta-llama/Meta-Llama-3.1-8B-Instruct2026-08-19
4 safetensor shards; shard bytes 16,060,556,376; tree files 17Hugging Face tree API meta-llama/Meta-Llama-3.1-8B-Instruct2026-08-19
08

Compare, FAQ, and Get Plus

No other card in this cluster shares a published DeepSWE official row we can put next to this one. We will not compare on vendor-blog numbers.

Continue through Meta's model family with Muse Glimmer 30B.

FAQ

Why does the lab blog disagree with DeepSWE or Scale?

Lab posts pick a harness, an effort, a split, and sometimes a private eval set. DeepSWE publishes the official mini-swe-agent row with cost, tokens, and steps. Scale SWE-bench Pro only counts when the public shared-harness board has a row we can fetch. A higher lab number is usually a different test, not a better one.

OpenAI-compatible call

No Continuum hosted id is verified for this model, so this page does not invent a curl target. Use the first-party API or the OpenRouter slug meta-llama/llama-3.1-8b-instruct.

Plus is $25/mo with $25 weekly hosted usage. The Mac app stays free with your own keys. Get Plus is not Download for Mac.