Models/Meta/Llama 3.1 70B Instruct

Llama 3.1 70B Instruct

Meta's latest class of model (Llama 3.1) launched with a variety of sizes & flavors. This 70B instruct-tuned version is optimized for high quality dialogue usecases. It has demonstrated strong...

01

Identity

131,072Context in
16,384Context out
2024-07-23 (OpenRouter created 1721692800)Released
OpenWeights
Canonical name
Llama 3.1 70B Instruct
Aliases
meta-llama/llama-3.1-70b-instruct, meta-llama/Meta-Llama-3.1-70B-Instruct
Continuum hosted id
Not listed as a Continuum hosted id
OpenRouter slug
meta-llama/llama-3.1-70b-instruct
Hugging Face
meta-llama/Meta-Llama-3.1-70B-Instruct
Modalities
text
Open weights Not on Continuum host list Open

Context window on the 2026-08-19 OpenRouter row: 131,072 tokens.

Max completion tokens on that row: 16,384.

Input / output on that row: $0.40 / $0.40 per 1M tokens.

Architecture fields: text->text · Llama3 · llama3.

Hugging Face id on that row: meta-llama/Meta-Llama-3.1-70B-Instruct.

02

Should I use this for coding agents

Meta's latest class of model (Llama 3.1) launched with a variety of sizes & flavors.

When you want the 2024-07-23 catalog SKU, not a later rename.

No independent row for this SKU

No independent board lists this exact SKU as of 2026-08-19.

Closest benchmarked sibling: Muse Spark 1.2, 55% ±2% Pass@1 at xhigh on the DeepSWE official board. That is a score for the sibling, not for this SKU.

Boards checked: DeepSWE official (2026-08-13), Artificial Analysis Intelligence Index, vals.ai hero index. Every card that does carry a row is on the coding leaderboard.

DeepSWE

No official DeepSWE mini-swe-agent row for this identity on the public board we fetched.

omitted, not invented as of 2026-08-13 source
Artificial Analysis

No Intelligence Index integer fetched for this identity. Chip omitted.

omitted, not invented as of 2026-08-19 source
Vals Index

We did not open a vals.ai card for this identity. Chip omitted.

omitted, not invented as of 2026-08-19 source
03

When not to use it

Trust this before you buy

  • When you need a Continuum-hosted SKU from this lab, use the hosted card on this hub instead of this catalog row.
  • $0.40 in / $0.40 out per 1M on the 2026-08-19 OpenRouter row. Context 131,072 in / 16,384 out.
  • Need a denser published sibling on this hub? Llama 4 Maverick is the card already on the cluster.
  • When the job needs a published DeepSWE, Artificial Analysis, or vals.ai chip: this page does not invent one. Those boards did not list this exact SKU on 2026-08-19.

Plus is $25/mo with $25 weekly hosted usage. The Mac app stays free with your own keys. Get Plus is not Download for Mac.

04

Artifact and Hugging Face downloads

Artifact

SPDX / license
llama3.1
Total params
HF safetensors.total: 70,553,706,496 stored tensors
Native precision
HF tensors BF16
Repo size
HF API usedStorage 282,237,450,046 bytes
HF created
2024-07-16T16:07:46.000Z
Official HF repo
meta-llama/Meta-Llama-3.1-70B-Instruct
HF downloads
766,891
HF likes
949

Official weights id on the 2026-08-19 OpenRouter row: meta-llama/Meta-Llama-3.1-70B-Instruct. Community GGUF is a quant, not a second Continuum card.

Hugging Face Hub API 2026-08-19: gated=manual.

Official weights

This is the model repo, not Download for Mac. Downloads and likes from the Hugging Face API on 2026-08-19.

Hub file tree

Safetensor shards
30
Shard bytes
141,107,497,872 bytes
Other file bytes
141,139,213,008 bytes
Tree file count
50
Hub usedStorage
282,237,450,046 bytes

usedStorage and the sum of .safetensors sizes can differ (LFS pointers, non-shard files, duplicate copies). Cited separately. Source: recursive tree API, 2026-08-19.

PathBytes
model-00001-of-00030.safetensors4,584,408,808
model-00002-of-00030.safetensors4,664,167,376
model-00003-of-00030.safetensors4,999,711,704
model-00004-of-00030.safetensors4,966,157,032
model-00005-of-00030.safetensors4,664,134,408
model-00006-of-00030.safetensors4,664,167,408
model-00007-of-00030.safetensors4,664,167,408
model-00008-of-00030.safetensors4,999,711,728
model-00009-of-00030.safetensors4,966,157,056
model-00010-of-00030.safetensors4,664,134,408
model-00011-of-00030.safetensors4,664,167,408
model-00012-of-00030.safetensors4,664,167,408
model-00013-of-00030.safetensors4,999,711,728
model-00014-of-00030.safetensors4,966,157,056
model-00015-of-00030.safetensors4,664,134,408
model-00016-of-00030.safetensors4,664,167,408
model-00017-of-00030.safetensors4,664,167,408
model-00018-of-00030.safetensors4,999,711,728
model-00019-of-00030.safetensors4,966,157,056
model-00020-of-00030.safetensors4,664,134,408
model-00021-of-00030.safetensors4,664,167,408
model-00022-of-00030.safetensors4,664,167,408
model-00023-of-00030.safetensors4,999,711,728
model-00024-of-00030.safetensors4,966,157,056
model-00025-of-00030.safetensors4,664,134,408
model-00026-of-00030.safetensors4,664,167,408
model-00027-of-00030.safetensors4,664,167,408
model-00028-of-00030.safetensors4,999,711,728
model-00029-of-00030.safetensors4,966,173,536
model-00030-of-00030.safetensors2,101,346,432
original/consolidated.00.pth17,640,971,024
original/consolidated.01.pth17,640,971,024
original/consolidated.02.pth17,640,971,024
original/consolidated.03.pth17,640,971,024
original/consolidated.04.pth17,640,971,024
original/consolidated.05.pth17,640,971,024
original/consolidated.06.pth17,640,971,024
original/consolidated.07.pth17,640,971,024
tokenizer.json9,085,657
original/tokenizer.model2,183,982
model.safetensors.index.json59,615
tokenizer_config.json55,351
README.md44,841
LICENSE7,627
USE_POLICY.md4,691
.gitattributes1,519
config.json855
special_tokens_map.json296
original/params.json199
generation_config.json183

All 50 files the tree API returned are listed: every one of the 30 safetensor shards plus 20 other files. Files tab: https://huggingface.co/meta-llama/Meta-Llama-3.1-70B-Instruct/tree/main.

05

Continuum serving

OpenRouter slug meta-llama/llama-3.1-70b-instruct on the 2026-08-19 catalog.

First party: https://www.llama.com/.

Not on the Continuum host list fetched 2026-08-19.

Continuum hosted id not on the host list

We list first-party and OpenRouter, plus Continuum only when the live public allowlist named the id. This is not a 15-host routing table.

06

Economics

OpenRouter 2026-08-19: $0.40 in / $0.40 out per 1M tokens. Context 131,072 in / 16,384 out.

No separate internal-reasoning price on that row.

Price this model Opens the pricing calculator preloaded with Llama 3.1 70B Instruct.
07

Provenance

ClaimSourceAs of
OpenRouter id meta-llama/llama-3.1-70b-instruct; context 131072; created 1721692800; $0.40 / $0.40 per 1MOpenRouter /api/v1/models2026-08-19
OpenRouter id meta-llama/llama-3.1-70b-instruct, context 131,072OpenRouter /api/v1/models2026-08-19
HF downloads 766,891Hugging Face API meta-llama/Meta-Llama-3.1-70B-Instruct2026-08-19
HF safetensors.total: 70,553,706,496 stored tensors; HF API usedStorage 282,237,450,046 bytes; created 2024-07-16T16:07:46.000Z; HF tensors BF16Hugging Face API meta-llama/Meta-Llama-3.1-70B-Instruct2026-08-19
30 safetensor shards; shard bytes 141,107,497,872; tree files 50Hugging Face tree API meta-llama/Meta-Llama-3.1-70B-Instruct2026-08-19
08

Compare, FAQ, and Get Plus

No other card in this cluster shares a published DeepSWE official row we can put next to this one. We will not compare on vendor-blog numbers.

Continue through Meta's model family with Muse Spark 1.2.

FAQ

Why does the lab blog disagree with DeepSWE or Scale?

Lab posts pick a harness, an effort, a split, and sometimes a private eval set. DeepSWE publishes the official mini-swe-agent row with cost, tokens, and steps. Scale SWE-bench Pro only counts when the public shared-harness board has a row we can fetch. A higher lab number is usually a different test, not a better one.

OpenAI-compatible call

No Continuum hosted id is verified for this model, so this page does not invent a curl target. Use the first-party API or the OpenRouter slug meta-llama/llama-3.1-70b-instruct.

Plus is $25/mo with $25 weekly hosted usage. The Mac app stays free with your own keys. Get Plus is not Download for Mac.