Models/Mistral/Mistral Small 3

Mistral Small 3

Mistral Small 3 is a 24B-parameter language model optimized for low-latency performance across common AI tasks. Released under the Apache 2.0 license, it features both pre-trained and instruction-tuned versions designed...

01

Identity

32,768Context in
16,384Context out
2025-01-30 (OpenRouter created 1738255409)Released
OpenWeights
Canonical name
Mistral Small 3
Aliases
mistralai/mistral-small-24b-instruct-2501, mistralai/Mistral-Small-24B-Instruct-2501
Continuum hosted id
Not listed as a Continuum hosted id
OpenRouter slug
mistralai/mistral-small-24b-instruct-2501
Hugging Face
mistralai/Mistral-Small-24B-Instruct-2501
Modalities
text
Open weights Not on Continuum host list Open

Context window on the 2026-08-19 OpenRouter row: 32,768 tokens.

Max completion tokens on that row: 16,384.

Input / output on that row: $0.05 / $0.08 per 1M tokens.

Architecture fields: text->text · Mistral.

Hugging Face id on that row: mistralai/Mistral-Small-24B-Instruct-2501.

02

Should I use this for coding agents

Mistral Small 3 is a 24B-parameter language model optimized for low-latency performance across common AI tasks.

When you want the 2025-01-30 catalog SKU, not a later rename.

No independent row for this SKU

No independent board lists this exact SKU as of 2026-08-19.

No other card from this lab carries an independent board row either, so there is no sibling worth pointing at.

Boards checked: DeepSWE official (2026-08-13), Artificial Analysis Intelligence Index, vals.ai hero index. Every card that does carry a row is on the coding leaderboard.

DeepSWE

No official DeepSWE mini-swe-agent row for this identity on the public board we fetched.

omitted, not invented as of 2026-08-13 source
Artificial Analysis

No Intelligence Index integer fetched for this identity. Chip omitted.

omitted, not invented as of 2026-08-19 source
Vals Index

We did not open a vals.ai card for this identity. Chip omitted.

omitted, not invented as of 2026-08-19 source
03

When not to use it

Trust this before you buy

  • When you need a Continuum-hosted SKU from this lab, use the hosted card on this hub instead of this catalog row.
  • $0.05 in / $0.08 out per 1M on the 2026-08-19 OpenRouter row. Context 32,768 in / 16,384 out.
  • Need a denser published sibling on this hub? Mistral Large 2512 is the card already on the cluster.
  • When the job needs a published DeepSWE, Artificial Analysis, or vals.ai chip: this page does not invent one. Those boards did not list this exact SKU on 2026-08-19.

Plus is $25/mo with $25 weekly hosted usage. The Mac app stays free with your own keys. Get Plus is not Download for Mac.

04

Artifact and Hugging Face downloads

Artifact

SPDX / license
apache-2.0
Total params
HF safetensors.total: 23,572,403,200 stored tensors
Native precision
HF tensors BF16
Repo size
HF API usedStorage 94,321,574,156 bytes
HF created
2025-01-28T13:30:13.000Z
Official HF repo
mistralai/Mistral-Small-24B-Instruct-2501
HF downloads
61,825
HF likes
966

Official weights id on the 2026-08-19 OpenRouter row: mistralai/Mistral-Small-24B-Instruct-2501. Community GGUF is a quant, not a second Continuum card.

Official weights

This is the model repo, not Download for Mac. Downloads and likes from the Hugging Face API on 2026-08-19.

Hub file tree

Safetensor shards
11
Shard bytes
94,289,694,896 bytes
Other file bytes
32,148,585 bytes
Tree file count
22
Hub usedStorage
94,321,574,156 bytes

usedStorage and the sum of .safetensors sizes can differ (LFS pointers, non-shard files, duplicate copies). Cited separately. Source: recursive tree API, 2026-08-19.

PathBytes
consolidated.safetensors47,144,846,024
model-00001-of-00010.safetensors4,781,571,736
model-00002-of-00010.safetensors4,781,592,784
model-00003-of-00010.safetensors4,781,592,800
model-00004-of-00010.safetensors4,886,471,600
model-00005-of-00010.safetensors4,781,592,824
model-00006-of-00010.safetensors4,781,592,816
model-00007-of-00010.safetensors4,886,471,600
model-00008-of-00010.safetensors4,781,592,824
model-00009-of-00010.safetensors4,781,592,816
model-00010-of-00010.safetensors3,900,777,072
tokenizer.json17,078,037
tekken.json14,801,223
tokenizer_config.json199,695
model.safetensors.index.json29,894
special_tokens_map.json21,311
README.md14,757
.gitattributes1,618
SYSTEM_PROMPT.txt997
config.json623
params.json270
generation_config.json160

All 22 files the tree API returned are listed: every one of the 11 safetensor shards plus 11 other files. Files tab: https://huggingface.co/mistralai/Mistral-Small-24B-Instruct-2501/tree/main.

05

Continuum serving

OpenRouter slug mistralai/mistral-small-24b-instruct-2501 on the 2026-08-19 catalog.

First party: https://docs.mistral.ai/getting-started/models/.

Not on the Continuum host list fetched 2026-08-19.

Continuum hosted id not on the host list

We list first-party and OpenRouter, plus Continuum only when the live public allowlist named the id. This is not a 15-host routing table.

06

Economics

OpenRouter 2026-08-19: $0.05 in / $0.08 out per 1M tokens. Context 32,768 in / 16,384 out.

No separate internal-reasoning price on that row.

Price this model Opens the pricing calculator preloaded with Mistral Small 3.
07

Provenance

ClaimSourceAs of
OpenRouter id mistralai/mistral-small-24b-instruct-2501; context 32768; created 1738255409; $0.05 / $0.08 per 1MOpenRouter /api/v1/models2026-08-19
OpenRouter id mistralai/mistral-small-24b-instruct-2501, context 32,768OpenRouter /api/v1/models2026-08-19
HF downloads 61,825Hugging Face API mistralai/Mistral-Small-24B-Instruct-25012026-08-19
HF safetensors.total: 23,572,403,200 stored tensors; HF API usedStorage 94,321,574,156 bytes; created 2025-01-28T13:30:13.000Z; HF tensors BF16Hugging Face API mistralai/Mistral-Small-24B-Instruct-25012026-08-19
11 safetensor shards; shard bytes 94,289,694,896; tree files 22Hugging Face tree API mistralai/Mistral-Small-24B-Instruct-25012026-08-19
08

Compare, FAQ, and Get Plus

No other card in this cluster shares a published DeepSWE official row we can put next to this one. We will not compare on vendor-blog numbers.

Continue through Mistral's model family with Mistral Large 2407.

FAQ

Why does the lab blog disagree with DeepSWE or Scale?

Lab posts pick a harness, an effort, a split, and sometimes a private eval set. DeepSWE publishes the official mini-swe-agent row with cost, tokens, and steps. Scale SWE-bench Pro only counts when the public shared-harness board has a row we can fetch. A higher lab number is usually a different test, not a better one.

OpenAI-compatible call

No Continuum hosted id is verified for this model, so this page does not invent a curl target. Use the first-party API or the OpenRouter slug mistralai/mistral-small-24b-instruct-2501.

Plus is $25/mo with $25 weekly hosted usage. The Mac app stays free with your own keys. Get Plus is not Download for Mac.