Models/Xiaomi/MiMo-V2.5

MiMo-V2.5

MiMo-V2.5 is the Xiaomi week-volume model. Open MIT weights. AA Intelligence Index 38; vals hero index 39.91%. No official DeepSWE row.

01

Identity

1,050,000Context in
131,072Context out
2026-04-22 (OpenRouter created 1776874269)Released
OpenWeights
Canonical name
MiMo-V2.5
Aliases
mimo-v2.5, xiaomi/mimo-v2.5
Continuum hosted id
Not listed as a Continuum hosted id
OpenRouter slug
xiaomi/mimo-v2.5
Hugging Face
XiaomiMiMo/MiMo-V2.5
Modalities
text, image, audio, video
Open weights Not on Continuum host list Vendor lab

Not listed as a Continuum hosted id. MIT on the official repo.

HF config.json (fetched 19 Aug 2026): MiMoV2ForCausalLM, 48 layers, hidden 4096, 256 routed experts, 8 experts/token, YaRN window 1,048,576. Vision config is present. FP8 quantization_config on the Hub card. Chat template is present.

02

Should I use this for coding agents

Use it when you want Xiaomi’s open multimodal checkpoint. AA 38 and vals 39.91% are cited; we do not invent a DeepSWE percent.

DeepSWE

No official DeepSWE mini-swe-agent row for this identity on the public board we fetched.

omitted, not invented as of 2026-08-13 source
Artificial Analysispublished integer
38
bench
Artificial Analysis
version
Intelligence Index
harness
published integer
metric
Intelligence Index
value
38
independent as of 2026-08-19 source
Vals Indexvals.ai card
39.91%
bench
Vals Index
version
hero index
harness
vals.ai card
metric
Vals Index
value
39.91%
independent as of 2026-08-19 source

vals.ai MiMo-V2.5 card, fetched 19 Aug 2026: Index 39.91% ±1.31, $0.061 per Index test, 18 min 13 s latency. Those are Index-page extras, not DeepSWE, and not hero chips. No Updates prose with named non-zero benches was present on that HTML. The Accuracy Rankings bars printed 0.0% placeholders: omitted.

Ranked view: Best coding models: the independent leaderboard puts this row and every other card that carries an independent coding score on one board, with the confidence intervals left visible.

Named benches, printed only

Copied from AA or vals HTML we opened. 0.0% placeholder bars are omitted. Do not average these into the hero tiles, and do not invent CI, $/task, or steps for Artificial Analysis.

BenchPrintedNoteSourceURLAs of
AA list pair $0.14 / $0.28 per 1M Printed AA pricing sentence. Matches the OpenRouter catalog pair. Artificial Analysis MiMo-V2.5 source 2026-08-19
AA throughput 58.4 tok/s Printed Speed row. Not a coding score. Artificial Analysis MiMo-V2.5 source 2026-08-19
AA Index eval cost $25.18 Printed “it cost $25.18 to evaluate MiMo-V2.5 on the Intelligence Index.” Artificial Analysis MiMo-V2.5 source 2026-08-19
vals Index extras $0.061/test · 18 min 13 s Printed Cost / Test and Latency on the vals hero. Not DeepSWE. vals.ai MiMo-V2.5 source 2026-08-19
03

When not to use it

Trust this before you buy

  • No official DeepSWE row. AA 38 and vals 39.91% ±1.31 are different harnesses; do not average them into a fake SWE percent.
  • vals printed $0.061 per Index test and 18 min 13 s latency. Those extras are not DeepSWE $/task and not a hero chip.
  • MIT open weights. No official GGUF on this repo. Community quants are community. Cache read on OpenRouter is $0.0028/M. Do not invent a first-party Xiaomi list; the pages we opened did not print one.
  • Not a Continuum hosted id. Sibling: none carded on this hub.

Plus is $25/mo with $25 weekly hosted usage. The Mac app stays free with your own keys. Get Plus is not Download for Mac.

04

Artifact and Hugging Face downloads

Artifact

SPDX / license
MIT
Total params
HF safetensors.total: 310,775,040,000 stored tensors
Architecture
MiMoV2ForCausalLM · 48 layers · 256 routed experts · 8 experts/tok · hidden 4096 · 1,048,576 positions · vision config present
Native precision
HF tensors F8_E4M3 + BF16 + F32. Hub quantization_config quant_method fp8
Files
18 safetensor files: 16 pipeline/expert shards (model_pp0_ep0_shard0 … ep7_shard1) plus model_mtp.safetensors and audio_tokenizer/model.safetensors
Repo size
HF API usedStorage 315,693,598,475 bytes (294 GiB)
HF created
2026-04-27T13:37:38Z
Chat template
Jinja present on the Hub card, including tool and multimodal branches.
Paper
No arXiv tag on the Hugging Face API payload we fetched.
Official HF repo
XiaomiMiMo/MiMo-V2.5
HF downloads
426,872
HF likes
400

Community quants: Community quants are community.

Official weights

This is the model repo, not Download for Mac. Downloads and likes from the Hugging Face API on 2026-08-19.

Hub file tree

Safetensor shards
18
Shard bytes
315,693,004,496 bytes
Other file bytes
21,048,906 bytes
Tree file count
39
Hub usedStorage
315,693,598,475 bytes

usedStorage and the sum of .safetensors sizes can differ (LFS pointers, non-shard files, duplicate copies). Cited separately. Source: recursive tree API, 2026-08-19.

PathBytes
audio_tokenizer/model.safetensors652,622,472
model_mtp.safetensors1,189,405,960
model_pp0_ep0_shard0.safetensors34,369,159,872
model_pp0_ep0_shard1.safetensors14,463,302,008
model_pp0_ep1_shard0.safetensors34,369,162,432
model_pp0_ep1_shard1.safetensors3,490,619,024
model_pp0_ep2_shard0.safetensors34,369,162,432
model_pp0_ep2_shard1.safetensors3,490,619,024
model_pp0_ep3_shard0.safetensors34,369,169,600
model_pp0_ep3_shard1.safetensors3,490,619,752
model_pp0_ep4_shard0.safetensors34,369,170,624
model_pp0_ep4_shard1.safetensors3,490,619,856

Showing the first 12 of 18 safetensor shards. The tree also holds 21 non-shard files, none of them listed here. The snapshot was capped at 12 shards when the tree API was read on 2026-08-19, so the remaining 6 are on the Files tab, not on this page. Totals above are the full tree API sum, not this slice. Byte sizes are Hub tree API size fields, never invented.

05

Continuum serving

Not listed as a Continuum hosted id. OpenRouter xiaomi/mimo-v2.5.

Continuum hosted id not on the host list

We list first-party and OpenRouter, plus Continuum only when the live public allowlist named the id. This is not a 15-host routing table.

06

Economics

OpenRouter catalog 19 Aug 2026: $0.14 / $0.28 per million, cache read $0.0028 /M, created 1776874269, context 1,050,000 / 131,072. No DeepSWE $/task.

Artificial Analysis MiMo-V2.5 page (19 Aug 2026) prints the same $0.14 / $0.28 pair, 58.4 tok/s, and “it cost $25.18 to evaluate MiMo-V2.5 on the Intelligence Index.”

Price this model Opens the pricing calculator preloaded with MiMo-V2.5.

Artificial Analysis Intelligence Index 38, fetched 19 Aug 2026. Same page prints $0.14 / $0.28, 58.4 tok/s, and $25.18 to run the Index. The AA tile is the published integer 38. AA did not print a DeepSWE-style CI or step count we will copy onto the chip.

07

Provenance

ClaimSourceAs of
OR catalog $0.14 / $0.28, cache read $0.0028/M, 1,050,000 / 131,072, created 1776874269OpenRouter catalog2026-08-19
5.46T week tokensOpenRouter rankings This Week2026-08-19
HF safetensors.total 310,775,040,000; usedStorage 315,693,598,475; 18 safetensor files; likes 400; created 2026-04-27T13:37:38Z; MITHugging Face API XiaomiMiMo/MiMo-V2.52026-08-19
AA Index 38; printed $0.14 / $0.28; 58.4 tok/s; $25.18 to evaluate IndexArtificial Analysis MiMo-V2.52026-08-19
vals Index 39.91% ±1.31; $0.061/test; 18 min 13 svals.ai MiMo-V2.52026-08-19
OpenRouter id xiaomi/mimo-v2.5, context 1,050,000OpenRouter /api/v1/models2026-08-19
HF downloads 426,872Hugging Face API XiaomiMiMo/MiMo-V2.52026-08-19
18 safetensor shards; shard bytes 315,693,004,496; tree files 39Hugging Face tree API XiaomiMiMo/MiMo-V2.52026-08-19
Spaces API returned 2; official collection MiMo-V2.5 (4 items, 91 upvotes)Hugging Face Spaces / collections API2026-08-19
08

Compare, FAQ, and Get Plus

No other card in this cluster shares a published DeepSWE official row we can put next to this one. We will not compare on vendor-blog numbers.

Continue through Xiaomi's model family with MiMo-V2.5-Pro.

FAQ

Why no DeepSWE chip?

The official DeepSWE board updated 13 Aug 2026 did not list MiMo-V2.5. Omitting is the rule.

Why is vals 39.91% next to AA 38?

Different harnesses. We cite both integers and do not average them.

Why does the lab blog disagree with DeepSWE or Scale?

Lab posts pick a harness, an effort, a split, and sometimes a private eval set. DeepSWE publishes the official mini-swe-agent row with cost, tokens, and steps. Scale SWE-bench Pro only counts when the public shared-harness board has a row we can fetch. A higher lab number is usually a different test, not a better one.

OpenAI-compatible call

No Continuum hosted id is verified for this model, so this page does not invent a curl target. Use the first-party API or the OpenRouter slug xiaomi/mimo-v2.5.

Plus is $25/mo with $25 weekly hosted usage. The Mac app stays free with your own keys. Get Plus is not Download for Mac.