Models/Moonshot/Kimi K2 0711

Kimi K2 0711

Kimi K2 Instruct is a large-scale Mixture-of-Experts (MoE) language model developed by Moonshot AI, featuring 1 trillion total parameters with 32 billion active per forward pass. It is optimized for...

01

Identity

131,072Context in
100,352Context out
2025-07-11 (OpenRouter created 1752263252)Released
OpenWeights
Canonical name
Kimi K2 0711
Aliases
moonshotai/kimi-k2, moonshotai/Kimi-K2-Instruct
Continuum hosted id
Not listed as a Continuum hosted id
OpenRouter slug
moonshotai/kimi-k2
Hugging Face
moonshotai/Kimi-K2-Instruct
Modalities
text
Open weights Not on Continuum host list Open

Context window on the 2026-08-19 OpenRouter row: 131,072 tokens.

Max completion tokens on that row: 100,352.

Input / output on that row: $0.57 / $2.30 per 1M tokens.

Architecture fields: text->text · Other.

Hugging Face id on that row: moonshotai/Kimi-K2-Instruct.

02

Should I use this for coding agents

Kimi K2 Instruct is a large-scale Mixture-of-Experts (MoE) language model developed by Moonshot AI, featuring 1 trillion total parameters with 32 billion active per forward pass.

When you want the 2025-07-11 catalog SKU, not a later rename.

DeepSWE

No official DeepSWE mini-swe-agent row for this identity on the public board we fetched.

omitted, not invented as of 2026-08-13 source
Artificial Analysispublished integer
20
bench
Artificial Analysis
version
Intelligence Index
harness
published integer
metric
Intelligence Index
value
20
independent as of 2026-08-19 source
Vals Index

We did not open a vals.ai card for this identity. Chip omitted.

omitted, not invented as of 2026-08-19 source

Ranked view: Best coding models: the independent leaderboard puts this row and every other card that carries an independent coding score on one board, with the confidence intervals left visible.

Named benches, printed only

Copied from AA or vals HTML we opened. 0.0% placeholder bars are omitted. Do not average these into the hero tiles, and do not invent CI, $/task, or steps for Artificial Analysis.

BenchPrintedNoteSourceURLAs of
SciCode 34.5% Printed on the AA Intelligence Evaluations grid for Kimi K2. Not a DeepSWE chip. Artificial Analysis Kimi K2 source 2026-08-19
Humanity's Last Exam 7.4% Printed on the AA Intelligence Evaluations grid for Kimi K2. Not a DeepSWE chip. Artificial Analysis Kimi K2 source 2026-08-19
GPQA Diamond 76.6% Printed on the AA Intelligence Evaluations grid for Kimi K2. Not a DeepSWE chip. Artificial Analysis Kimi K2 source 2026-08-19
AA-Omniscience Accuracy 27.4% Printed on the AA Intelligence Evaluations grid for Kimi K2. Not a DeepSWE chip. Artificial Analysis Kimi K2 source 2026-08-19
AA-LCR 53% Printed on the AA Intelligence Evaluations grid for Kimi K2. Not a DeepSWE chip. Artificial Analysis Kimi K2 source 2026-08-19
03

When not to use it

Trust this before you buy

  • When you need a Continuum-hosted SKU from this lab, use the hosted card on this hub instead of this catalog row.
  • $0.57 in / $2.30 out per 1M on the 2026-08-19 OpenRouter row. Context 131,072 in / 100,352 out.
  • Need a denser published sibling on this hub? Kimi K3 is the card already on the cluster.

Plus is $25/mo with $25 weekly hosted usage. The Mac app stays free with your own keys. Get Plus is not Download for Mac.

04

Artifact and Hugging Face downloads

Artifact

SPDX / license
other
Total params
HF safetensors.total: 1,026,408,235,864 stored tensors
Native precision
HF tensors F32 + BF16 + F8_E4M3
Repo size
HF API usedStorage 1,029,226,387,740 bytes
HF created
2025-07-11T00:55:12.000Z
Official HF repo
moonshotai/Kimi-K2-Instruct
HF downloads
182,842
HF likes
2,374

Official weights id on the 2026-08-19 OpenRouter row: moonshotai/Kimi-K2-Instruct. Community GGUF is a quant, not a second Continuum card.

Official weights

This is the model repo, not Download for Mac. Downloads and likes from the Hugging Face API on 2026-08-19.

Hub file tree

Safetensor shards
61
Shard bytes
1,029,190,981,272 bytes
Other file bytes
16,194,861 bytes
Tree file count
83
Hub usedStorage
1,029,226,387,740 bytes

usedStorage and the sum of .safetensors sizes can differ (LFS pointers, non-shard files, duplicate copies). Cited separately. Source: recursive tree API, 2026-08-19.

PathBytes
model-1-of-61.safetensors2,846,451,040
model-10-of-61.safetensors17,066,593,104
model-11-of-61.safetensors17,066,595,432
model-12-of-61.safetensors17,066,595,432
model-13-of-61.safetensors17,066,595,432
model-14-of-61.safetensors17,066,595,432
model-15-of-61.safetensors17,066,595,432
model-16-of-61.safetensors17,066,595,432
model-17-of-61.safetensors17,066,595,432
model-18-of-61.safetensors17,066,595,432
model-19-of-61.safetensors17,066,595,432
model-2-of-61.safetensors17,066,593,104
model-20-of-61.safetensors17,066,595,432
model-21-of-61.safetensors17,066,595,432
model-22-of-61.safetensors17,066,595,432
model-23-of-61.safetensors17,066,595,432
model-24-of-61.safetensors17,066,595,432
model-25-of-61.safetensors17,066,595,432
model-26-of-61.safetensors17,066,595,432
model-27-of-61.safetensors17,066,595,432
model-28-of-61.safetensors17,066,595,432
model-29-of-61.safetensors17,066,595,432
model-3-of-61.safetensors17,066,593,104
model-30-of-61.safetensors17,066,595,432
model-31-of-61.safetensors17,066,595,432
model-32-of-61.safetensors17,066,595,432
model-33-of-61.safetensors17,066,595,432
model-34-of-61.safetensors17,066,595,432
model-35-of-61.safetensors17,066,595,432
model-36-of-61.safetensors17,066,595,432
model-37-of-61.safetensors17,066,595,432
model-38-of-61.safetensors17,066,595,432
model-39-of-61.safetensors17,066,595,432
model-4-of-61.safetensors17,066,593,104
model-40-of-61.safetensors17,066,595,432
model-41-of-61.safetensors17,066,595,432
model-42-of-61.safetensors17,066,595,432
model-43-of-61.safetensors17,066,595,432
model-44-of-61.safetensors17,066,595,432
model-45-of-61.safetensors17,066,595,432
model-46-of-61.safetensors17,066,595,432
model-47-of-61.safetensors17,066,595,432
model-48-of-61.safetensors17,066,595,432
model-49-of-61.safetensors17,066,595,432
model-5-of-61.safetensors17,066,593,104
model-50-of-61.safetensors17,066,595,432
model-51-of-61.safetensors17,066,595,432
model-52-of-61.safetensors17,066,595,432
model-53-of-61.safetensors17,066,595,432
model-54-of-61.safetensors17,066,595,432
model-55-of-61.safetensors17,066,595,432
model-56-of-61.safetensors17,066,595,432
model-57-of-61.safetensors17,066,595,432
model-58-of-61.safetensors17,066,595,432
model-59-of-61.safetensors17,066,595,432
model-6-of-61.safetensors17,066,593,104
model-60-of-61.safetensors17,066,595,432
model-61-of-61.safetensors19,415,420,696
model-7-of-61.safetensors17,066,593,104
model-8-of-61.safetensors17,066,593,104
model-9-of-61.safetensors17,066,593,104
model.safetensors.index.json12,529,626
tiktoken.model2,795,286
figures/banner.png291,736
figures/Base-Evaluation.png245,449
figures/kimi-logo.png87,988
kimi-logo.png87,988
modeling_deepseek.py75,769
README.md25,517
tokenization_kimi.py12,586
configuration_deepseek.py10,652
docs/tool_call_guidance.md10,280
docs/deploy_guidance.md8,903
tokenizer_config.json3,695
chat_template.jinja2,021
config.json1,725
.gitattributes1,695
THIRD_PARTY_NOTICES.md1,664
LICENSE1,463
.eval_results/terminal_bench.yaml277
.eval_results/apex-swe.yaml268
.eval_results/swe_bench_pro.yaml221
generation_config.json52

All 83 files the tree API returned are listed: every one of the 61 safetensor shards plus 22 other files. Files tab: https://huggingface.co/moonshotai/Kimi-K2-Instruct/tree/main.

05

Continuum serving

OpenRouter slug moonshotai/kimi-k2 on the 2026-08-19 catalog.

First party: https://platform.moonshot.ai/docs/overview.

Not on the Continuum host list fetched 2026-08-19.

Continuum hosted id not on the host list

We list first-party and OpenRouter, plus Continuum only when the live public allowlist named the id. This is not a 15-host routing table.

06

Economics

OpenRouter 2026-08-19: $0.57 in / $2.30 out per 1M tokens. Context 131,072 in / 100,352 out.

No separate internal-reasoning price on that row.

Price this model Opens the pricing calculator preloaded with Kimi K2 0711.
07

Provenance

ClaimSourceAs of
OpenRouter id moonshotai/kimi-k2; context 131072; created 1752263252; $0.57 / $2.30 per 1MOpenRouter /api/v1/models2026-08-19
AA Intelligence Index 20; 5 printed Index benchesArtificial Analysis model page2026-08-19
OpenRouter id moonshotai/kimi-k2, context 131,072OpenRouter /api/v1/models2026-08-19
HF downloads 182,842Hugging Face API moonshotai/Kimi-K2-Instruct2026-08-19
HF safetensors.total: 1,026,408,235,864 stored tensors; HF API usedStorage 1,029,226,387,740 bytes; created 2025-07-11T00:55:12.000Z; HF tensors F32 + BF16 + F8_E4M3Hugging Face API moonshotai/Kimi-K2-Instruct2026-08-19
61 safetensor shards; shard bytes 1,029,190,981,272; tree files 83Hugging Face tree API moonshotai/Kimi-K2-Instruct2026-08-19
08

Compare, FAQ, and Get Plus

No other card in this cluster shares a published DeepSWE official row we can put next to this one. We will not compare on vendor-blog numbers.

Continue through Moonshot's model family with Kimi K3.

FAQ

Why does the lab blog disagree with DeepSWE or Scale?

Lab posts pick a harness, an effort, a split, and sometimes a private eval set. DeepSWE publishes the official mini-swe-agent row with cost, tokens, and steps. Scale SWE-bench Pro only counts when the public shared-harness board has a row we can fetch. A higher lab number is usually a different test, not a better one.

OpenAI-compatible call

No Continuum hosted id is verified for this model, so this page does not invent a curl target. Use the first-party API or the OpenRouter slug moonshotai/kimi-k2.

Plus is $25/mo with $25 weekly hosted usage. The Mac app stays free with your own keys. Get Plus is not Download for Mac.