Models/NVIDIA/Nemotron 3 Super (free)

Nemotron 3 Super (free)

NVIDIA Nemotron 3 Super is a 120B-parameter open hybrid MoE model, activating just 12B parameters for maximum compute efficiency and accuracy in complex multi-agent applications. Built on a hybrid Mamba-Transformer...

01

Identity

262,144Context in
262,144Context out
2026-03-11 (OpenRouter created 1773245239)Released
OpenWeights
Canonical name
Nemotron 3 Super (free)
Aliases
nvidia/nemotron-3-super-120b-a12b:free, nvidia/NVIDIA-Nemotron-3-Super-120B-A12B-FP8
Continuum hosted id
nvidia/nemotron-3-super-120b-a12b:free
OpenRouter slug
nvidia/nemotron-3-super-120b-a12b:free
Hugging Face
nvidia/NVIDIA-Nemotron-3-Super-120B-A12B-FP8
Modalities
text
Open weights Hosted on Continuum Hosted

Context window on the 2026-08-19 OpenRouter row: 262,144 tokens.

Max completion tokens on that row: 262,144.

Input / output on that row: $0 / $0 per 1M tokens.

Architecture fields: text->text · Other.

Hugging Face id on that row: nvidia/NVIDIA-Nemotron-3-Super-120B-A12B-FP8.

02

Should I use this for coding agents

NVIDIA Nemotron 3 Super is a 120B-parameter open hybrid MoE model, activating just 12B parameters for maximum compute efficiency and accuracy in complex multi-agent applications.

When you want the 2026-03-11 catalog SKU, not a later rename.

DeepSWE

No official DeepSWE mini-swe-agent row for this identity on the public board we fetched.

omitted, not invented as of 2026-08-13 source
Artificial Analysispublished integer
26
bench
Artificial Analysis
version
Intelligence Index
harness
published integer
metric
Intelligence Index
value
26
independent as of 2026-08-19 source
Vals Index

We did not open a vals.ai card for this identity. Chip omitted.

omitted, not invented as of 2026-08-19 source

Ranked view: Best coding models: the independent leaderboard puts this row and every other card that carries an independent coding score on one board, with the confidence intervals left visible.

Named benches, printed only

Copied from AA or vals HTML we opened. 0.0% placeholder bars are omitted. Do not average these into the hero tiles, and do not invent CI, $/task, or steps for Artificial Analysis.

BenchPrintedNoteSourceURLAs of
GDPval-AA v2 9.9% Printed on the AA Intelligence Evaluations grid for Nemotron 3 Super 120B A12B (Reasoning). Not a DeepSWE chip. Artificial Analysis Nemotron 3 Super source 2026-08-19
τ³-Banking 10.3% Printed on the AA Intelligence Evaluations grid for Nemotron 3 Super 120B A12B (Reasoning). Not a DeepSWE chip. Artificial Analysis Nemotron 3 Super source 2026-08-19
Terminal-Bench v2.1 38.6% Printed on the AA Intelligence Evaluations grid for Nemotron 3 Super 120B A12B (Reasoning). Not a DeepSWE chip. Artificial Analysis Nemotron 3 Super source 2026-08-19
SciCode 36% Printed on the AA Intelligence Evaluations grid for Nemotron 3 Super 120B A12B (Reasoning). Not a DeepSWE chip. Artificial Analysis Nemotron 3 Super source 2026-08-19
Humanity's Last Exam 20.8% Printed on the AA Intelligence Evaluations grid for Nemotron 3 Super 120B A12B (Reasoning). Not a DeepSWE chip. Artificial Analysis Nemotron 3 Super source 2026-08-19
GPQA Diamond 80% Printed on the AA Intelligence Evaluations grid for Nemotron 3 Super 120B A12B (Reasoning). Not a DeepSWE chip. Artificial Analysis Nemotron 3 Super source 2026-08-19
CritPt 3.1% Printed on the AA Intelligence Evaluations grid for Nemotron 3 Super 120B A12B (Reasoning). Not a DeepSWE chip. Artificial Analysis Nemotron 3 Super source 2026-08-19
AA-Omniscience Accuracy 24.3% Printed on the AA Intelligence Evaluations grid for Nemotron 3 Super 120B A12B (Reasoning). Not a DeepSWE chip. Artificial Analysis Nemotron 3 Super source 2026-08-19
AA-LCR 60.3% Printed on the AA Intelligence Evaluations grid for Nemotron 3 Super 120B A12B (Reasoning). Not a DeepSWE chip. Artificial Analysis Nemotron 3 Super source 2026-08-19
03

When not to use it

Trust this before you buy

  • This SKU is on the Continuum host list as nvidia/nemotron-3-super-120b-a12b:free. Use a different card on this hub when you need another lab SKU.
  • $0 in / $0 out per 1M on the 2026-08-19 OpenRouter row. Context 262,144 in / 262,144 out.
  • Need a denser published sibling on this hub? Nemotron 3 Ultra is the card already on the cluster.

Plus is $25/mo with $25 weekly hosted usage. The Mac app stays free with your own keys. Get Plus is not Download for Mac.

04

Artifact and Hugging Face downloads

Artifact

SPDX / license
other
Total params
HF safetensors.total: 123,611,012,096 stored tensors
Native precision
HF tensors F32 + BF16 + F8_E4M3
Repo size
HF API usedStorage 128,379,469,916 bytes
HF created
2026-03-10T18:32:42.000Z
Official HF repo
nvidia/NVIDIA-Nemotron-3-Super-120B-A12B-FP8
HF downloads
175,033
HF likes
274

Official weights id on the 2026-08-19 OpenRouter row: nvidia/NVIDIA-Nemotron-3-Super-120B-A12B-FP8. Community GGUF is a quant, not a second Continuum card.

Official weights

This is the model repo, not Download for Mac. Downloads and likes from the Hugging Face API on 2026-08-19.

Hub file tree

Safetensor shards
26
Shard bytes
128,350,001,680 bytes
Other file bytes
29,946,903 bytes
Tree file count
45
Hub usedStorage
128,379,469,916 bytes

usedStorage and the sum of .safetensors sizes can differ (LFS pointers, non-shard files, duplicate copies). Cited separately. Source: recursive tree API, 2026-08-19.

PathBytes
model-00001-of-00026.safetensors5,000,243,808
model-00002-of-00026.safetensors4,993,713,376
model-00003-of-00026.safetensors5,000,119,040
model-00004-of-00026.safetensors4,999,076,840
model-00005-of-00026.safetensors5,000,270,768
model-00006-of-00026.safetensors4,999,077,288
model-00007-of-00026.safetensors4,998,813,984
model-00008-of-00026.safetensors5,000,533,616
model-00009-of-00026.safetensors4,999,077,224
model-00010-of-00026.safetensors4,998,813,992
model-00011-of-00026.safetensors5,000,533,608
model-00012-of-00026.safetensors4,999,077,168
model-00013-of-00026.safetensors4,998,813,984
model-00014-of-00026.safetensors5,000,533,616
model-00015-of-00026.safetensors4,999,077,104
model-00016-of-00026.safetensors4,998,814,008
model-00017-of-00026.safetensors5,000,533,600
model-00018-of-00026.safetensors4,999,077,048
model-00019-of-00026.safetensors4,998,814,040
model-00020-of-00026.safetensors5,000,533,560
model-00021-of-00026.safetensors4,999,076,984
model-00022-of-00026.safetensors4,998,814,112
model-00023-of-00026.safetensors5,000,533,512
model-00024-of-00026.safetensors4,999,076,944
model-00025-of-00026.safetensors4,998,841,656
model-00026-of-00026.safetensors3,368,110,800
tokenizer.json17,077,484
model.safetensors.index.json12,390,752
tokenizer_config.json177,209
modeling_nemotron_h.py82,338
accuracy_chart.png79,162
README.md79,162
configuration_nemotron_h.py19,822
chat_template.jinja10,771
config.json8,440
hf_quant_config.json6,888
explainability.md3,155
privacy.md2,688
bias.md2,626
safety.md2,120
super_v3_reasoning_parser.py1,878
.gitattributes1,635
special_tokens_map.json563
generation_config.json210
__init__.py0

All 45 files the tree API returned are listed: every one of the 26 safetensor shards plus 19 other files. Files tab: https://huggingface.co/nvidia/NVIDIA-Nemotron-3-Super-120B-A12B-FP8/tree/main.

05

Continuum serving

OpenRouter slug nvidia/nemotron-3-super-120b-a12b:free on the 2026-08-19 catalog.

First party: https://build.nvidia.com/.

Continuum hosts nvidia/nemotron-3-super-120b-a12b:free on the live public allowlist fetched 2026-08-19.

Continuum hosted id nvidia/nemotron-3-super-120b-a12b:free

click to select

We list first-party and OpenRouter, plus Continuum only when the live public allowlist named the id. This is not a 15-host routing table.

06

Economics

OpenRouter 2026-08-19: $0 in / $0 out per 1M tokens. Context 262,144 in / 262,144 out.

No separate internal-reasoning price on that row.

Price this model Opens the pricing calculator preloaded with Nemotron 3 Super (free).
07

Provenance

ClaimSourceAs of
OpenRouter id nvidia/nemotron-3-super-120b-a12b:free; context 262144; created 1773245239; $0 / $0 per 1MOpenRouter /api/v1/models2026-08-19
AA Intelligence Index 26; 9 printed Index benchesArtificial Analysis model page2026-08-19
OpenRouter id nvidia/nemotron-3-super-120b-a12b:free, context 262,144OpenRouter /api/v1/models2026-08-19
HF downloads 175,033Hugging Face API nvidia/NVIDIA-Nemotron-3-Super-120B-A12B-FP82026-08-19
HF safetensors.total: 123,611,012,096 stored tensors; HF API usedStorage 128,379,469,916 bytes; created 2026-03-10T18:32:42.000Z; HF tensors F32 + BF16 + F8_E4M3Hugging Face API nvidia/NVIDIA-Nemotron-3-Super-120B-A12B-FP82026-08-19
26 safetensor shards; shard bytes 128,350,001,680; tree files 45Hugging Face tree API nvidia/NVIDIA-Nemotron-3-Super-120B-A12B-FP82026-08-19
08

Compare, FAQ, and Get Plus

No other card in this cluster shares a published DeepSWE official row we can put next to this one. We will not compare on vendor-blog numbers.

Continue through NVIDIA's model family with Nemotron 3 Nano Omni (free).

FAQ

Why does the lab blog disagree with DeepSWE or Scale?

Lab posts pick a harness, an effort, a split, and sometimes a private eval set. DeepSWE publishes the official mini-swe-agent row with cost, tokens, and steps. Scale SWE-bench Pro only counts when the public shared-harness board has a row we can fetch. A higher lab number is usually a different test, not a better one.

OpenAI-compatible call

Verified against the live public allowlist on 2026-08-19. Base URL is https://continuumcode.ai/v1. Keys are cont_sk_ from Settings, Account, Inference API. Personal keys need Plus or above.

curl https://continuumcode.ai/v1/chat/completions \
  -H "Authorization: Bearer $CONTINUUM_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{"model":"nvidia/nemotron-3-super-120b-a12b:free","messages":[{"role":"user","content":"Review this diff."}]}'

Plus is $25/mo with $25 weekly hosted usage. The Mac app stays free with your own keys. Get Plus is not Download for Mac.