Models/NVIDIA/Nemotron Nano 12B 2 VL (free)

Nemotron Nano 12B 2 VL (free)

NVIDIA Nemotron Nano 2 VL is a 12-billion-parameter open multimodal reasoning model designed for video understanding and document intelligence. It introduces a hybrid Transformer-Mamba architecture, combining transformer-level accuracy with Mamba’s...

01

Identity

128,000Context in
128,000Context out
2025-10-28 (OpenRouter created 1761675565)Released
OpenWeights
Canonical name
Nemotron Nano 12B 2 VL (free)
Aliases
nvidia/nemotron-nano-12b-v2-vl:free, nvidia/NVIDIA-Nemotron-Nano-12B-v2-VL-BF16
Continuum hosted id
Not listed as a Continuum hosted id
OpenRouter slug
nvidia/nemotron-nano-12b-v2-vl:free
Hugging Face
nvidia/NVIDIA-Nemotron-Nano-12B-v2-VL-BF16
Modalities
image, text, video
Open weights Not on Continuum host list Open

Context window on the 2026-08-19 OpenRouter row: 128,000 tokens.

Max completion tokens on that row: 128,000.

Input / output on that row: $0 / $0 per 1M tokens.

Architecture fields: text+image+video->text · Other.

Hugging Face id on that row: nvidia/NVIDIA-Nemotron-Nano-12B-v2-VL-BF16.

02

Should I use this for coding agents

NVIDIA Nemotron Nano 2 VL is a 12-billion-parameter open multimodal reasoning model designed for video understanding and document intelligence.

When you want the 2025-10-28 catalog SKU, not a later rename.

DeepSWE

No official DeepSWE mini-swe-agent row for this identity on the public board we fetched.

omitted, not invented as of 2026-08-13 source
Artificial Analysispublished integer
4
bench
Artificial Analysis
version
Intelligence Index
harness
published integer
metric
Intelligence Index
value
4
independent as of 2026-08-19 source
Vals Index

We did not open a vals.ai card for this identity. Chip omitted.

omitted, not invented as of 2026-08-19 source

Ranked view: Best coding models: the independent leaderboard puts this row and every other card that carries an independent coding score on one board, with the confidence intervals left visible.

Named benches, printed only

Copied from AA or vals HTML we opened. 0.0% placeholder bars are omitted. Do not average these into the hero tiles, and do not invent CI, $/task, or steps for Artificial Analysis.

BenchPrintedNoteSourceURLAs of
SciCode 17.6% Printed on the AA Intelligence Evaluations grid for NVIDIA Nemotron Nano 12B v2 VL (Non-reasoning). Not a DeepSWE chip. Artificial Analysis NVIDIA Nemotron Nano 12B v2 VL source 2026-08-19
Humanity's Last Exam 4.3% Printed on the AA Intelligence Evaluations grid for NVIDIA Nemotron Nano 12B v2 VL (Non-reasoning). Not a DeepSWE chip. Artificial Analysis NVIDIA Nemotron Nano 12B v2 VL source 2026-08-19
GPQA Diamond 43.9% Printed on the AA Intelligence Evaluations grid for NVIDIA Nemotron Nano 12B v2 VL (Non-reasoning). Not a DeepSWE chip. Artificial Analysis NVIDIA Nemotron Nano 12B v2 VL source 2026-08-19
AA-Omniscience Accuracy 11.9% Printed on the AA Intelligence Evaluations grid for NVIDIA Nemotron Nano 12B v2 VL (Non-reasoning). Not a DeepSWE chip. Artificial Analysis NVIDIA Nemotron Nano 12B v2 VL source 2026-08-19
AA-LCR 19.7% Printed on the AA Intelligence Evaluations grid for NVIDIA Nemotron Nano 12B v2 VL (Non-reasoning). Not a DeepSWE chip. Artificial Analysis NVIDIA Nemotron Nano 12B v2 VL source 2026-08-19
03

When not to use it

Trust this before you buy

  • When you need a Continuum-hosted SKU from this lab, use the hosted card on this hub instead of this catalog row.
  • $0 in / $0 out per 1M on the 2026-08-19 OpenRouter row. Context 128,000 in / 128,000 out.
  • Need a denser published sibling on this hub? Nemotron 3 Ultra is the card already on the cluster.

Plus is $25/mo with $25 weekly hosted usage. The Mac app stays free with your own keys. Get Plus is not Download for Mac.

04

Artifact and Hugging Face downloads

Artifact

SPDX / license
other
Total params
HF safetensors.total: 13,181,860,358 stored tensors
Native precision
HF tensors F32 + BF16
Repo size
HF API usedStorage 26,400,962,337 bytes
HF created
2025-10-21T18:11:05.000Z
Official HF repo
nvidia/NVIDIA-Nemotron-Nano-12B-v2-VL-BF16
HF downloads
120,675
HF likes
88

Official weights id on the 2026-08-19 OpenRouter row: nvidia/NVIDIA-Nemotron-Nano-12B-v2-VL-BF16. Community GGUF is a quant, not a second Continuum card.

Official weights

This is the model repo, not Download for Mac. Downloads and likes from the Hugging Face API on 2026-08-19.

Hub file tree

Safetensor shards
7
Shard bytes
26,363,817,792 bytes
Other file bytes
32,051,940 bytes
Tree file count
54
Hub usedStorage
26,400,962,337 bytes

usedStorage and the sum of .safetensors sizes can differ (LFS pointers, non-shard files, duplicate copies). Cited separately. Source: recursive tree API, 2026-08-19.

PathBytes
model-00001-of-00007.safetensors3,831,944,248
model-00002-of-00007.safetensors3,908,097,776
model-00003-of-00007.safetensors3,908,097,808
model-00004-of-00007.safetensors3,824,190,384
model-00005-of-00007.safetensors3,908,097,808
model-00006-of-00007.safetensors3,908,097,808
model-00007-of-00007.safetensors3,075,291,960
tokenizer.json17,079,976
images/demo.mp47,333,182
images/demo_frames/frame_0000.jpg532,176
images/demo_frames/frame_0001.jpg525,666
images/demo_frames/frame_0002.jpg522,853
images/demo_frames/frame_0003.jpg510,824
images/demo_frames/frame_0004.jpg494,798
images/demo_frames/frame_0006.jpg466,415
images/demo_frames/frame_0007.jpg451,954
images/demo_frames/frame_0008.jpg446,495
images/demo_frames/frame_0009.jpg442,849
images/demo_frames/frame_0005.jpg435,068
images/demo_frames/frame_0014.jpg425,102
images/demo_frames/frame_0011.jpg399,758
images/demo_frames/frame_0010.jpg399,499
images/demo_frames/frame_0012.jpg374,855
images/demo_frames/frame_0013.jpg361,900
images/tech.png222,054
tokenizer_config.json185,845
images/table.png131,014
modeling_nemotron_h.py78,570
model.safetensors.index.json72,736
README.md20,007
images/example1a.jpeg14,890

Showing the first 7 of 7 safetensor shards, plus 24 other files. The snapshot was capped at 7 shards when the tree API was read on 2026-08-19, so the remaining 0 are on the Files tab, not on this page. Totals above are the full tree API sum, not this slice. Byte sizes are Hub tree API size fields, never invented.

05

Continuum serving

OpenRouter slug nvidia/nemotron-nano-12b-v2-vl:free on the 2026-08-19 catalog.

First party: https://build.nvidia.com/.

Not on the Continuum host list fetched 2026-08-19.

Continuum hosted id not on the host list

We list first-party and OpenRouter, plus Continuum only when the live public allowlist named the id. This is not a 15-host routing table.

06

Economics

OpenRouter 2026-08-19: $0 in / $0 out per 1M tokens. Context 128,000 in / 128,000 out.

No separate internal-reasoning price on that row.

Price this model Opens the pricing calculator preloaded with Nemotron Nano 12B 2 VL (free).
07

Provenance

ClaimSourceAs of
OpenRouter id nvidia/nemotron-nano-12b-v2-vl:free; context 128000; created 1761675565; $0 / $0 per 1MOpenRouter /api/v1/models2026-08-19
AA Intelligence Index 4; 5 printed Index benchesArtificial Analysis model page2026-08-19
OpenRouter id nvidia/nemotron-nano-12b-v2-vl:free, context 128,000OpenRouter /api/v1/models2026-08-19
HF downloads 120,675Hugging Face API nvidia/NVIDIA-Nemotron-Nano-12B-v2-VL-BF162026-08-19
HF safetensors.total: 13,181,860,358 stored tensors; HF API usedStorage 26,400,962,337 bytes; created 2025-10-21T18:11:05.000Z; HF tensors F32 + BF16Hugging Face API nvidia/NVIDIA-Nemotron-Nano-12B-v2-VL-BF162026-08-19
7 safetensor shards; shard bytes 26,363,817,792; tree files 54Hugging Face tree API nvidia/NVIDIA-Nemotron-Nano-12B-v2-VL-BF162026-08-19
08

Compare, FAQ, and Get Plus

No other card in this cluster shares a published DeepSWE official row we can put next to this one. We will not compare on vendor-blog numbers.

Continue through NVIDIA's model family with Nemotron Nano 9B V2 (free).

FAQ

Why does the lab blog disagree with DeepSWE or Scale?

Lab posts pick a harness, an effort, a split, and sometimes a private eval set. DeepSWE publishes the official mini-swe-agent row with cost, tokens, and steps. Scale SWE-bench Pro only counts when the public shared-harness board has a row we can fetch. A higher lab number is usually a different test, not a better one.

OpenAI-compatible call

No Continuum hosted id is verified for this model, so this page does not invent a curl target. Use the first-party API or the OpenRouter slug nvidia/nemotron-nano-12b-v2-vl:free.

Plus is $25/mo with $25 weekly hosted usage. The Mac app stays free with your own keys. Get Plus is not Download for Mac.