Models/Google/Gemma 4 31B (free)

Gemma 4 31B (free)

Gemma 4 31B Instruct is Google DeepMind's 30.7B dense multimodal model supporting text and image input with text output. Features a 256K token context window, configurable thinking/reasoning mode, native function...

01

Identity

262,144Context in
32,768Context out
2026-04-02 (OpenRouter created 1775148486)Released
OpenWeights
Canonical name
Gemma 4 31B (free)
Aliases
google/gemma-4-31b-it:free, google/gemma-4-31B-it
Continuum hosted id
google/gemma-4-31b-it:free
OpenRouter slug
google/gemma-4-31b-it:free
Hugging Face
google/gemma-4-31B-it
Modalities
image, text, video
Open weights Hosted on Continuum Hosted

Context window on the 2026-08-19 OpenRouter row: 262,144 tokens.

Max completion tokens on that row: 32,768.

Input / output on that row: $0 / $0 per 1M tokens.

Architecture fields: text+image+video->text · Gemma.

Hugging Face id on that row: google/gemma-4-31B-it.

02

Should I use this for coding agents

Gemma 4 31B Instruct is Google DeepMind's 30.7B dense multimodal model supporting text and image input with text output.

When you want the 2026-04-02 catalog SKU, not a later rename.

No independent row for this SKU

No independent board lists this exact SKU as of 2026-08-19.

Closest benchmarked sibling: Gemini 3.7 Flash, 65% ±2% Pass@1 at high on the DeepSWE official board. That is a score for the sibling, not for this SKU.

Boards checked: DeepSWE official (2026-08-13), Artificial Analysis Intelligence Index, vals.ai hero index. Every card that does carry a row is on the coding leaderboard.

DeepSWE

No official DeepSWE mini-swe-agent row for this identity on the public board we fetched.

omitted, not invented as of 2026-08-13 source
Artificial Analysis

No Intelligence Index integer fetched for this identity. Chip omitted.

omitted, not invented as of 2026-08-19 source
Vals Index

We did not open a vals.ai card for this identity. Chip omitted.

omitted, not invented as of 2026-08-19 source
03

When not to use it

Trust this before you buy

  • This SKU is on the Continuum host list as google/gemma-4-31b-it:free. Use a different card on this hub when you need another lab SKU.
  • $0 in / $0 out per 1M on the 2026-08-19 OpenRouter row. Context 262,144 in / 32,768 out.
  • Need a denser published sibling on this hub? Gemini 3.7 Flash is the card already on the cluster.
  • When the job needs a published DeepSWE, Artificial Analysis, or vals.ai chip: this page does not invent one. Those boards did not list this exact SKU on 2026-08-19.

Plus is $25/mo with $25 weekly hosted usage. The Mac app stays free with your own keys. Get Plus is not Download for Mac.

04

Artifact and Hugging Face downloads

Artifact

SPDX / license
apache-2.0
Total params
HF safetensors.total: 31,273,088,876 stored tensors
Native precision
HF tensors BF16
Repo size
HF API usedStorage 174,974,072,605 bytes
HF created
2026-03-11T18:22:36.000Z
Official HF repo
google/gemma-4-31B-it
HF downloads
9,311,525
HF likes
3,600

Official weights id on the 2026-08-19 OpenRouter row: google/gemma-4-31B-it. Community GGUF is a quant, not a second Continuum card.

Official weights

This is the model repo, not Download for Mac. Downloads and likes from the Hugging Face API on 2026-08-19.

Hub file tree

Safetensor shards
2
Shard bytes
62,546,338,248 bytes
Other file bytes
32,348,008 bytes
Tree file count
12
Hub usedStorage
174,974,072,605 bytes

usedStorage and the sum of .safetensors sizes can differ (LFS pointers, non-shard files, duplicate copies). Cited separately. Source: recursive tree API, 2026-08-19.

PathBytes
model-00001-of-00002.safetensors49,784,788,364
model-00002-of-00002.safetensors12,761,549,884
tokenizer.json32,169,626
model.safetensors.index.json120,246
README.md27,963
chat_template.jinja18,683
config.json4,621
tokenizer_config.json3,082
.gitattributes1,708
processor_config.json1,689
generation_config.json208
.eval_results/mmmu_pro.yaml182

All 12 files the tree API returned are listed: every one of the 2 safetensor shards plus 10 other files. Files tab: https://huggingface.co/google/gemma-4-31B-it/tree/main.

05

Continuum serving

OpenRouter slug google/gemma-4-31b-it:free on the 2026-08-19 catalog.

First party: https://ai.google.dev/gemini-api/docs/models.

Continuum hosts google/gemma-4-31b-it:free on the live public allowlist fetched 2026-08-19.

Continuum hosted id google/gemma-4-31b-it:free

click to select

We list first-party and OpenRouter, plus Continuum only when the live public allowlist named the id. This is not a 15-host routing table.

06

Economics

OpenRouter 2026-08-19: $0 in / $0 out per 1M tokens. Context 262,144 in / 32,768 out.

No separate internal-reasoning price on that row.

Price this model Opens the pricing calculator preloaded with Gemma 4 31B (free).
07

Provenance

ClaimSourceAs of
OpenRouter id google/gemma-4-31b-it:free; context 262144; created 1775148486; $0 / $0 per 1MOpenRouter /api/v1/models2026-08-19
OpenRouter id google/gemma-4-31b-it:free, context 262,144OpenRouter /api/v1/models2026-08-19
HF downloads 9,311,525Hugging Face API google/gemma-4-31B-it2026-08-19
HF safetensors.total: 31,273,088,876 stored tensors; HF API usedStorage 174,974,072,605 bytes; created 2026-03-11T18:22:36.000Z; HF tensors BF16Hugging Face API google/gemma-4-31B-it2026-08-19
2 safetensor shards; shard bytes 62,546,338,248; tree files 12Hugging Face tree API google/gemma-4-31B-it2026-08-19
08

Compare, FAQ, and Get Plus

No other card in this cluster shares a published DeepSWE official row we can put next to this one. We will not compare on vendor-blog numbers.

Continue through Google's model family with Gemma 3 4B.

FAQ

Why does the lab blog disagree with DeepSWE or Scale?

Lab posts pick a harness, an effort, a split, and sometimes a private eval set. DeepSWE publishes the official mini-swe-agent row with cost, tokens, and steps. Scale SWE-bench Pro only counts when the public shared-harness board has a row we can fetch. A higher lab number is usually a different test, not a better one.

OpenAI-compatible call

Verified against the live public allowlist on 2026-08-19. Base URL is https://continuumcode.ai/v1. Keys are cont_sk_ from Settings, Account, Inference API. Personal keys need Plus or above.

curl https://continuumcode.ai/v1/chat/completions \
  -H "Authorization: Bearer $CONTINUUM_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{"model":"google/gemma-4-31b-it:free","messages":[{"role":"user","content":"Review this diff."}]}'

Plus is $25/mo with $25 weekly hosted usage. The Mac app stays free with your own keys. Get Plus is not Download for Mac.