Models/Qwen/Qwen3 VL 8B Thinking

Qwen3 VL 8B Thinking

Qwen3-VL-8B-Thinking is the reasoning-optimized variant of the Qwen3-VL-8B multimodal model, designed for advanced visual and textual reasoning across complex scenes, documents, and temporal sequences. It integrates enhanced multimodal alignment and...

01

Identity

131,072Context in
32,768Context out
2025-10-14 (OpenRouter created 1760463746)Released
OpenWeights
Canonical name
Qwen3 VL 8B Thinking
Aliases
qwen/qwen3-vl-8b-thinking, Qwen/Qwen3-VL-8B-Thinking
Continuum hosted id
Not listed as a Continuum hosted id
OpenRouter slug
qwen/qwen3-vl-8b-thinking
Hugging Face
Qwen/Qwen3-VL-8B-Thinking
Modalities
image, text
Open weights Not on Continuum host list Open

Context window on the 2026-08-19 OpenRouter row: 131,072 tokens.

Max completion tokens on that row: 32,768.

Input / output on that row: $0.18 / $2.10 per 1M tokens.

Architecture fields: text+image->text · Qwen3.

Hugging Face id on that row: Qwen/Qwen3-VL-8B-Thinking.

02

Should I use this for coding agents

Qwen3-VL-8B-Thinking is the reasoning-optimized variant of the Qwen3-VL-8B multimodal model, designed for advanced visual and textual reasoning across complex scenes, documents, and temporal sequences.

When you want the 2025-10-14 catalog SKU, not a later rename.

No independent row for this SKU

No independent board lists this exact SKU as of 2026-08-19.

Closest benchmarked sibling: Qwen3.8 Max, 57% ±3% Pass@1 at xhigh on the DeepSWE official board. That is a score for the sibling, not for this SKU.

Boards checked: DeepSWE official (2026-08-13), Artificial Analysis Intelligence Index, vals.ai hero index. Every card that does carry a row is on the coding leaderboard.

DeepSWE

No official DeepSWE mini-swe-agent row for this identity on the public board we fetched.

omitted, not invented as of 2026-08-13 source
Artificial Analysis

No Intelligence Index integer fetched for this identity. Chip omitted.

omitted, not invented as of 2026-08-19 source
Vals Index

We did not open a vals.ai card for this identity. Chip omitted.

omitted, not invented as of 2026-08-19 source
03

When not to use it

Trust this before you buy

  • When you need a Continuum-hosted SKU from this lab, use the hosted card on this hub instead of this catalog row.
  • $0.18 in / $2.10 out per 1M on the 2026-08-19 OpenRouter row. Context 131,072 in / 32,768 out.
  • Need a denser published sibling on this hub? Qwen3.8 Max is the card already on the cluster.
  • When the job needs a published DeepSWE, Artificial Analysis, or vals.ai chip: this page does not invent one. Those boards did not list this exact SKU on 2026-08-19.

Plus is $25/mo with $25 weekly hosted usage. The Mac app stays free with your own keys. Get Plus is not Download for Mac.

04

Artifact and Hugging Face downloads

Artifact

SPDX / license
apache-2.0
Total params
HF safetensors.total: 8,767,123,696 stored tensors
Native precision
HF tensors BF16
Repo size
HF API usedStorage 17,534,339,512 bytes
HF created
2025-10-11T07:24:34.000Z
Official HF repo
Qwen/Qwen3-VL-8B-Thinking
HF downloads
91,981
HF likes
220

Official weights id on the 2026-08-19 OpenRouter row: Qwen/Qwen3-VL-8B-Thinking. Community GGUF is a quant, not a second Continuum card.

Official weights

This is the model repo, not Download for Mac. Downloads and likes from the Hugging Face API on 2026-08-19.

Hub file tree

Safetensor shards
4
Shard bytes
17,534,339,512 bytes
Other file bytes
11,576,266 bytes
Tree file count
16
Hub usedStorage
17,534,339,512 bytes

usedStorage and the sum of .safetensors sizes can differ (LFS pointers, non-shard files, duplicate copies). Cited separately. Source: recursive tree API, 2026-08-19.

PathBytes
model-00001-of-00004.safetensors4,902,275,944
model-00002-of-00004.safetensors4,915,962,496
model-00003-of-00004.safetensors4,999,831,048
model-00004-of-00004.safetensors2,716,270,024
tokenizer.json7,032,403
vocab.json2,776,833
merges.txt1,671,839
model.safetensors.index.json67,759
tokenizer_config.json10,781
README.md7,201
chat_template.json5,412
.gitattributes1,519
config.json1,474
preprocessor_config.json390
video_preprocessor_config.json385
generation_config.json270

All 16 files the tree API returned are listed: every one of the 4 safetensor shards plus 12 other files. Files tab: https://huggingface.co/Qwen/Qwen3-VL-8B-Thinking/tree/main.

05

Continuum serving

OpenRouter slug qwen/qwen3-vl-8b-thinking on the 2026-08-19 catalog.

First party: https://qwen.ai/.

Not on the Continuum host list fetched 2026-08-19.

Continuum hosted id not on the host list

We list first-party and OpenRouter, plus Continuum only when the live public allowlist named the id. This is not a 15-host routing table.

06

Economics

OpenRouter 2026-08-19: $0.18 in / $2.10 out per 1M tokens. Context 131,072 in / 32,768 out.

No separate internal-reasoning price on that row.

Price this model Opens the pricing calculator preloaded with Qwen3 VL 8B Thinking.
07

Provenance

ClaimSourceAs of
OpenRouter id qwen/qwen3-vl-8b-thinking; context 131072; created 1760463746; $0.18 / $2.10 per 1MOpenRouter /api/v1/models2026-08-19
OpenRouter id qwen/qwen3-vl-8b-thinking, context 131,072OpenRouter /api/v1/models2026-08-19
HF downloads 91,981Hugging Face API Qwen/Qwen3-VL-8B-Thinking2026-08-19
HF safetensors.total: 8,767,123,696 stored tensors; HF API usedStorage 17,534,339,512 bytes; created 2025-10-11T07:24:34.000Z; HF tensors BF16Hugging Face API Qwen/Qwen3-VL-8B-Thinking2026-08-19
4 safetensor shards; shard bytes 17,534,339,512; tree files 16Hugging Face tree API Qwen/Qwen3-VL-8B-Thinking2026-08-19
08

Compare, FAQ, and Get Plus

No other card in this cluster shares a published DeepSWE official row we can put next to this one. We will not compare on vendor-blog numbers.

Continue through Qwen's model family with Qwen3 VL 8B Instruct.

FAQ

Why does the lab blog disagree with DeepSWE or Scale?

Lab posts pick a harness, an effort, a split, and sometimes a private eval set. DeepSWE publishes the official mini-swe-agent row with cost, tokens, and steps. Scale SWE-bench Pro only counts when the public shared-harness board has a row we can fetch. A higher lab number is usually a different test, not a better one.

OpenAI-compatible call

No Continuum hosted id is verified for this model, so this page does not invent a curl target. Use the first-party API or the OpenRouter slug qwen/qwen3-vl-8b-thinking.

Plus is $25/mo with $25 weekly hosted usage. The Mac app stays free with your own keys. Get Plus is not Download for Mac.