modelgrep
qwen logo

Qwen: Qwen3 VL 8B Instruct

qwen/qwen3-vl-8b-instruct

161st smartest of 182Cheaper than 82% of paidToolsJSONVision
Use via OpenRouter ↗
Intelligence
8.4
161st of 182
Design Elo
Speed
tokens/sec
Latency
first token
Input price
$0.117
61st cheapest
Context
262K
33K max out

How it compares

Smarter than12%
of all ranked models
Cheaper than82%
of all ranked models

Overview

Qwen3-VL-8B-Instruct is a multimodal vision-language model from the Qwen3-VL series, built for high-fidelity understanding and reasoning across text, images, and video. It features improved multimodal fusion with Interleaved-MRoPE for long-horizon...

Benchmarks

independent · Artificial Analysis & Design Arena
Artificial Analysis28th percentile
Intelligence Index
8.4
GPQA Diamond
43%
Humanity's Last Exam
3%
SciCode
17%
Tau²-Bench (agentic)
29%

Providers & pricing (2)

ProviderIn $/MOut $/MUptime
Alibabafp8$0.117$0.455100%
Parasailbf16$0.250$0.75095.5%

Specifications

Context window262K
Max output33K
Knowledge cutoff
Input modalitiesimage, text
Output modalitiestext
Prompt caching
Cache read price
ModeratedNo

Qwen3 VL 8B Instruct FAQ

How much does Qwen3 VL 8B Instruct cost?

Qwen3 VL 8B Instruct costs $0.117 per million input tokens and $0.455 per million output tokens via OpenRouter, making it 61st cheapest of 332 paid models.

How smart is Qwen3 VL 8B Instruct?

Qwen3 VL 8B Instruct scores 8.4 on the Artificial Analysis Intelligence Index, ranking 161st of 182 benchmarked models, with a GPQA Diamond score of 43%.

What is Qwen3 VL 8B Instruct's context window?

Qwen3 VL 8B Instruct supports a 262K-token context window and can output up to 33K tokens. It accepts image, text input.

Compare head-to-head