modelgrep
qwen logo

Qwen: Qwen3 VL 30B A3B Instruct

qwen/qwen3-vl-30b-a3b-instruct

152nd smartest of 182ToolsJSONVision
Use via OpenRouter ↗
Intelligence
10.0
152nd of 182
Design Elo
Speed
tokens/sec
Latency
first token
Input price
$0.150
85th cheapest
Context
262K
16K max out

How it compares

Smarter than16%
of all ranked models
Cheaper than74%
of all ranked models

Overview

Qwen3-VL-30B-A3B-Instruct is a multimodal model that unifies strong text generation with visual understanding for images and videos. Its Instruct variant optimizes instruction-following for general multimodal tasks. It excels in perception...

Benchmarks

independent · Artificial Analysis & Design Arena
Artificial Analysis36th percentile
Intelligence Index
10.0
GPQA Diamond
70%
Humanity's Last Exam
6%
SciCode
31%
Tau²-Bench (agentic)
19%

Providers & pricing (4)

ProviderIn $/MOut $/MUptime
Alibabafp8$0.130$0.520100%
DeepInfrafp8$0.150$0.60099.9%
Novitabf16$0.200$0.70094.7%
SiliconFlowfp8$0.290$1.0099.9%

Specifications

Context window262K
Max output16K
Knowledge cutoffMar 2025
Input modalitiestext, image
Output modalitiestext
Prompt caching
Cache read price
ModeratedNo

Qwen3 VL 30B A3B Instruct FAQ

How much does Qwen3 VL 30B A3B Instruct cost?

Qwen3 VL 30B A3B Instruct costs $0.150 per million input tokens and $0.600 per million output tokens via OpenRouter, making it 85th cheapest of 332 paid models.

How smart is Qwen3 VL 30B A3B Instruct?

Qwen3 VL 30B A3B Instruct scores 10.0 on the Artificial Analysis Intelligence Index, ranking 152nd of 182 benchmarked models, with a GPQA Diamond score of 70%.

What is Qwen3 VL 30B A3B Instruct's context window?

Qwen3 VL 30B A3B Instruct supports a 262K-token context window and can output up to 16K tokens. It accepts text, image input.

Compare head-to-head