qwen/qwen3-vl-30b-a3b-instruct
Qwen3-VL-30B-A3B-Instruct is a multimodal model that unifies strong text generation with visual understanding for images and videos. Its Instruct variant optimizes instruction-following for general multimodal tasks. It excels in perception...
| Provider | In $/M | Out $/M | Context | Uptime |
|---|---|---|---|---|
| Alibabafp8 | $0.130 | $0.520 | 131K | 100% |
| DeepInfrafp8 | $0.150 | $0.600 | 262K | 99.9% |
| Novitabf16 | $0.200 | $0.700 | 131K | 94.7% |
| SiliconFlowfp8 | $0.290 | $1.00 | 262K | 99.9% |
Qwen3 VL 30B A3B Instruct costs $0.150 per million input tokens and $0.600 per million output tokens via OpenRouter, making it 85th cheapest of 332 paid models.
Qwen3 VL 30B A3B Instruct scores 10.0 on the Artificial Analysis Intelligence Index, ranking 152nd of 182 benchmarked models, with a GPQA Diamond score of 70%.
Qwen3 VL 30B A3B Instruct supports a 262K-token context window and can output up to 16K tokens. It accepts text, image input.