qwen/qwen3-vl-235b-a22b-instruct
Qwen3-VL-235B-A22B Instruct is an open-weight multimodal model that unifies strong text generation with visual understanding across images and video. The Instruct model targets general vision-language use (VQA, document parsing, chart/table...
| Provider | In $/M | Out $/M | Context | Uptime |
|---|---|---|---|---|
| DeepInfrafp8 | $0.200 | $0.880 | 262K | 84.7% |
| Venicefp8 | $0.210 | $1.90 | 128K | 98.6% |
| Parasailfp8 | $0.210 | $1.90 | 131K | 98.3% |
| Alibabafp8 | $0.260 | $1.04 | 131K | 100% |
| Novitabf16 | $0.300 | $1.50 | 131K | 89.1% |
Qwen3 VL 235B A22B Instruct costs $0.210 per million input tokens and $1.90 per million output tokens via OpenRouter, making it 105th cheapest of 332 paid models.
Qwen3 VL 235B A22B Instruct scores 14.3 on the Artificial Analysis Intelligence Index, ranking 139th of 182 benchmarked models, with a GPQA Diamond score of 71%.
Qwen3 VL 235B A22B Instruct supports a 262K-token context window and can output up to 33K tokens. It accepts text, image input.