qwen/qwen3-vl-8b-instruct
Qwen3-VL-8B-Instruct is a multimodal vision-language model from the Qwen3-VL series, built for high-fidelity understanding and reasoning across text, images, and video. It features improved multimodal fusion with Interleaved-MRoPE for long-horizon...
| Provider | In $/M | Out $/M | Context | Uptime |
|---|---|---|---|---|
| Alibabafp8 | $0.117 | $0.455 | 131K | 100% |
| Parasailbf16 | $0.250 | $0.750 | 262K | 95.5% |
Qwen3 VL 8B Instruct costs $0.117 per million input tokens and $0.455 per million output tokens via OpenRouter, making it 61st cheapest of 332 paid models.
Qwen3 VL 8B Instruct scores 8.4 on the Artificial Analysis Intelligence Index, ranking 161st of 182 benchmarked models, with a GPQA Diamond score of 43%.
Qwen3 VL 8B Instruct supports a 262K-token context window and can output up to 33K tokens. It accepts image, text input.