Qwen: Qwen3 VL 235B A22B Instruct wins on more metrics (4 of 6), but the right pick depends on what you optimize for — see the breakdown below.
| Metric | Google: Gemini 2.5 Flash | Qwen: Qwen3 VL 235B A22B Instruct |
|---|---|---|
| Intelligence Index | 14.1 | 14.3✓ |
| Coding Index | — | — |
| GPQA Diamond | 68% | 71%✓ |
| Design Arena Elo | 1099✓ | — |
| Speed (tokens/sec) | — | — |
| Latency | — | — |
| Input price /M | $0.300 | $0.210✓ |
| Output price /M | $2.50 | $1.90✓ |
| Context window | 1.0M✓ | 262K |
| Capabilities | ReasoningToolsJSONVisionAudio | ToolsJSONVision |
Qwen: Qwen3 VL 235B A22B Instruct wins on more metrics (4 of 6), but the right pick depends on what you optimize for — see the breakdown below.
The table on this page compares Google: Gemini 2.5 Flash and Qwen: Qwen3 VL 235B A22B Instruct metric by metric — intelligence benchmarks, coding score, speed, latency, context window and per-token price — with the winner marked on each row. Data refreshes hourly.