OpenAI: gpt-oss-20b wins on more metrics (5 of 7), but the right pick depends on what you optimize for — see the breakdown below.
| Metric | OpenAI: gpt-oss-20b | Qwen: Qwen3 VL 235B A22B Instruct |
|---|---|---|
| Intelligence Index | 14.9✓ | 14.3 |
| Coding Index | 20.7✓ | — |
| GPQA Diamond | 69% | 71%✓ |
| Design Arena Elo | 963✓ | — |
| Speed (tokens/sec) | — | — |
| Latency | — | — |
| Input price /M | $0.030✓ | $0.210 |
| Output price /M | $0.130✓ | $1.90 |
| Context window | 131K | 262K✓ |
| Capabilities | ReasoningToolsJSON | ToolsJSONVision |
OpenAI: gpt-oss-20b wins on more metrics (5 of 7), but the right pick depends on what you optimize for — see the breakdown below.
The table on this page compares OpenAI: gpt-oss-20b and Qwen: Qwen3 VL 235B A22B Instruct metric by metric — intelligence benchmarks, coding score, speed, latency, context window and per-token price — with the winner marked on each row. Data refreshes hourly.