modelgrep

OpenAI: gpt-oss-20b vs Qwen: Qwen3 VL 235B A22B Instruct

OpenAI: gpt-oss-20b wins on more metrics (5 of 7), but the right pick depends on what you optimize for — see the breakdown below.

MetricOpenAI: gpt-oss-20bQwen: Qwen3 VL 235B A22B Instruct
Intelligence Index14.914.3
Coding Index20.7
GPQA Diamond69%71%
Design Arena Elo963
Speed (tokens/sec)
Latency
Input price /M$0.030$0.210
Output price /M$0.130$1.90
Context window131K262K
CapabilitiesReasoningToolsJSONToolsJSONVision

Frequently asked

Is OpenAI: gpt-oss-20b better than Qwen: Qwen3 VL 235B A22B Instruct?

OpenAI: gpt-oss-20b wins on more metrics (5 of 7), but the right pick depends on what you optimize for — see the breakdown below.

What are the biggest differences between OpenAI: gpt-oss-20b and Qwen: Qwen3 VL 235B A22B Instruct?

The table on this page compares OpenAI: gpt-oss-20b and Qwen: Qwen3 VL 235B A22B Instruct metric by metric — intelligence benchmarks, coding score, speed, latency, context window and per-token price — with the winner marked on each row. Data refreshes hourly.