modelgrep

OpenAI: gpt-oss-120b vs Qwen: Qwen3 Max

OpenAI: gpt-oss-120b wins on more metrics (4 of 7), but the right pick depends on what you optimize for — see the breakdown below.

MetricOpenAI: gpt-oss-120bQwen: Qwen3 Max
Intelligence Index23.824.0
Coding Index30.4
GPQA Diamond78%76%
Design Arena Elo10291143
Speed (tokens/sec)
Latency
Input price /M$0.037$0.780
Output price /M$0.170$3.90
Context window131K262K
CapabilitiesReasoningToolsJSONToolsJSON

Frequently asked

Is OpenAI: gpt-oss-120b better than Qwen: Qwen3 Max?

OpenAI: gpt-oss-120b wins on more metrics (4 of 7), but the right pick depends on what you optimize for — see the breakdown below.

What are the biggest differences between OpenAI: gpt-oss-120b and Qwen: Qwen3 Max?

The table on this page compares OpenAI: gpt-oss-120b and Qwen: Qwen3 Max metric by metric — intelligence benchmarks, coding score, speed, latency, context window and per-token price — with the winner marked on each row. Data refreshes hourly.