OpenAI: gpt-oss-20b wins on more metrics (6 of 6), but the right pick depends on what you optimize for — see the breakdown below.
| Metric | Mistral: Mistral Medium 3.1 | OpenAI: gpt-oss-20b |
|---|---|---|
| Intelligence Index | 14.7 | 14.9✓ |
| Coding Index | 20.5 | 20.7✓ |
| GPQA Diamond | 59% | 69%✓ |
| Design Arena Elo | — | 963✓ |
| Speed (tokens/sec) | — | — |
| Latency | — | — |
| Input price /M | $0.400 | $0.030✓ |
| Output price /M | $2.00 | $0.130✓ |
| Context window | 131K | 131K |
| Capabilities | ToolsJSONVision | ReasoningToolsJSON |
OpenAI: gpt-oss-20b wins on more metrics (6 of 6), but the right pick depends on what you optimize for — see the breakdown below.
The table on this page compares Mistral: Mistral Medium 3.1 and OpenAI: gpt-oss-20b metric by metric — intelligence benchmarks, coding score, speed, latency, context window and per-token price — with the winner marked on each row. Data refreshes hourly.